The question was not whether a visible agent helps. It was when its presence is worth the attention it takes
Pedagogical agents can demonstrate behavior, provide social cues, and make instruction feel more personal. The same visual presence can also become extra stimulus when the learner is already doing abstract mental work.
For my MSc Human-Technology Interaction thesis at TU/e, I tested whether that tradeoff changes with the task. I compared a visible pedagogical agent with the same instruction delivered without the agent across psychomotor and cognitive Tetris-based tasks.
I designed the experiment around an interaction rather than a universal winner. If the agent was useful because learners could model what they saw, its effect should be stronger when the task actually required physical imitation.
A visible agent should earn its place through the task it supports, not simply because an interface can show one.
A mixed-factorial experiment separated agent presence from task demand
Agent visibility was between participants. Each participant completed both task types in randomized order.
The agent helped most when learners had something physical to model
The headline result was an interaction: visibility did not help every task in the same way.
The same Tetris rules were adapted into two different kinds of work
Keeping one learning domain helped isolate a more useful question: what changes when the learner has to imitate movement rather than reason through a problem?
Learn through hand movement
Participants used predefined hand gestures to control Tetris pieces. A visible agent could provide a behavioral model that was directly relevant to the action being learned.
The visual demonstration was part of how the task could be understood and imitated.Learn through mathematical problem-solving
Participants solved mathematical problems to perform game actions. The task depended more on abstract reasoning, making an additional visual agent less directly relevant to the work itself.
This let the study test whether visual presence becomes less useful when modeling is not the main need.The design separated a between-person condition from a within-person task comparison
Participants were randomly assigned to one visibility condition, then completed both cognitive and psychomotor tasks in randomized order.
157 participants completed the experiment.
Each participant is assigned to one visibility condition.
Instruction includes the on-screen pedagogical agent.
The same instructional content is shown without the visible agent.
Cognitive and psychomotor tasks are completed in randomized order, with basic and difficult levels.
Task performance, self-efficacy, recall, and affective beliefs are analyzed across visibility and task type.
Flow connections
- participants to assign
- assign to visible, visible
- assign to absent, absent
- visible to tasks
- absent to tasks
- tasks to outcomes
The diagram is a portfolio view of the experimental structure. It keeps the between-subject visibility condition separate from the within-subject task comparison.
I owned the study from experimental structure through analysis
The project combined HCI research design with experimental materials, automated measurement, and quantitative analysis.
The study kept the learning domain stable while changing task demand and agent visibility
These original project artifacts show how the factorial design and the Tetris-based learning tasks were represented in the study.
One experiment crossed visibility with task type
The design compared a visible versus absent agent while participants completed psychomotor and cognitive learning tasks.

The psychomotor condition needed a scalable score without treating AI as ground truth
Psychomotor performance depended on whether a participant performed specific hand movements correctly. The sessions were video recorded and an AI-based gesture-recognition system was used to score those performance assignments.
Automating the score only helped if it behaved closely enough to manual assessment. I therefore treated validation as part of the measurement design rather than assuming the model output was correct because it was automated.
The AI score was a measurement instrument. It still needed evidence that it tracked the human reference closely enough for the study.
Automated gesture scores were checked against manual scoring before being trusted
A validation sample compared the AI-generated scores with manual ratings of the same performance data.
Visibility improved psychomotor performance, not cognitive performance
Mean task-performance scores by task type and agent condition.
The important result is the interaction. The visible agent materially helped the psychomotor task, while cognitive performance did not significantly improve.
View data
| Task type | Visible agent | No agent |
|---|---|---|
| Cognitive | 6.13 | 6.47 |
| Psychomotor | 5.59 | 4.39 |
Visibility × task type interaction: F(1,150)=10.72, p<.0013. Cognitive visible vs absent was not significant (p=.27); psychomotor visible vs absent was significant (p<.001).
The same interaction appeared in how capable learners felt
Mean self-efficacy ratings by task type and visibility condition.
Visible agents increased psychomotor self-efficacy, but cognitive self-efficacy moved in the opposite direction. Presence was not uniformly beneficial even when the content stayed the same.
View data
| Task type | Visible agent | No agent |
|---|---|---|
| Cognitive | 5.66 | 5.97 |
| Psychomotor | 6.31 | 5.64 |
Visibility × task type interaction: F(1,465)=9.73, p<.0022. Psychomotor visible vs absent p<.001; cognitive visible vs absent p=.043, with the visible condition lower.
The broader pattern reinforced a task-dependent interpretation
Null results and affective outcomes matter because they stop the conclusion from becoming 'show an agent everywhere.'
When the assumptions were messy, I checked whether the conclusion survived a different analysis
Several outcome distributions contained outliers or did not meet normality and variance assumptions cleanly. Relying on one conventional ANOVA result would therefore give more confidence than the data justified.
I used the traditional analyses, then checked key conclusions with robust mixed-effects models. The central interaction for task performance and self-efficacy remained: the effect of visibility depended on the type of task.
The research claim came from a pattern that survived a robustness check, not from one convenient p-value.
Agent visibility should behave like a task-level design decision, not a default feature
The study supports a more conditional rule for learning interfaces with visible AI or instructional characters.
Behavioral modeling and physical imitation
When learners need to observe and reproduce an action, a visible model can provide task-relevant information. In this study, that was where performance, self-efficacy, and affective responses improved most clearly.
Use presence because it contributes to the task.Abstract reasoning without a modeling need
The visible agent did not improve cognitive task performance and cognitive self-efficacy was slightly lower. Extra visual or social presence should therefore justify what it adds before occupying attention.
Do not equate visibility with support.This was a controlled learning study, not proof that one agent pattern works everywhere
The experiment used Tetris-based tasks in a controlled lab setting. Self-efficacy and affective beliefs were self-reported, outliers were retained, and the AI scoring validation applies to this measurement setup rather than every gesture-recognition system.
Cognitive load theory helped explain why an additional visual model might be less useful during abstract reasoning, but the four dependent variables reported in the study were task performance, self-efficacy, recall, and affective beliefs. I therefore treat cognitive load as an interpretive lens here, not as a directly measured outcome.
A stronger next study would test other learning domains, longer-term exposure, and different agent appearances or behaviors, while continuing to validate any automated performance measure against a human reference.
