Hero
02 — UX Research · Mixed Methods · Goldsmiths MSc · 2025

Visual Cues Study

How visual cues in fitness instructional videos affect beginner motor learning — a three-phase mixed-methods study producing an original design framework.

Role Lead Researcher
& Designer
Institution Goldsmiths,
University of London
Duration 6 Months
MSc Dissertation
Methods Survey · Usability Testing
Expert Evaluation
58 Participants · 3 Phases
3 Research Phases
4 Visual Cue Types
1 Design Framework
Overview

This study investigates how visual cues in fitness instructional videos affect beginners' motor learning. Through a three-phase approach, I evaluated four types of visual cues — directional, measurement, body highlight, and metaphor — across quantitative performance data and qualitative user insights. The study produced an original design framework: the Dynamic Balance Model of Cognitive Load and Learning Effectiveness.

Research Gap

86% of people use video to learn physical skills. Almost no one has studied whether those videos actually work.

The Problem

Digital learning has fundamentally changed how beginners approach physical training. Over 70% of users turn to YouTube to learn exercises like squats and stretching. Yet without professional guidance, learners routinely imitate incorrectly — leading to poor form and risk of injury. The deeper problem is systemic: the design of visual cues in existing tutorial videos is driven by intuition and aesthetic preference, not evidence. There are no established guidelines for what works.

Research Context

Prior research has focused on VR/AR and interactive systems, leaving video-based motor learning systematically underexplored. This study addresses that gap by examining four cue types through the lens of Cognitive Load Theory and the theory of external attention.

Research Question

How do four types of visual cues — directional, measurement, body highlighting, and metaphorical — support and improve beginner motor learning in instructional videos?

Methodology

A three-phase mixed-methods approach — quantitative and qualitative, beginners and experts.

01 Online Survey
38 beginners · Quantitative

Distributed to fitness beginners aged 18–40 via social media and university networks. The survey mapped learning habits, video-watching behaviour, and initial preferences for six types of visual cues — including skeletal overlays and muscle activation highlights. Participant feedback identified four cue types as the focus for subsequent testing.

02 Usability Testing
10 beginners · Observation + Interview

Each participant watched four short squat tutorial videos, each embedding a different visual cue type. Performance was observed and recorded. Post-task interviews collected qualitative insights on clarity, helpfulness, and perceived cognitive load. A counterbalanced design controlled for order effects.

03 Expert Evaluation
10 coaches & practitioners · Triangulation

Coaches, personal trainers, and experienced exercisers evaluated an optimised video integrating all four cue types. This phase gathered professional validation and critical perspective, allowing cross-comparison between novice and expert responses — which revealed significant differences in how cues are perceived across experience levels.

Core Premise

Most instructional video designs are built on intuition and aesthetic preference, not evidence. This study produced a systematic ranking of the four cue types for beginner motor learning in instructional fitness videos — grounding design decisions in observed data rather than assumption.

The Experiment

Four cue types. One squat. Measured against the same beginners, under the same conditions.

Directional cue — green arrows showing downward foot pressure
01 Directional Cues

Arrows and motion lines that indicate the direction and path of movement. Immediately communicates where the body should go, reducing ambiguity about force and trajectory.

Optimal Zone — Best for beginners
Body highlight cue — anatomical muscle labels overlaid on the body
02 Body Highlight Cues

Glow or colour emphasis on specific muscle groups or joints during movement. Draws attention to where activation should occur, supporting body awareness and proprioception.

Useful — Frequency must be controlled
Metaphor cue — a virtual chair illustrating the sit-back movement
03 Metaphor Cues

Life-like comparisons that translate abstract movement concepts into familiar imagery — e.g. "sit back into an invisible chair." Relies on prior experience to be decoded correctly.

Expert-first — Novices find harder to interpret
Measurement cue — 45-degree angle line overlaid on torso lean
04 Measurement Cues

Angle lines, degree markers, and depth indicators overlaid on the body. Provides precise numerical reference but requires abstract interpretation that overwhelms beginners.

Confusing Zone — Precise but overloading
Qualitative Findings

Gioia coding made the hierarchy visible: each cue has a different role — and a different cost for beginners.

What the coding revealed

Directional

Provides real-time movement guidance and corrective feedback. Participants found it immediately clear and lowest in cognitive cost — but weak contrast, missing motion, or poor timing could make the arrow feel like UI chrome rather than instruction.

Body Highlight

Directs attention to muscle activation and postural control. Useful when timed well; pacing and overuse emerged as key qualitative risks.

Metaphor

Conveys movement via analogy. Intuitive for participants with prior exercise experience — ambiguous and harder to embody for true beginners.

Measurement

Offers precision and standards, but abstract numbers are hard to translate into physical movement. High cognitive load with limited beginner payoff.

Cross-cutting
From coding to model

Comparing these roles across the four coding trees — not researcher intuition alone — underpinned the Dynamic Balance Model (learning effectiveness × cognitive load). The quantitative scores in the next section align with these qualitative roles.

Method

Inductive Gioia coding (1st-order concepts → 2nd-order themes → aggregate dimensions) was applied to post-task interview transcripts from beginner usability testing. The analysis ran per cue, producing four comparable coding trees. Staying in participant language before theorising ensured the framework reflected what users said, not what was assumed.

How conclusions were derived

01 1st Order Concepts

Raw participant quotes tagged with emergent descriptive codes — language stays close to the informant's own words

"more clear and intuitive" "easier to understand and implement" "helps me achieve the right posture" "specific numbers are of little help"
02 2nd Order Themes

1st order codes grouped by conceptual similarity into researcher-identified theoretical categories

Clarity & Understanding Practical Application Feature Preference Personal Exercise Experience
03 Aggregate Dimensions

2nd order themes consolidated into high-level theoretical constructs that explain the underlying phenomenon

Cognitive Preference Embodied Experience Instructional Value
Scroll to explore coding evidence
Gioia Coding Analysis — Directional Cues
01 — Directional Cues Clear real-time guidance — but can be misread as UI chrome
Gioia Coding Analysis — Body Highlight Cues
02 — Body Highlight Cues Focuses attention on form — pacing can overload
Gioia Coding Analysis — Metaphor Cues
03 — Metaphor Cues Relatable for some — ambiguous for true beginners
Gioia Coding Analysis — Measurement Cues
04 — Measurement Cues Precise on paper — hard to embody; high load
Findings

A clear hierarchy emerged. Not all visual cues help beginners equally.

Data Source

These scores align with the qualitative roles above. Scores are Likert mean ratings from post-condition self-report questionnaires (n=10 beginners, Iteration 01 usability testing). Each participant watched all four cue videos in counterbalanced order and rated each immediately after. Effectiveness summarises the Preference dimension from dissertation Table 2; Cognitive Load summarises the self-reported Mental Effort dimension. Full six-dimension data (Usefulness, Clarity, Distraction, Mental Effort, Engagement, Preference) with standard deviations are reported in dissertation Table 2.

Visual Cue Effectiveness (M) Cognitive Load (M) Zone
Directional Cues 6.5 1.5 Optimal
Body Highlight Cues 4.7 4.2 Useful
Metaphor Cues 4.4 3.6 Expert-first
Measurement Cues 3.1 4.8 Confusing
Beginners vs Experts

Experience changes everything. The same cue lands completely differently depending on who is watching.

Key Divergence

Expert evaluators rated metaphor cues significantly higher than beginners — because they already possess the conceptual framework to decode abstract comparisons. For beginners, the same cue created confusion. Measurement cues showed the opposite pattern: experts valued precision, but beginners were overwhelmed by numbers they couldn't contextualise into physical movement. This divergence confirms that visual cue design cannot be one-size-fits-all.

Visual Cue Beginners — Effectiveness (M, usability testing, n=10) Experts — Perceived Reception (qualitative, n=10)
Directional
M = 6.5
Strongly positive Agreement across groups — unanimously the clearest cue
Body Highlight
M = 4.7
Moderately positive Experts more precise in feedback — noted timing and overuse risks
Metaphor
M = 4.4
Expert preferred ↑↑ Experts found it intuitive and memorable; beginners found it ambiguous without context
Measurement
M = 3.1
Tolerated by experts Experts valued precision; beginners were overwhelmed by numbers without physical reference
Original Framework

The Dynamic Balance Model — plotting visual cues against the tension between learning effectiveness and cognitive cost.

The Model

This study's primary contribution is a two-axis framework for evaluating instructional visual cues. By mapping Learning Effectiveness against Cognitive Load, the model creates four zones — Optimal, Overload, Ineffective, and Confusing — that provide actionable guidance for video designers. The model also identifies the Double-Edged Sword Effect: integrating multiple cues simultaneously increases perceived professionalism but raises cognitive load, potentially undermining learning efficiency.

Optimal Zone
High effectiveness,
low cognitive load
Directional
Overload Zone
High effectiveness but
costly for learners
Ineffective Zone
Insufficient to
support learning
Confusing Zone
Low effectiveness,
high cognitive cost
Measurement
Body Highlight
Metaphor
Low High
Cognitive Load →
Key Finding

The Double-Edged Sword Effect: integrating multiple visual cues simultaneously enhances professionalism and trust — but raises cognitive load and may fragment focus, ultimately reducing overall learning efficiency.

Design Principles

Eight evidence-based guidelines for instructional video creators.

Cue Selection
  • Prioritise directional cues for clear and immediate movement guidance.
  • Use body highlighting to focus on key body parts — but avoid overuse.
  • Be cautious with measurement cues; numbers may confuse beginners.
  • Apply metaphorical cues contextually — they work better for experienced learners.
Delivery & Structure
  • Ensure timely appearance of cues at critical movement moments.
  • Keep videos short (~90 seconds) to match beginner attention and preference.
  • Do not display more than two cue types simultaneously — cognitive overload risk.
  • Incorporate multiple camera angles to reduce blind spots and support diverse learner needs.
Reflection

What it means to design with evidence, not assumption.

As a Researcher

This project deepened my conviction that design decisions must be grounded in evidence. The multi-phase methodology — survey to usability testing to expert evaluation — showed me how much a single method misses, and how triangulation builds credibility. The Gioia coding process transformed raw participant quotes into a structured framework, revealing patterns invisible to intuition.

As a Designer

The most surprising finding was the experience gap: cues that experts found intuitive were actively confusing to beginners. This reinforced a principle I now carry into every design project — that your own familiarity with a system is the least reliable measure of its usability. Designing for someone who doesn't yet know what you know requires deliberate research, not empathy alone.

Limitations

The study has inherent constraints worth acknowledging. The sample sizes were small — 10 beginners for usability testing, 10 experts for evaluation — limiting statistical generalisability. The findings are based on a single movement (the squat), and may not transfer directly to other exercises or motor skill domains. Cognitive load was measured through self-report scales rather than physiological methods, introducing subjective bias. Additionally, the controlled research setting differs from real-world video consumption conditions. These constraints are common to early-stage exploratory research and point to productive directions for future work.

Next Project
Health GeneTech