Home
Bookmarks
Leagues
LEARN
Courses
Career Paths
Assessments
Tutorials
Arcade
Glossary
GROW
Certifications
Log in
Glossary
Log inSign up
Courses/UX Design Foundations/Usability Testing Protocols
Level 6 · Workflows and Deliverables

Usability Testing Protocols

14 min read
6 exercises
250 XP
Start lesson

Ready to test what you learned?

Complete the lesson quiz and earn 250 XP.

Start quiz

Topics in this lesson

Why We Test: The Gap Between Saying and DoingModerated vs. Unmoderated Usability TestingFormulating Objective, Non-Leading Usability TasksThe Think-Aloud Protocol: Accessing Cognitive LogicQuantitative Usability Metrics: Task Success, Time, and SUSSynthesizing Insights: Affinity Mapping and Triage

From Course

📘
UX Design FoundationsBeginner · 25 lessons
Usability Testing Protocols

Lesson

Usability Testing Protocols

Why We Test: The Gap Between Saying and DoingModerated vs. Unmoderated Usability TestingFormulating Objective, Non-Leading Usability TasksThe Think-Aloud Protocol: Accessing Cognitive LogicQuantitative Usability Metrics: Task Success, Time, and SUSSynthesizing Insights: Affinity Mapping and Triage
Start lesson
Complete lesson and earn 250 PX
Exercise #1

Why We Test: The Gap Between Saying and Doing

In the words of legendary usability pioneer Jakob Nielsen:

"The first rule of usability is: never listen to what users say. Watch what they actually do."

Human beings are notoriously poor at predicting their own future behavior. In focus groups or surveys, people claim they would love a feature; when put in front of the actual software, they struggle to find it.

Usability testing is not an opinion survey; it is a scientific observation of human interaction with a product.

Exercise #2

Moderated vs. Unmoderated Usability Testing

Choosing between testing protocols depends on project goals and research budget:

Moderated Testing

A live facilitator guides the participant through the test session:

  • Pros: Can probe into unexpected behaviors ('What went through your mind just now?'), rescue participants from technical bugs, and observe subtle nonverbal cues.
  • Cons: Time-intensive and expensive.

Unmoderated Testing

Participants complete tasks autonomously on their own devices using automated testing platforms:

  • Pros: Fast, scalable, tests in natural real-world home/office environments, and delivers hundreds of data points quickly.
  • Cons: Cannot clarify confusion or ask spontaneous follow-up questions.
Exercise #3

Formulating Objective, Non-Leading Usability Tasks

How you phrase a task script directly determines the validity of your testing data:

The Leading Question Defect

  • Leading (Bad): 'Go to the search bar at the top right, search for Nike sneakers, and click the blue checkout button.' This tests whether the user can follow directions, not whether the interface is usable!
  • Non-Leading (Good): 'You want to buy a pair of running shoes for under $100. Show me how you would find and purchase them on this site.'

A sound task specifies a realistic human goal without mentioning interface elements, button labels, or navigation paths.

Exercise #4

The Think-Aloud Protocol: Accessing Cognitive Logic

Introduced to usability by Clayton Lewis in 1982, the Think-Aloud Protocol asks participants to continuously verbalize their thoughts, intentions, and reactions as they perform tasks:

  • 'I'm looking for the pricing tab... I see 'Plans' here, let me check if that's the same thing... I'm clicking here because I expect to see enterprise tiers...'

This provides an uninterrupted window into the user's mental model, exposing exactly where interface cues contradict human expectations.

Exercise #5

Quantitative Usability Metrics: Task Success, Time, and SUS

Usability is not purely qualitative; it is measured with rigorous mathematical metrics:

  1. Task Completion Rate (Binary 0 or 1): The percentage of participants who successfully achieve the task goal without facilitator intervention (benchmark: > 78%).
  2. Time on Task (Seconds): The duration required to complete the task.
  3. Error Frequency: Count of mis-clicks, wrong turns, and error messages triggered.
  4. System Usability Scale (SUS): A validated 10-item questionnaire yielding a score from 0 to 100. A score of 68 is the global average; 80+ indicates world-class usability.
Exercise #6

Synthesizing Insights: Affinity Mapping and Triage

After completing 5 to 10 usability sessions, research teams synthesize raw observations:

  • Affinity Mapping: Grouping individual observations, pain points, and quotes onto a virtual canvas by thematic cluster.
  • Usability Triage Matrix: Prioritizing fixes by plotting Severity (Catastrophe to Minor) against Engineering Effort (Low to High).
  • Quick Wins: High-severity, low-effort usability flaws (such as ambiguous button copy or missing focus outlines) must be remediated immediately in the next sprint.