Prototyping and Usability Evaluation
Running usability tests with five participants: what you learn and what you miss
Five participants is often enough to catch major usability problems, but not every problem. This post walks through when a small sample is sufficient and when it is not.
Overview
A rule that is often applied too broadly
The idea that five participants is enough to find most usability problems comes from a widely cited model showing that the majority of usability issues are discovered within the first five sessions of a study, with diminishing returns after that. The finding is accurate for what it measures. It is often applied to situations it was never meant to cover, which leads teams to treat five as a fixed rule rather than a starting point that depends on what is actually being tested.
Origins
What the five participant model actually shows
The original research behind this guidance modeled a single user group completing tasks against a single design, using a specific definition of what counts as a usability problem. Under those conditions, a small number of sessions surfaces most of the high frequency issues, because the same problems tend to repeat across participants who share a similar mental model and similar goals. This is genuinely useful for catching the most disruptive issues early and cheaply.
The model says less about issues that are rare but severe, issues specific to a subgroup of users, or comparative questions between multiple design directions. Five participants drawn from a single user group will not reliably surface a problem that only affects a smaller segment, such as screen reader users, or older adults with less familiarity with a particular interaction pattern.
Limits
When five participants is not enough
Three situations call for more than a single round of five. Testing across genuinely different user groups, such as new users and experienced users, or riders and drivers in a two sided marketplace, requires enough participants within each group to see repeated patterns, not just five total split across groups. Comparing two or more design directions requires enough participants per direction to distinguish a real difference from individual variation. Evaluating a feature where errors carry serious consequences, such as a financial transaction or a medical decision, calls for a larger sample specifically because rare and severe problems matter as much as common ones.
Five participants will tell you what most people struggle with. It will not tell you what a specific group of people struggles with.
Approach
A more useful question than how many
Rather than starting from a fixed number, a more reliable approach starts from the decision the research needs to support. A study meant to catch obvious usability problems before a design review can often stop at five, provided the participants represent one coherent user group. A study meant to choose between two prototypes, or to evaluate accessibility for a specific population, needs a sample built around that specific question, which may be larger, smaller, or structured differently than five.
Takeaways
What to carry forward
The five participant guidance applies to catching common problems within a single user group, not to every research question.
Testing across multiple user groups or comparing design directions requires participants within each condition, not five total.
High stakes features benefit from larger samples because rare and severe problems matter as much as common ones.
Choose a sample size based on the decision the study needs to support, not a fixed rule of thumb.
References
Further reading
Nielsen, Jakob. Why You Only Need to Test with 5 Users. Nielsen Norman Group, 2000.
Nielsen, Jakob, and Landauer, Thomas K. A mathematical model of the finding of usability problems. Proceedings of ACM INTERCHI 93 Conference, Amsterdam, 1993, pages 206 to 213.
Budiu, Raluca. Why 5 Participants Are Okay in a Qualitative Study, but Not in a Quantitative One. Nielsen Norman Group.
Nielsen Norman Group. How Many Test Users in a Usability Study.