Group name

Website heatmap

Behavior analytics

Voice of customer

Research methods

User journey map

Research methods

User behavior analytics

Behavior analytics

Usability testing

Research methods

Trust signals

Page levers

Tree testing

Research methods

Time on page

Metrics and funnel

Survey design

Research methods

Social proof

Page levers

Session replay

Behavior analytics

Session recording

Behavior analytics

Segmentation analysis

Metrics and funnel

Scroll map

Behavior analytics

Scroll depth

Behavior analytics

Revenue per visitor

Metrics and funnel

Rage click

Behavior analytics

PIE framework

Research methods

Mobile conversion rate

Metrics and funnel

Micro conversion

Metrics and funnel

Message match

Page levers

Macro conversion

Metrics and funnel

LIFT model

Research methods

ICE score

Research methods

Hotjar

Tools

Hick's law

Page levers

Goal completion

Metrics and funnel

Funnel analysis

Metrics and funnel

Form analytics

Behavior analytics

Form abandonment

Behavior analytics

Five second test

Research methods

Fitts's law

Page levers

Exit rate

Metrics and funnel

Event tracking

Metrics and funnel

Drop-off rate

Metrics and funnel

Dead click

Behavior analytics

CRO audit

Research methods

Conversion funnel

Metrics and funnel

Cohort analysis

Metrics and funnel

Cognitive load

Page levers

Click map

Behavior analytics

Cart abandonment

Metrics and funnel

Bounce rate

Metrics and funnel

Average order value

Metrics and funnel

Attention map

Behavior analytics

Anchoring bias

Page levers

Above the fold

Page levers
This is some text inside of a div block.
Created:
updated:

What is usability testing?

Usability testing observes real people attempting real tasks on a site to find where the interface fails them. Participants are given a goal rather than instructions, and the researcher watches what they do and where they struggle. It reveals why something fails, which behavioral analytics cannot.

How many participants do you need?

Fewer than most people expect for finding problems, and more than most people expect for measuring anything.

For finding problems: five per user group. A widely used planning heuristic rather than a threshold. The reasoning behind it is that the same problems recur across participants quickly, so each additional session returns less than the one before. Running three rounds of five, fixing between rounds, beats one round of fifteen.

For measuring anything: far more. Comparing task completion rates between two designs is a statistical comparison and needs a sample sized accordingly. Usability testing is a qualitative instrument, and treating five participants' completion rates as a measurement is a common error.

Distinct user groups need their own five. If technical evaluators and procurement buyers use the site differently, they are two groups.

Key takeaway: five participants to find problems, a sample sized for the comparison to measure whether one design beats another.

Moderated or unmoderated?

ModeratedUnmoderated
Researcher presentYes, liveNo, recorded
Can probe reasoningYesNo
Cost per sessionHigherLower
SpeedDays to scheduleHours
Best forComplex flows, understanding whyStraightforward tasks, larger samples

Moderated allows follow-up. When a participant hesitates, you can ask what they expected. That single capability is what makes it the better instrument for complex B2B flows.

Unmoderated is faster and cheaper and scales, at the cost of losing the ability to ask why. It suits comparing two navigation structures or checking whether a specific task is completable.

A practical combination is unmoderated first to find where people struggle, then moderated sessions focused on those points.

What does usability testing not tell you?

Whether people want it. Usability testing measures whether an interface can be used, not whether the offer is compelling. A perfectly usable page for a product nobody wants converts at zero.

How common a problem is. Five participants hitting an issue tells you it exists. Quantifying prevalence requires analytics or heatmaps.

What people would actually do. Participants know they are observed and are attempting an assigned task rather than pursuing their own goal. They persist longer than real visitors, so tasks that get completed in testing can be abandoned in reality.

Whether your fix works. That is an A/B test.

Anything reliable from the wrong participants. Testing a B2B SaaS buying flow with people who would never buy the product produces confident findings about the wrong audience. Recruitment quality decides whether the study was worth running at all.

Related terms

Five second test · Tree testing · Voice of customer · Session recording · Cognitive load

Service: Conversion Rate Optimization.

FAQ

Is five participants really enough?

For surfacing the problems that recur across users in one group, yes, and returns diminish quickly after that. It is a planning heuristic rather than a guarantee: a problem that one visitor in fifty hits is unlikely to appear in five sessions. For measuring differences between designs, five is far too few, since that is a statistical comparison.

How is usability testing different from session recordings?

Session recordings capture real visitors pursuing their own goals, with no ability to ask questions. Usability testing uses recruited participants on assigned tasks, with the ability to probe reasoning. Recordings show what happens at scale, usability testing explains why.

Can you run usability testing on a live site?

Yes, and it is the usual approach. Testing a prototype before building is cheaper, since problems surface before engineering time is spent. On Webflow there is a third option: publish the redesign to the staging subdomain, which is served separately from the live domain and kept out of search results, and run sessions against that.

Ask AI about this term

Want more revenue from your existing traffic?

We run CRO for Webflow sites — from audit to A/B testing.

Work with us

Work with us