Key takeaways
- **What are the 4 types of usability test questions?** Ask about the task, the expectation, the obstacle, and the user's confidence after trying the next step.
- **What are the 7 methods of usability testing?** Moderated sessions, unmoderated tests, first-click tests, five-second tests, tree testing, card sorting, and session recordings each answer a different question.
- **Which tool is used for usability testing?** Use the tool that fits the method; the important part is the task design and the fix list that follows.
Website usability testing means watching real people attempt real tasks on your site and recording where they struggle. It is the fastest way to find why users fail, because it shows behavior instead of opinion. You do not need a lab or a large budget: 5 participants, a clear task list, and a screen recorder surface the majority of usability problems in a given flow.
This is a practical walkthrough of how to run a test and turn what you see into fixes. It sits alongside broader UI UX design services work, and pairs directly with a UI UX design services.
Why 5 users is usually enough
The often-cited finding from Nielsen Norman Group is that 5 users uncover roughly 85 percent of the usability problems in a single flow. The reason: the same core problems recur across participants quickly. After the fifth person hits the same wall, more testing on that flow adds little. It is more valuable to fix what you found and test again than to keep adding participants.
Test more only when you have distinct user groups. If beginners and experts behave differently, run 5 of each.
The math is simple. If each user has about a 31 percent chance of hitting a given problem, five users have a high chance that at least one hits it. Push to 15 users and you spend triple the time to find the same defects three times over. Spread that budget across three rounds of five instead, and each round tests a better version of the site.
1. Pick the method that fits your question
| Method | Best for | Cost |
|---|---|---|
| Moderated in person or remote | Deep why, follow-up questions | Higher time |
| Unmoderated remote | Speed, volume, natural setting | Lower |
| Guerrilla (quick intercepts) | Fast, cheap early signal | Lowest |
| First-click testing | Whether navigation points work | Low |
Moderated testing lets you probe hesitation in real time. Unmoderated tools record users completing tasks alone, which scales and removes the moderator's influence. For most small teams, a mix works: a few moderated sessions for depth, then unmoderated for volume.
Match the method to the stage. Early, when you are testing whether a flow makes sense at all, moderated sessions give you the "why" behind every stumble. Later, when you are confirming a fix worked across many users, unmoderated volume gives you cleaner numbers.
2. Write tasks, not questions
The most common mistake is asking opinions ("Do you like this page?") instead of setting tasks. Opinions are unreliable; behavior is not. Write realistic, goal-based tasks that mirror why people actually visit:
- "You need a waterproof jacket for a trip. Find one in your size and add it to the cart."
- "You want to know if this product ships to your address. Find out."
- "You changed your mind. Remove the item and start a return."
Good tasks have a clear end state so you can score success or failure. Avoid leading language that hints at the path. Never say "click the menu"; say what the user wants and watch how they look for it.
Give the task a short scenario so the participant has a reason to act. "You are planning a weekend hike and it might rain" puts them in a mindset closer to a real visitor than a bare instruction does. Keep the scenario neutral. Do not describe the feature you want them to find.
3. Recruit participants who resemble real users
Five participants who match your audience beat 20 who do not. Screen for the traits that matter: are they in your customer segment, do they shop this category, are they new or returning. Avoid testing only colleagues, who know the product and the jargon.
Offer a small incentive to reduce no-shows. For most consumer sites, recruiting through a testing panel or your own customer list both work.
Write a short screener of three or four questions to filter for fit. Ask about recent behavior, not intent. "Have you booked a hotel online in the last three months" is a better filter than "would you ever book a hotel online." Over-recruit by one, because a no-show in a five-person round costs 20 percent of your data.
4. Run the session without helping
The hardest discipline in moderating is silence. When a participant struggles, the instinct is to guide them. Resist it, because their struggle is the data. Use these rules:
- Ask them to think aloud: narrate what they expect and why.
- When they get stuck, ask "what would you do next?" rather than pointing.
- Note the moment of hesitation, the wrong turn, and any workaround.
- Record the screen and, if possible, the face for reaction.
- Keep sessions to 30 to 45 minutes to avoid fatigue.
Record everything. You will catch struggles on review that you missed live.
5. Score what you see
For each task and participant, capture:
- Success, partial, or failure against the end state.
- Time on task and number of wrong turns.
- The specific point of hesitation or error.
- A direct quote that captures the confusion.
Patterns emerge fast. When 4 of 5 users miss the same button, that is not a preference issue, it is a design defect. Frequency and severity together tell you what matters. A problem that blocks checkout for every user outranks a minor confusion one person hit.
6. Turn findings into ranked fixes
A pile of observations is not useful until it is prioritized. For each problem, record severity (how badly it hurts the task), frequency (how many users hit it), and rough effort to fix. Rank by severity times frequency, then sort the top items by effort so quick wins go first.
Write each finding as a specific, actionable item: the problem, the evidence (which users, what they did), and a recommended direction. "Users could not find the size selector, 4 of 5 scrolled past it; move it above the fold" is actionable. "Improve the product page" is not.
7. Test again after fixing
Usability testing is iterative. Fix the top problems, then run 5 new participants on the same tasks to confirm the fix worked and to surface the next layer of problems that the first round hid. Two or three rounds on a critical flow like checkout typically produce the largest conversion gains.
Common mistakes to avoid
- Asking for opinions instead of setting tasks.
- Helping the participant, which erases the evidence.
- Testing only internal staff who know the product.
- Leading tasks that reveal the path.
- Running one round and never retesting after fixes.
- Collecting observations but never ranking or acting on them.
A worked test of a checkout flow
Seeing one round play out shows why five participants and a task list beat a stack of opinions. Take a store that suspected its checkout was losing people but could not say where.
- Task. "You want to buy this jacket in your size and have it shipped to your home. Complete the purchase." No mention of the path, no hints.
- What happened. Four of five participants filled the shipping address, saw a shipping charge appear for the first time, and paused. Two said some version of "I did not know it would cost that much" out loud. A third stalled on the phone field, retyping the number three times because it rejected the format silently.
- The pattern. The same two walls, surprise shipping cost and the strict phone field, recurred across participants. That frequency is what turns an anecdote into a design defect worth fixing.
- The fixes. Show shipping cost before the address step, and accept any phone format. Both were quick, and a second round of five new participants completed the task without the same stalls.
The whole first round took a day of sessions and cost five incentives. It found the two problems that analytics had only hinted at, because watching a person hesitate tells you why in a way a drop-off percentage never can.
Running useful tests on a small budget
Usability testing has a reputation for needing a lab and a research team. It does not. The method scales down to almost nothing without losing most of its value.
- Recruit from your own customer list or a low-cost testing panel. Five people who match your audience are enough for one flow.
- Use a screen recorder and a video call for moderated sessions, or an unmoderated testing tool for volume. Neither is expensive.
- Keep tasks to the one flow that matters most, usually checkout or signup, rather than trying to test the whole site at once.
- Offer a small incentive to cut no-shows, and keep each session to 30 to 45 minutes.
The cost of a round is a day of time and a handful of incentives. The cost of not testing is shipping fixes based on opinion and finding out from the revenue whether they worked. Watching five real people is the cheaper way to learn, and it is available to any team willing to write good tasks and stay quiet during the session.
Related terms
You may see this topic described with related searches like how to do website usability testing, remote usability testing, site usability testing, usability testing questions, and usability testing template. Those phrases are useful when they clarify what the reader needs next, but they should still point back to one clear plan.
Related searches such as user interview questions, web usability testing, website usability testing examples, and website usability testing services are useful when they clarify what the reader needs next. They should support the same plan rather than pulling the page in several directions at once.
Frequently asked questions
How many users do I need for usability testing?
Five per user group uncovers about 85 percent of usability problems in a flow, because the same issues recur quickly. Fix what you find and retest rather than adding more participants to a single round. Use 5 of each group only when segments behave differently.
What is the difference between moderated and unmoderated testing?
Moderated testing has a facilitator present to ask follow-up questions and probe hesitation in real time. Unmoderated testing records users completing tasks alone, which scales and removes moderator influence. Many teams run a few moderated sessions for depth, then unmoderated for volume.
What tasks should I ask users to complete?
Realistic, goal-based tasks with a clear end state, such as "find a jacket in your size and add it to the cart." Avoid opinion questions and avoid naming the path. You want to watch how people pursue a goal, not confirm they can follow instructions.
How is usability testing different from a UX audit?
Usability testing is one method: watching real users attempt tasks. A UX audit is the broader review that combines testing with heuristics and analytics to find and rank problems. Testing supplies the behavioral evidence the audit needs.
How much does a round of usability testing cost?
A five-person unmoderated round can cost as little as incentive money plus a platform fee, often a few hundred dollars. Moderated sessions cost more in time than money. The expensive path is skipping testing and rebuilding after launch when the flow fails.
Can I run usability testing on a prototype before building?
Yes, and you should. A clickable prototype tests a flow before a line of production code ships, which is when fixes are cheapest. Link your mockups into a prototype, run five users on the core task, and fix what breaks before development starts.
What are the 5 E's of usability?
Website usability testing shows where real users hesitate, misread, or fail to complete a task. The value comes from turning those observations into ranked fixes, then checking whether the next version is easier to use.
What is unmoderated user testing?
Unmoderated user testing asks participants to complete tasks without a facilitator present. It is useful for clear flows, quick comparison tests, and repeatable tasks, while moderated testing is better when you need to ask follow-up questions.
Useful usability test question types
When people ask "what are the 4 types of usability test questions?", they usually need a simple structure. Cover the task, the expectation, the obstacle, and the confidence level after the task is complete. Ask what the user thought would happen, where they hesitated, what felt unclear, and whether they would take the next step.
Common usability testing methods
The answer to "what are the 7 methods of usability testing?" is usually moderated sessions, unmoderated tests, first-click tests, five-second tests, tree testing, card sorting, and session recordings. Pick the method based on the question you need answered, not because the method sounds more advanced.
Choosing a usability testing tool
If the question is "which tool is used for usability testing?", use the tool that fits the test: Lookback or UserTesting for recorded sessions, Maze for prototype and task tests, Hotjar or Microsoft Clarity for recordings, and Optimal Workshop for card sorting or tree testing. The tool matters less than writing clear tasks and acting on the findings.
Getting help
If you want tasks written, participants recruited, and findings turned into a ranked fix list, our UI UX design services team can run testing rounds on your critical flows. Pair it with a UI UX design services to combine behavior with analytics, and see UI UX design services to know which layer each finding belongs to.
