UX research and usability testing: how to assess whether a digital product works as intended by the team

Monika

The product team is usually convinced that a new app is intuitive – after all, they designed it themselves and know every screen by heart. The problem is that a new user does not have this knowledge and often gets lost precisely where the creators see something as obvious. UX research and usability testing are methods that reveal real barriers to use before they affect conversion, retention, or customer support costs.

When is it worth using UX research and usability testing?

The point at which UX research and usability testing often deliver the greatest return comes when there is a gap between the team’s assumptions and users’ actual behavior. The signals are usually clear: users abandon their cart at a specific step, fail to complete registration, cannot find a feature the team considers central, or the support team receives repeated questions about the same issue. Analytics then shows where something is going wrong, but does not explain why. Usability testing helps answer precisely that second question.

It is worth distinguishing between two typical contexts. The first is the design phase – a prototype or wireframe that has not yet gone into production. Testing at this stage is inexpensive because revisions concern the design rather than deployed code. The second context is an already functioning product in which metrics signal a problem, but the team cannot pinpoint it. In both cases, usability testing provides observations of behavior rather than stated opinions – which is its advantage over a satisfaction survey.

Digital product research makes particular sense when the stakes of the decision are high: redesigning the purchasing process, introducing new onboarding, or changing the information architecture. The larger the implementation investment, the greater the risk of replicating an incorrect assumption at scale. Usability testing reduces this risk by confronting the design with users before the decision becomes costly to reverse.

How is usability testing of a digital product conducted?

The method is based on observing a user performing real tasks in a product. The participant does not evaluate the interface in the abstract – they try to achieve a specific goal, such as finding a product, completing a payment, or changing account settings. The moderator observes where the user hesitates, clicks in the wrong place, and at what point they give up. This makes it possible to reconstruct the actual user journey in the app and compare it with the one designed by the team.

A well-prepared test consists of several elements that should be planned before fieldwork begins:

  • Task scenario – a list of real goals phrased in the user’s language, rather than using the names of interface features. Instead of saying “use the category filter,” the task would be “find running shoes in your size.”
  • Success criteria – defining what constitutes task completion and which behaviors are treated as difficulties or errors.
  • Participant recruitment – selecting people who match the real target group rather than random internal testers.
  • Think-aloud technique – asking the participant to comment on their decisions, which reveals how they interpret the interface.
  • Recording and analysis – recording sessions and organizing observations by frequency and severity of the issue.

Tests are divided into moderated and unmoderated. In moderated testing, the researcher leads the session live and can ask about the reason behind a particular decision, which provides deeper insight. In unmoderated testing, the participant completes tasks independently and the researcher analyzes the recordings – this approach is faster and less expensive when there are more sessions. The choice depends on whether understanding the reasons or the scale of observation is the priority.

As Hume’s Institute experts point out, a team that has built a product eventually loses the ability to see it through the eyes of a new user – it knows the shortcuts, assumptions, and logic that the user does not. Usability testing restores this outside perspective because it shows the interface through the eyes of someone seeing it for the first time. This shift in perspective is often the most valuable outcome of the research, rather than the list of issues itself.

The results of usability testing are organized according to two dimensions: how frequently a problem occurs and its impact on task completion. A problem that affects many users and blocks a key goal has a different priority than a minor inconvenience noticed by one person. This hierarchy enables the product team to make decisions about improvements based on evidence rather than intuition or the loudest voice on the team.

What distinguishes a good usability test from superficial research?

The biggest pitfall is testing with people from within the organization. Colleagues know the product, context, and terminology, so they complete tasks without the difficulties a real user would experience. This creates a false sense that the interface works well. The essence of UX research and usability testing is confronting the product with people outside the project who match the target group.

Another common mistake is asking for opinions instead of observing behavior. Asking “do you like this screen?” leads to stated opinions that are poor predictors of actual behavior. Instead, the user is asked to complete a task and their actions are observed. This distinction separates usability-oriented digital product research from traditional declarative research.

Other limitations and mistakes that should be addressed before starting include:

  • Leading participants – a moderator who gives hints or steers the participant invalidates the observation; tasks should be neutral.
  • Tasks detached from reality – an overly artificial scenario does not replicate the user’s real motivation.
  • Confusing usability with aesthetics – a test examines whether a product can be used, not whether it looks good; these are two different questions.
  • Treating the result as a quantitative verdict – a qualitative test indicates the existence and nature of a problem, not its precise scale in the population.

It is worth understanding how usability testing differs from digital analytics. Analytics shows the scale of a phenomenon – how many users abandon a given step – but usually does not explain the cause on its own. In Hume’s Institute projects, combining both approaches produces very good results: analytics identifies the problematic point in the journey, while usability testing explains the mechanism behind it. These methods complement rather than replace one another.

What should be included in a usability test plan?

Before commissioning or launching research, it is worth checking whether the plan covers the elements that determine the quality of the results. The following list serves as a checkpoint for a team preparing usability testing:

  1. Research objective – a clearly defined question the test is intended to answer, rather than a general “let’s check the UX.”
  2. Target group definition – who represents real users and how they will be recruited.
  3. Task scenario – real goals described in the audience’s language and ordered by importance.
  4. Session format – a decision between moderated and unmoderated testing depending on the need to understand the reasons.
  5. Recording method – recording the screen and user behavior to enable subsequent analysis.
  6. Prioritization method – established criteria for the severity and frequency of problems.
  7. Results format – a report format that translates observations into specific areas for improvement.

This framework ensures that UX research and usability testing provide actionable observations rather than a loose collection of impressions. A precisely defined objective and a realistic task scenario are the foundation without which even well-conducted sessions lose their analytical value.

Frequently asked questions

What does usability testing involve?

Usability testing involves observing users as they complete real tasks in a digital product. The researcher analyzes where the user hesitates, makes errors, or gives up in order to identify actual barriers to use. The key is observing behavior rather than collecting stated opinions about the interface.

How many users are enough for a UX test?

In qualitative research, even a small group of participants can reveal many recurring usability issues because the same barriers appear across successive users. The goal is not statistical representativeness, but identifying mechanisms that make a product difficult to use. When the scale of a phenomenon in the population is needed, qualitative testing is combined with analytics or quantitative research.

How does UX research differ from satisfaction research?

Satisfaction research measures stated feelings and attitudes toward a product, while UX research, especially usability testing, observes actual user behavior while tasks are being completed. Satisfaction answers the question of how a product is evaluated, while usability answers whether it can actually be used. These approaches answer different questions and work best as complements to one another.

Ask about usability testing for your digital product

If metrics indicate that users are getting lost in your product and the team cannot identify the cause, contact Hume’s Institute to ask about usability testing tailored to the stage your digital product has reached.