When a website underperforms, a UX audit can sound like the sensible first purchase. It is fast, tangible and often produces a long list of issues. Usability testing can sound more rigorous because real people are involved. Neither label tells you whether the method will answer the decision in front of the organisation.

An expert review identifies likely friction by examining an experience against established interaction principles and the project context. Usability testing observes representative people attempting realistic tasks. The first gives efficient breadth and hypotheses. The second reveals behaviour, language, expectations and recovery that expert judgement can miss.

The credible choice depends on what is uncertain, which audience and journey are affected, the consequence of being wrong and what evidence already exists. Sometimes one method is enough. Sometimes the strongest sequence is review, test, redesign, retest and monitor.

The short answer: choose evidence for the decision

Use a UX audit when the organisation needs efficient triage across a defined experience, wants to identify likely friction or needs hypotheses to focus later research. Use usability testing when understanding, sequence, expectations or task completion needs observed evidence from relevant users. Combine methods when the decision is consequential, the problem is broad or competing evidence needs to be reconciled.

Start by writing the decision, not the deliverable: which journey may change, which audience is affected, what the team believes, what it needs to learn and what action will follow. Do not commission a report if no one owns prioritisation and implementation.

  • Define the redesign decision and affected audience
  • Review the evidence already available
  • Use expert review for breadth and testable hypotheses
  • Use usability testing for observed task behaviour
  • Keep accessibility, analytics and technical evaluation explicit
  • Choose a proportionate sequence and name the implementation owner

Comparison cards showing UX audit strengths and usability testing strengths, with shared limits around accessibility, analytics and technical evidence.

Neither method answers every experience question alone.

Compare the evidence, not the label

UX audit is not a consistently defined industry product. It may describe a heuristic review, expert review, analytics review or mixed diagnostic. Compare the journeys examined, evidence used, specialist inputs, outputs and limitations rather than assuming the title guarantees a method.

Usability testing and user acceptance testing also answer different questions. Observed usability testing examines whether representative users understand and complete tasks. UAT checks whether delivered functionality meets agreed acceptance criteria. Research planning should also define recruitment, incentives, consent, recording access, retention and secure disposal.

What an expert UX audit can reveal

A UX specialist reviews representative pages, interfaces and journeys against recognised usability principles, design conventions and the known audience and business context. The review may examine hierarchy, labels, navigation, forms, error handling, calls to action, consistency, responsive behaviour and the clarity of the proposition. A useful commercial context is why beautiful websites underperform.

Its main strength is efficient coverage. An experienced reviewer can identify likely friction and repeated patterns across many interfaces without recruiting participants for every section. It can expose obvious defects, inconsistent components and assumptions that deserve further investigation.

The findings remain expert judgement. A review cannot prove how often an issue occurs, whether a particular audience will understand a label or what contextual behaviour will emerge. Strong reports distinguish observation, principle, likely impact, evidence confidence and recommended next validation. Weak reports present every preference as a defect and every defect as equal priority.

What usability testing can reveal

Moderated usability testing involves watching participants try to complete specific tasks using a live service or prototype. GOV.UK guidance describes asking people to think aloud so researchers can understand what they are doing, thinking and feeling. Tasks should be relevant, believable and neutral rather than leading people to the intended answer.

Observation reveals where people form a different expectation from the design, misunderstand language, choose an unexpected route, fail to notice an action, recover from errors or bring context the team did not anticipate. It can uncover the difference between what people say they prefer and what they do during a realistic task.

The quality depends on the research question, participant relevance, task design, facilitation, environment and interpretation. There is no universal participant number that fits every decision. A small qualitative study can discover important issues, but it is not a statistically representative conversion estimate. State what the evidence supports and where uncertainty remains.

Understand the blind spots of both methods

An expert review can miss context-specific behaviour and overvalue conventions that do not fit the audience. Usability testing can miss issues outside the selected tasks or participant groups. A participant completing a journey in a research session does not prove that every customer can complete it at scale or under different conditions. Technical evidence may require a separate technical SEO audit, while accessibility planning should follow WCAG 2.2 and accessibility in Discovery.

Neither method automatically validates analytics, search visibility, technical performance, security, privacy, content accuracy or operational handovers. A broken form notification may look like a conversion problem while the interface works as designed. A slow third-party script may affect behaviour without appearing in a static audit.

UX research also does not equal accessibility conformance. W3C guidance recommends combining appropriate evaluation tools, knowledgeable human review and involvement of people with disabilities. Usability sessions with disabled participants can provide valuable experience evidence, but the sample and tasks do not replace a defined accessibility evaluation.

When a UX audit is the credible first step

An audit can be proportionate when the organisation needs to triage a broad but controlled experience, identify repeated interface issues or establish hypotheses before committing research capacity. It can also suit a reversible decision where existing user and analytics evidence is reasonably strong.

Define representative journeys rather than asking for the whole site in the abstract. Supply available audience research, analytics, brand and content constraints. Ask the reviewer to state the evidence basis, confidence and consequence of each finding. Separate clear defects, plausible friction and subjective preference.

The output should support action. Group findings by journey or problem, identify cross-cutting patterns, distinguish quick repairs from design questions and name what needs user evidence. A site-specific audit and prioritised roadmap are paid professional work; an agency should not promise them as a free pre-sales review.

When usability testing is necessary

Use observed research where user comprehension or behaviour is the central uncertainty. Examples include contested navigation labels, a complex eligibility journey, a high-value enquiry process, unfamiliar product selection, account tasks or a redesigned workflow that changes the sequence people must follow.

Prioritise the journeys where failure has business, service or inclusion consequences. Recruit actual or likely users who reflect the relevant contexts. Include people with appropriate access needs when the design decision involves accessibility and inclusion, and support them to use familiar devices or assistive technology where practical.

Research can begin with a prototype before development hardens. It can also diagnose a live service. The goal is not to ask whether participants like the design. It is to observe whether they understand the proposition, find the right path, complete the task and recover when something goes wrong.

Decision flow for choosing a UX audit, usability testing or a combined research sequence based on uncertainty and risk.

The smallest credible pathway may be one method, a sequence or paid Full Website Discovery.

Use a combined evidence sequence when risk justifies it

A combined programme often starts with the material already available: analytics, search terms, support themes, sales feedback, prior research, accessibility findings and technical performance. The team defines the decision and gaps rather than gathering data without a question.

An expert review can then identify patterns and form testable hypotheses. Usability testing explores the priority assumptions with representative users. Material changes can be prototyped and tested before implementation, then monitored after release. Trace each decision back to evidence and record what was not tested.

This is not a demand for a large research programme on every redesign. Use the smallest sequence that can reduce the consequential uncertainty. Full Website Discovery is appropriate when user research must be integrated with content, information architecture, data, complex integrations, ecommerce, portals, significant migration or governance. A contained question may suit a focused research engagement.

Commission research that can be acted on

A useful brief names the decision, audience, priority journeys, existing evidence, constraints, research questions, intended output and implementation owner. It should explain how participants will be recruited, how tasks will avoid leading language, how privacy and consent will be managed and how limitations will be reported.

Ask how accessibility, analytics and technical evidence fit. They may be parallel workstreams, specialist reviews or explicit exclusions. Avoid assuming that a broad UX label covers them. Define whether the work produces hypotheses, prioritised findings, prototype recommendations or implementation requirements.

Plan a decision meeting, not only a presentation. Agree which findings trigger repair, further validation, prototype work, backlog entry or no action. Preserve the evidence and reasoning so future teams do not repeat the same debate.

Prioritise findings without manufacturing precision

A long audit can make every issue appear equally actionable. Prioritisation should consider affected audience, task consequence, frequency evidence, confidence, breadth, risk, effort and dependency. A severe barrier in a low-volume but critical task may outrank a common cosmetic inconsistency. A likely issue with weak evidence may deserve testing before development.

Separate severity from certainty. Severity asks how damaging the problem would be if the interpretation is correct. Certainty asks how strongly the available evidence supports that interpretation. This distinction prevents an expert’s confident wording from becoming false proof and helps teams choose between repair, research and monitoring.

Avoid converting scores into an automatic roadmap. The solution may affect brand, content, technology, accessibility or operations. Group related findings into underlying problems and confirm dependencies before estimating. A prioritised findings report can recommend a next decision without pretending that every implementation detail is settled.

Use research to challenge the solution, not validate a preferred design

Research briefs sometimes begin with a solution that a senior stakeholder wants approved. Tasks then direct participants towards the intended feature, and positive comments are reported as validation. This creates confirmation rather than evidence.

Write neutral research questions around behaviour and understanding. Instead of asking whether people like a new menu, ask where they would go to complete a realistic task and what they expect under each label. Instead of asking whether a form is easy, observe completion, hesitation, errors and recovery. Ask follow-up questions about what happened, not whether the participant agrees with the team.

Include the delivery team in observation and synthesis without letting observers intervene. Shared exposure to behaviour reduces the chance that a written report becomes abstract or contested. The researcher should still lead method, privacy, facilitation and interpretation so stakeholder enthusiasm does not distort the session.

Evaluate the quality of a proposed UX engagement

Ask a prospective partner to explain the decision it believes the work should support. A credible proposal names the research or review scope, representative journeys, audience assumptions, evidence inputs, method, outputs, limitations, specialist boundaries and client responsibilities. It should explain how findings move into design and implementation.

Challenge generic promises. An audit that covers every page may provide less value than a focused review of the highest-consequence journeys. A test with more participants is not automatically stronger if recruitment is irrelevant or tasks are leading. A visually polished report is not useful if findings cannot be traced to evidence and action.

Confirm ownership of recordings, notes, personal information, prototypes and findings. Define access, retention and disposal consistent with the research context and applicable obligations. Participant incentives, recruitment, accessibility support and specialist services should be visible in scope rather than treated as incidental extras.

Retest the changes that matter

Research should not end when findings are accepted. Where a change is material, evaluate the revised prototype or implementation against the original problem. Check whether the fix created a new barrier, shifted confusion elsewhere or works only for the most familiar audience.

Retesting does not mean repeating the entire programme. Focus on the changed assumptions and high-risk journeys. Combine observation with technical, accessibility and analytics checks where relevant. Document what improved, what remains unresolved and what should be monitored after release.

This closes the evidence loop. The organisation learns not only which interface had a problem, but whether its reasoning and intervention were effective. That learning improves later redesign decisions and reduces dependence on individual opinion.

Preserve the original tasks, prototype version and decision record so the comparison remains meaningful. If the team changes the audience, proposition and interaction at once, describe the result as a new evidence round rather than a clean confirmation of one fix. Note any remaining risk, affected audience, decision owner and monitoring responsibility before release and review it after launch.

Six-part evidence table covering usability testing, expert review, analytics, accessibility evaluation, technical performance and operational feedback.

Evidence should be traced to the decision it supports.

Questions to ask before buying either method

  • What decision will this evidence change?
  • Which audience and tasks matter most?
  • What do analytics, support and existing research already show?
  • Which assumptions need observation rather than expert judgement?
  • How will accessibility and technical evidence be evaluated?
  • Who will prioritise and implement the findings?
  • What limitations and residual uncertainty will be recorded?

Frequently asked questions

Is a UX audit the same as usability testing?

No. A UX audit is an expert review that identifies likely issues and hypotheses. Usability testing observes representative participants attempting realistic tasks. They provide different evidence.

Which method should come first?

It depends on the decision and evidence. An audit can efficiently focus later testing. Direct testing may be better when a specific audience assumption is the main uncertainty. Existing evidence may also make one step unnecessary.

How many people should take part in usability testing?

There is no universal number. It depends on the research question, diversity of relevant users, tasks, risk and whether findings stabilise. The plan should explain what the chosen sample can and cannot support.

Can a UX audit prove why conversion is low?

It can identify plausible friction, but it cannot prove causation alone. Combine it with trustworthy measurement, user observation, operational evidence and proportionate validation.

Does usability testing prove accessibility?

No. Research with people with disabilities is valuable, but accessibility evaluation requires an appropriate scope, standards knowledge, tools and human review. The two evidence streams complement each other.

When should UX research be included in Discovery?

When the user evidence must inform broader content, architecture, workflow, integration or governance decisions. A narrow, well-defined question can be commissioned as a contained activity.

How Emote can help

Emote can frame the decision first, then combine relevant UX/UI, analytics, content, technical and delivery evidence to select a research method that fits the uncertainty.

The smallest credible paid step may be a focused expert review, usability testing or a combined sequence. Broader journey, content, architecture or integration questions may justify paid Full Website Discovery before redesign work is fixed.

If you need stronger evidence before changing an important website journey, book an initial meeting with Emote.

Up next: Website information architecture: turning audience needs into a useful sitemap and navigation

Read More