Agent Reliability LabRequest a scope
Web

Usability audit: what it finds and what it's worth

A usability audit is a job-completion check, not a taste note. Formats are slice, funnel, or fixes — not a public price list.

A usability audit asks whether a visitor can find, understand, and finish the job they came for: send a request, book a service, buy. It is not “is it pretty,” and it is not a meta-tag SEO pass.

If ads land on a page where the button has slid off a phone screen, the form hangs on a phone field, and delivery terms are in the footer, the budget is already burning. The audit names those barriers as tasks: what is broken, on which screen, what to change.

Buying the PDF does not raise conversion. A list of faults is a list. Sales move when the patches are in the code and the numbers are read again.

ISO without the seminar

ISO 9241-11 defines usability as effectiveness, efficiency, and satisfaction for specified people, goals, and context.

  • Effectiveness — did they finish (form sent, paid).
  • Efficiency — how many extra steps and minutes.
  • Satisfaction — did the process make them close the tab.

ISO 9241-110:2020 (self-descriptiveness, conformity with expectation, error tolerance) is common sense with a standard number, not a hundred-point checklist.

Do not buy the neighbouring product by accident:

  • SEO audit — indexation, duplicates, titles. Robots need clean code; people need a path.
  • Technical audit — server time, script errors, load.
  • Analytics setup — events and goals. A counter records; it does not interpret.
  • CRO / A/B — an ongoing process on a real traffic stream.

A usability audit is a snapshot: where a person stumbles now.

Nielsen’s ten are not a hundred-item sheet

The working base for an expert pass remains Jakob Nielsen’s ten heuristics (1990–1994; NN/g restated them in 2020 for modern UI). They are logic, not a scorecard:

  1. Visibility of status — a loader, “order received.”
  2. Match to the real world — the customer’s words, not warehouse jargon.
  3. User control — undo, dismiss the modal.
  4. Consistency — the same control behaves the same.
  5. Error prevention — a phone mask beats a scolding after submit.
  6. Recognition over recall — keep choices in view.
  7. Flexibility — reorder, saved details, for people who already know the path.
  8. Minimal distraction — kill the banners that compete with the job.
  9. Plain errors — what broke and how to fix it.
  10. Help in place — a hint at the field, not a PDF in the footer.

The expert walks the buyer’s path against those rules.

The myth of five users and 85 percent

“Five users find 85 percent of problems” is a compression of Nielsen, 2000, itself from a 1993 model that put the chance of one person hitting a given fault at about 31 percent.

The rest of the sentence matters:

  • Five people find ~85 percent in one iteration.
  • To find almost all faults in a single pass you need about fifteen.
  • Nielsen’s actual advice: three rounds of five (find, patch, retest), not one heavy fifteen.
  • Separate audiences (wholesale and retail) are separate tests.

Laura Faulkner’s 2003 work (see the HFI summary): five users averaged 85 percent, but some groups hit 55 percent. Ten users averaged 95 percent, with a floor of 82 percent. Perfetti and Landesman (2001) saw five users on a music site find 35 percent.

Five is not a guarantee. For a small commercial audit, a people-lab is uncommon: it is slow and it is a different product. Most barriers are found by an expert walk plus behaviour data.

Search: convenience is not a ranking lever you can buy

No search engine will promote a page because a button was repainted. The link from comfort to rank is indirect.

Google’s page experience docs are explicit: there is no single page-experience signal. Relevance stays primary. What Google does use are Core Web Vitals: LCP, CLS, and INP (INP replaced FID in the set on 12 March 2024; Chrome dropped FID from PSI and CrUX later that year). Good CWV do not buy first place. They remove technical stall. An awkward page can still win if it answers the query better.

A “fix the button, climb the SERP” formula does not exist. Promising positions after an audit is a different service, and a dishonest one.

What session replay sees, and what it will not say

GA4 and Plausible record events you asked them to record. Session replay (whatever vendor you have) shows scroll, field fill, the mobile keyboard, idle stretches.

Limits to keep honest:

  1. Sampling. Many tools do not write every visit. Short bounces often never become a video.
  2. Blockers. Ad blockers and corporate networks drop trackers. Your “heatmap” is a biased subset.
  3. Bots. Some sessions are not people.
  4. Motive. A cluster of clicks on a non-clickable heading is a fact. It does not explain why they left.

Replay shows what happened on the glass. It does not show why. Watching hundreds of videos without a walkthrough protocol is a way to spend the week.

If there is no analytics, the brief says so. Inventing a replay you do not have is not an audit.

The market, without a tariff card

The market splits by what you actually receive:

  • Freelance lots — an automated checklist or a thin PDF. No funnel, no implementation.
  • A working studio pass — key scenarios, analytics if you grant it, a task list a developer can execute.
  • A thick agency binder — funnel narrative; implementation almost always a second invoice.
  • A UX lab — live participants. That is not a PDF “by eye,” and it is not the same product.

Landing-page claims of “+50 percent conversion after the audit” are ads. Conversion moves when the faults are patched.

A useless audit is forty slides: “dated design,” “not enough air,” “dull buttons.” A developer cannot schedule that.

A useful audit is a cognitive walkthrough. Each fault is a ticket:

  1. Finding — page, block, device.
  2. Evidence — annotated screenshot or a session id.
  3. Patch — phone mask, drop the patronymic field, put the request button in the first screen.
  4. Priorityblocks the job / adds friction / cosmetic.

That document can go to a typesetter the same day.

What an owner actually chooses

New site, little traffic. Skip empty charts. An expert pass on Nielsen: terms readable, cart or request usable, forms on a phone.

Traffic exists, requests do not. Open the funnel in GA4 or Plausible. See the step that dies. Check whether the leak is the interface or the delivery chain behind the form.

Do not buy a cheap PDF as a substitute for either. Contrast plugins find contrast. Why a person did not send a request is a different object.

Budget the patches. An audit with no implementation slot is a file.

This studio’s formats match the SEO-and-growth service: slice (one landing or a small site, expert pass, no people-lab), funnel (landing to request, behaviour data required), fixes (funnel plus the first-wave repairs — forms, offer, buttons, mobile hits, obvious stalls; not a redesign). No public price list. No conversion guarantee attached to a PDF.

The point of the audit is not the thickness of the report. It is fewer obstacles between the offer and the decision.

What to check in your system
  • Does the report name page, block, and device, with a screenshot and a concrete patch?
  • Are findings ranked as blocks the job / adds friction / cosmetic?
  • Is behaviour data from GA4 or Plausible — and does the brief say so if there is none?
  • Is implementation in the engagement, or is the PDF expected to raise conversion by itself?
Agent Reliability Lab
● Online