Case studies

Four products, four different gaps

Same instrument, opposite prescriptions. Each engagement opens with a score pair and closes with one, because that is the only claim worth making, and the “what did not work” section is never omitted.

Clients are identified by industry only; names are withheld and surfaces simplified.

B2B SaaS · reconciliation

Silent friction

Invoicing platform

Reconciliation was costing nine minutes a batch and nobody had ever complained about it.

Before

2.9@0.81

After · 9 weeks

3.6@0.87

Delta

+0.7

9m → 70s
Median time to reconcile a batch
+18pts
Batch completion
2.4 → 3.9
Task surface, behaviour

Read the breakdown

Ecommerce · checkout trust

Phantom friction

Homeware brand

A checkout that worked, and a checkout nobody trusted. The planned redesign was cancelled; copy shipped instead.

Before

3.1@0.74

After · 7 weeks

3.4@0.79

Delta

+0.3

2.3 → 3.5
Perceived trust reading
6% → 5%
Payment drop-off, as predicted barely moved
0
Screens rebuilt

Read the breakdown

Media · activation

Loud friction

Subscription news

Everyone knew onboarding was broken. The research budget went to the thing nobody knew: where readers actually stalled.

Before

2.6@0.68

After · 10 weeks

3.2@0.83

Delta

+0.6

40% → 61%
New cohorts reaching the activation event
+4pts
Week-three retention
0.68 → 0.83
Confidence, from instrumenting depth of use

Read the breakdown

HR platform · time off and onboarding

Silent friction Loud friction

HR platform

Two populations on the same screens, in opposite quadrants. The four per cent who had been asked were describing a different product.

Before

2.7@0.72

After · 11 weeks

3.5@0.86

Delta

+0.8

4m10s → 50s
Median time to file a time-off request
31% → 9%
Requests entered by an operator on the employee’s behalf
1
Form field deleted, after four weeks spent redesigning the wrong screen

Read the breakdown

Next up

This site is scoring itself in public. The pre-launch baseline is described in Scoring our own product in public; the rescore against real behavioural data will be published here as a further case.

Which one is yours?

Twelve signals, two passes, about two minutes.