AnswerRange
Prove

When did it actually change? A test decided before the work, not a chart drawn after

The fourth step of the loop is the one the category skips. Re-measurement uses the same questions and repeats as the baseline and a change rule fixed in advance — so a before-and-after is evidence, not a story.

Read the change rule See it on our own store

The analysis of our own intervention study was pre-registered on 11 September 2026 and cannot be edited after the data came in.

Why it matters

Why most before-and-afters are not evidence

Variance is larger than the effect

A platform that contradicts itself on 40% of repeats moves your one-sample score several points a day on its own. A two-point "improvement" measured once is indistinguishable from that.

The questions changed between runs

If the question set is refreshed, expanded or re-picked by the tool, the second number is a different measurement. Provenance moved the same store by nineteen points in one week.

The rule was chosen after the result

A threshold set once the numbers are in can always be set where the result clears it. A rule registered first cannot.

How it works

How a change is judged

  1. 01

    Freeze the instrument

    The baseline's question set, platforms, markets and repeat count are stored with the project. Re-measurement uses exactly those; new questions start a new baseline, never a comparison.

  2. 02

    State the test in advance

    Two-sided Fisher exact test on cited-versus-not counts per platform, at the sample size the project has. The smallest detectable change is printed with the baseline so nobody expects a two-point move to register.

  3. 03

    Re-measure on schedule

    Thirty days for retainers, six weeks for the Sprint; monthly or fortnightly on the Standard plan. Same time of day where the platform allows it.

  4. 04

    Report one of three words

    Per platform: changed, within noise, or new baseline. The report draws no arrow through a change the test did not support, and the CSV of both runs is attached.

What it does

What "proof" looks like in the report

Paired intervals

Baseline and re-measurement drawn on the same track. If the bands overlap, the page says so before it says anything else.

Minimum detectable change

Printed with every baseline: at 144 answers, about six points. The number you are told to expect before the work, not after.

Per-question movement

Which lost questions were won, which were not, and the sources cited in each run — so the next five briefs are chosen from evidence.

Corpus provenance on the label

Every run is labelled category, own-search-queries or natively authored. Two runs with different labels are never compared as before-and-after.

Open analysis

The intervention study's analysis script and its checksum were published before the data existed. The same code judges client re-measurements.

Client-ready

The print layout carries the interval, the test result and the plain-language verdict; an agency's name and colour where the brand goes.

Questions

Asked before buying

What if nothing changed?
Then the report says "within noise", the retainer's next five briefs are chosen from the per-question detail, and nobody is told a story. That outcome is common in month one and is the reason the Sprint runs six weeks.
Why not track daily?
Because daily single samples cannot detect anything smaller than the platform's own variance. Monthly runs with repeats can. The arithmetic is on the comparison section.
Is the test published?
Yes — the method paper is open access and the intervention study's pre-registration, with the analysis script checksum, is on the methodology page.

Read the change rule