Every test should start from a one-page brief. Hypothesis, variants, sample size, success metric, guardrails. This template removes every ambiguity that causes tests to stall mid-flight.
The hypothesis is what separates a disciplined experiment from a random change. Format: 'We believe [change] will cause [metric] to move by [amount] because [reason].' If you can't write it in that format, you don't understand the test yet.
Specify the control and every variant explicitly. Include screenshots or Figma links. Write down what is changing and what is staying the same. The most common test failure mode is 'we thought we were testing X but actually tested X+Y'.
Every experiment needs three categories of metrics. Primary (what you're optimizing), secondary (what you're watching), guardrail (what would cause you to halt the test even if primary is winning). Skipping guardrails is how teams ship wins that hurt the business.
Pre-commit to the sample size before the test runs. Otherwise 'we'll stop when we hit significance' becomes 'we stopped when we saw the number we wanted'. The statistical framework you use (Frequentist vs Bayesian) doesn't matter much — consistency matters.
Every test should have a single owner accountable for shipping, analyzing, and communicating the result. Committees don't own tests — people do. Name the person. Include the shipping engineer, the analyst, and the comms owner explicitly.
The highest-signal test briefs we've seen fit on one printed page. The lowest-signal ones sprawl across five Notion pages and include a 'background' section nobody reads.
Google Doc, Notion, and PDF versions. Shared with you in one email.
Download →Enter what you pay Optimizely, Crayon, Hotjar, and Ahrefs today. See what Optimize Pilot would cost instead — and how many headcount the delta covers.
Enter your baseline conversion rate, minimum detectable effect, and weekly traffic. Get the required sample size per variant and an estimated test duration.
How high-performing CRO teams ship more experiments without sacrificing statistical rigor. Includes the idea-to-ship workflow we see work in practice.
Navigator AI produces pre-filled experiment briefs from your top-ranked hypotheses. Review, edit, ship — no blank page.