Why stopping an A/B test at first significance backfires
A subscription news app is testing a redesigned paywall against the old one, with subscription rate as the metric. The plan was to run for 21 days. The analyst opens the dashboard every morning and intends to call a winner the first day the p-value drops below 0.05. On day 4, treatment is ahead with , and the team wants to ship.
What is wrong with declaring victory now, and what should the team have done instead?
Your answer
Solving needs a free account
Answers, streaks and solutions unlock when you are signed in. Reading the question and the hint stays free.
Discussion
Sign in to join the discussion · reading is open to everyone
💡 Discussion rules
- No full solutions here. Hints and approaches only.
- Complexity, edge cases and intuition are the point.
- Interview experiences are welcome. Respect your NDAs.
Loading discussion…
Learn the concepts
The theory behind this question.
Related questions
Why you can't stop an A/B test when it "hits significance"The test isn't significant, can we just run it longer?An automated alert that pings the moment p drops below 0.05Checking a retention test every morning for a monthKilling a subject-line test the instant it shows a winner
All questions →