Qm

Why stopping an A/B test at first significance backfires

A subscription news app is testing a redesigned paywall against the old one, with subscription rate as the metric. The plan was to run for 21 days. The analyst opens the dashboard every morning and intends to call a winner the first day the p-value drops below 0.05. On day 4, treatment is ahead with p=0.04p = 0.04, and the team wants to ship.

What is wrong with declaring victory now, and what should the team have done instead?

Your answer

Solving needs a free account

Answers, streaks and solutions unlock when you are signed in. Reading the question and the hint stays free.

Discussion

Sign in to join the discussion · reading is open to everyone

💡 Discussion rules

  1. No full solutions here. Hints and approaches only.
  2. Complexity, edge cases and intuition are the point.
  3. Interview experiences are welcome. Respect your NDAs.

Loading discussion…

Learn the concepts

The theory behind this question.

Related questions

Why you can't stop an A/B test when it "hits significance"The test isn't significant, can we just run it longer?An automated alert that pings the moment p drops below 0.05Checking a retention test every morning for a monthKilling a subject-line test the instant it shows a winner
All questions →