Qm

Checking a retention test every morning for a month

A mobile game studio tests a new onboarding tutorial, measuring day-7 retention. The experiment is meant to run for 30 days. The product manager checks the dashboard every morning and has told the team: "the first day it shows a significant win, we lock it in." On day 11 the tutorial is ahead with p=0.045p = 0.045.

Is a day-11 significant result trustworthy under this habit, and what is the sound way to run the test?

Your answer

Solving needs a free account

Answers, streaks and solutions unlock when you are signed in. Reading the question and the hint stays free.

Discussion

Sign in to join the discussion · reading is open to everyone

💡 Discussion rules

  1. No full solutions here. Hints and approaches only.
  2. Complexity, edge cases and intuition are the point.
  3. Interview experiences are welcome. Respect your NDAs.

Loading discussion…

Learn the concepts

The theory behind this question.

Related questions

Why you can't stop an A/B test when it "hits significance"The test isn't significant, can we just run it longer?60 neutral button colors at the 5% levelWhy stopping an A/B test at first significance backfiresAn automated alert that pings the moment p drops below 0.05
All questions →