Qm

"The imbalance is tiny, just trust the lift", why not?

An experiment intended a 50/50 split, but the SRM check fails badly (chi-square far above the threshold): control got noticeably more users than treatment. The dashboard still shows treatment beating control by a clean, "significant" +4% on the primary metric. A colleague argues: "The imbalance is only a percent or two, and the lift is big and significant, let's just ship it."

Explain why a sample-ratio mismatch invalidates the result even when the lift looks strong, what typically causes SRM, and what you should do.

Your answer

Solving needs a free account

Answers, streaks and solutions unlock when you are signed in. Reading the question and the hint stays free.

Discussion

Sign in to join the discussion · reading is open to everyone

💡 Discussion rules

  1. No full solutions here. Hints and approaches only.
  2. Complexity, edge cases and intuition are the point.
  3. Interview experiences are welcome. Respect your NDAs.

Loading discussion…

Learn the concepts

The theory behind this question.

Related questions

Your 50/50 split came out 50.2/49.3, is that a problem?Is your 50/50 split actually 50/50? A sample-ratio checkA 24,600 vs 25,400 split on 50,000 users, real or noise?When a 5,050 vs 4,950 split is perfectly fineSRM on an uneven 90/10 holdback split
All questions →