"The imbalance is tiny, just trust the lift", why not?
An experiment intended a 50/50 split, but the SRM check fails badly (chi-square far above the threshold): control got noticeably more users than treatment. The dashboard still shows treatment beating control by a clean, "significant" +4% on the primary metric. A colleague argues: "The imbalance is only a percent or two, and the lift is big and significant, let's just ship it."
Explain why a sample-ratio mismatch invalidates the result even when the lift looks strong, what typically causes SRM, and what you should do.
Your answer
Solving needs a free account
Answers, streaks and solutions unlock when you are signed in. Reading the question and the hint stays free.
Discussion
Sign in to join the discussion · reading is open to everyone
💡 Discussion rules
- No full solutions here. Hints and approaches only.
- Complexity, edge cases and intuition are the point.
- Interview experiences are welcome. Respect your NDAs.
Loading discussion…
Learn the concepts
The theory behind this question.
Related questions
Your 50/50 split came out 50.2/49.3, is that a problem?Is your 50/50 split actually 50/50? A sample-ratio checkA 24,600 vs 25,400 split on 50,000 users, real or noise?When a 5,050 vs 4,950 split is perfectly fineSRM on an uneven 90/10 holdback split
All questions →