Probability Lab // Cabinet 09
False Positive Arcade
Run the same A/B test again and again. Watch chance turn identical truths into conflicting headlines.
01 — Load the experiment
2%50%
−50%+100%
50010k
80%99%
02 — Race one test
True rates loaded: A 10.00% · B 10.00%
A
—converts
B
—converts
Insert a probability token.
Choose settings, then run a race. The same seed and settings replay the same result.
03 — The wall of outcomes
A wins
—
B wins
—
Inconclusive
—
No batch loaded yet.
Batch runs reveal how often one fixed experiment design points in each direction.
Significant for ASignificant for BNot significant
Educational simplification: independent binomial samples, a pooled two-proportion z test, and a two-sided confidence threshold. Controls keep expected conversions at five or more per arm, but this remains an approximation. Real experiment decisions also need power planning, data-quality checks, guardrails, and correction for repeated peeking or multiple comparisons.