The Peeking Problem: A/B Test Significance (2D)

Synthetic users stream into a control funnel and a variant funnel at rates you choose, a live two-proportion z-test tracks the p-value after every batch, and a growing strip-chart shows exactly when the result crosses the significance threshold. Switch between "peek every batch" and a pre-committed fixed sample size to watch, over dozens of repeated experiments, why checking a dashboard every day and stopping the moment it turns green inflates the real false-positive rate far above the advertised α — the optional-stopping bug behind a lot of unreliable mobile A/B test results.