Statistical Significance in A/B Testing, Explained
What p-values, confidence levels, and significance actually mean in an A/B test — and the mistakes that make "significant" results wrong.
Companion tool: The A/B Test Calculator
IV · LEARN
Essays on the statistics of experimentation, written for people who run tests.
What p-values, confidence levels, and significance actually mean in an A/B test — and the mistakes that make "significant" results wrong.
Companion tool: The A/B Test Calculator
A practical walkthrough of sample size calculations: choosing an MDE, setting power, and translating the answer into test duration.
Companion tool: Sample Size
How Bayesian probability-to-be-best and expected loss compare to p-values and confidence intervals, and when each approach fits.
Companion tool: Bayesian
Why repeatedly checking a fixed-horizon test inflates false positives, with simulations, and how sequential methods like mSPRT fix it.
Companion tool: Sequential
How to detect sample ratio mismatch, the most common causes, and why you should throw away results from a test that fails an SRM check.
Companion tool: SRM Checker