Sample Size
Binary independent proportions with a two-proportion normal approximation. Multi-arm designs apply Bonferroni correction across arms − 1 comparisons.
Experiment design & evaluation
Plan the sample your experiment needs, then evaluate Control versus Variant results with a transparent two-proportion methodology — including Bonferroni multi-arm correction and Newcombe intervals.
Methodology posture
Fixed two-proportion formulas, explicit validation, and conservative withhold rules when asymptotic assumptions break down — so results stay credible for planning and evaluation.
Workspace
Choose a mode, enter values, then calculate. Inputs stay in memory on this page only — nothing is written to the URL or storage.
Switching modes does not recalculate. Each mode keeps its own inputs and last result until you reset it or refresh the page.
Use this mode before launch to estimate the sample and duration needed to detect a meaningful conversion difference at your chosen Confidence level and Statistical power.
Experiment planning preview
Preview the conversion pathway from your baseline and MDE. Required sample and duration appear only after you press Calculate sample size.
Conversion pathway
Current conversion
—
Minimum detectable effect
—
Target conversion
—
After calculation
Daily experiment traffic is divided across the selected arms to estimate sample and whole-day duration.
Example: A 5% baseline with a 20% relative uplift means planning to detect an increase from 5% to 6% — not a jump to 25%.
Experiment comparison preview
Preview each arm's conversion rate as you enter counts. Difference, uncertainty and decision appear only after Calculate significance.
Control
Variant
After calculation
Sparse or degenerate data may produce a withheld result instead of an unreliable winner — a methodological safeguard, not an application error.
Methodology overview
A concise overview of the statistical approach behind sample-size planning and significance evaluation.
Binary independent proportions with a two-proportion normal approximation. Multi-arm designs apply Bonferroni correction across arms − 1 comparisons.
Pooled two-proportion z-test, Wilson intervals per arm, and a Newcombe hybrid-score interval for the absolute difference.
Sparse expected counts, degenerate variance, or p-value/CI conflict produce a withheld decision rather than an overconfident winner.
Statistical values keep full precision internally. Rounding, percentage formatting and “Unavailable” labels apply only when results are shown.
Key assumptions
Practical limitations
Common mistakes
Use the calculator as a planning and evaluation aid — not as a substitute for experiment design discipline.
Repeated peeking without a sequential design inflates false positives. Size the test first, then evaluate at the planned sample.
Extra variants increase the chance of a spurious winner. Sample Size mode applies Bonferroni correction when the experiment has more than two arms.
Tiny effects can require impractical samples. Use an MDE tied to a business decision, not the smallest imaginable lift.
A decision needs rates, uncertainty, assumptions and an explicit significance rule — not a green/red chart cue.
Related Growth Tools
Move from experiment maths into measurement maturity, unit economics, or the broader Growth Tools overview.
Audit whether tracking, attribution and revenue data are reliable enough to trust experiment outcomes.
Open Growth AssessmentTranslate conversion lifts into unit-economics impact through contribution LTV and CAC payback.
Open LTV & Payback CalculatorBrowse the live decision tools and see how assessment, modelling and experimentation fit together.
View Tools OverviewDisclaimer
Results are provided for informational and analytical purposes only. They are based on user-provided data and statistical assumptions and do not constitute financial, legal, medical or other professional advice. Read the full Disclaimer
From the practice
The A/B Test Calculator gives you a disciplined sample-size and significance read. Turning experiment design into a reliable measurement and decision system is what the Lazarevych growth-analytics practice does next.
Service
Explore Analytics Infrastructure
Ensure experiment tracking, assignment and conversion events are trustworthy before scaling tests.
Case study
Read: Growth Data Infrastructure & Funnel Economics Audit
See how measurement gaps that undermine experimentation were diagnosed in practice.
Insight
Read: GA4 Audit Checklist
Strengthen event and conversion measurement literacy around experiment reads.
Need implementation support? Turn the result into a scoped analytics, measurement or unit-economics workstream.
Discuss Implementation with Maksym