Glossary · Analytics

Sample Size

SAM-pul syznoun

Sample size is the number of observations or visitors included in a test or study.

Part of speech
noun
Pronunciation
SAM-pul syz
Origin
From 'sample,' Old French 'essample' meaning example, plus 'size.' It refers to the number of observations included in a study.

What is Sample Size?

Sample size is the number of observations, visitors, or data points included in a test or study. Whenever you measure something about a large group by looking at only a portion of it, the sample size is how big that portion is. In marketing this usually means the number of visitors who saw each version of a page in an experiment, the number of people surveyed, or the number of sessions counted before drawing a conclusion. It is one of the most important and most underestimated factors in whether the results of any test can be trusted, because the entire logic of sampling rests on having enough of it.

The mechanics come down to how sample size interacts with noise. Any measurement drawn from a sample carries random variation, and the smaller the sample, the more that variation can distort the picture. A conversion rate calculated from thirty visitors can swing wildly on the strength of a single extra purchase, while the same rate measured across many thousands of visitors settles down and reflects reality far more closely. Statisticians use the intended confidence level, the smallest difference worth detecting, and the natural variability of the data to calculate the sample size a test needs before it starts. Reaching that predetermined number is what gives a result the statistical power to be believed, and falling short leaves the outcome dominated by chance.

The term is plainly built. Sample descends from the Old French essample, meaning an example or specimen, and it is paired with size, the measure of magnitude. Together they simply name how large your example is. The concept is as old as formal sampling itself, growing out of the same early twentieth century statistical work that gave us significance and confidence, and it applies identically whether the sample is people in a poll or visitors in a web experiment.

For a business, sample size determines whether testing produces knowledge or illusion. Teams eager for answers often end experiments the moment one version pulls ahead, but with a small sample that lead is frequently random and evaporates when more data arrives. Committing to an adequate sample size before launching a test protects against acting on false winners, and it also sets realistic expectations for how long a test must run given the available traffic. Sites with modest traffic must accept that reliable testing takes time or focus only on changes large enough to detect quickly, which is itself a valuable planning insight.

The nuances and common mistakes are consequential. The biggest error is stopping too early, which manufactures false positives and undermines the whole practice of experimentation. Another is imagining that a huge sample fixes everything, when a large but biased sample, drawn from an unrepresentative slice of your audience, will confidently mislead you no matter how many observations it contains. Sample size must also match the effect you hope to find: detecting a tiny improvement requires far more data than detecting a large one. Calculating the required size in advance, keeping the sample representative, and holding the discipline to wait for it are what turn raw traffic into dependable conclusions rather than expensive guesswork.

Why it matters

Sample size determines whether test results can be trusted, so planning it upfront prevents costly decisions based on too little data.