Sample size is the number of people enrolled in a study. Statistical power is a study’s ability to detect a real effect if one truly exists. When a study enrolls too few people, it can miss real effects, produce results that look bigger than they really are, or generate “positive” findings that don’t hold up when other researchers try to repeat them. Understanding these two ideas helps you judge whether a study’s conclusions deserve your confidence.
This guide explains what sample size and statistical power mean, why small studies are more likely to mislead, and what questions to ask before you trust a result.
The Short Answer
A small study is not automatically wrong, but it is more fragile. With fewer participants, chance plays a bigger role in the outcome. A single unusual participant, a coincidence in timing, or normal day-to-day variation can swing the results. Larger, well-designed studies average out that noise, so their findings are more likely to reflect what actually happens across a wider population.
What Sample Size Actually Means
Sample size is simply how many people were studied. It matters because humans vary a lot — in genetics, health history, diet, and lifestyle. A study needs enough participants to represent that variation. If a study only includes 15 or 20 people, the results may reflect the quirks of that small group rather than a general pattern.
Researchers typically calculate a target sample size before a study begins. This calculation is based on how big an effect they expect to find and how much natural variation exists in the population. Studies that skip this step, or that fall short of their planned enrollment, are at higher risk of unreliable results.
What Statistical Power Means
Statistical power is the probability that a study will detect a real effect, if that effect truly exists. Researchers usually aim for at least 80% power, meaning the study has an 80% chance of finding a true effect rather than missing it.
Low-powered studies have two main problems:
- They can miss real effects. A treatment might genuinely work, but the study is too small to prove it, leading to a false “no effect” conclusion.
- When they do find a “significant” result, it is more likely to be a false positive or an exaggerated effect. This is a well-documented pattern in research on study design: small, low-powered studies that report a positive finding tend to overstate how large that effect really is.
Statistical Significance Is Not the Same as Clinical Importance
A result can be “statistically significant” — meaning it’s unlikely to be due to chance alone — without being meaningful in real life. For example, a supplement might show a statistically significant drop in a lab marker, but the actual size of that change could be too small to matter for a person’s health or daily symptoms.
Always ask two separate questions: Is this result statistically significant? And separately, is the size of the effect actually large enough to matter? A small study is more likely to blur these two questions together, especially in marketing materials that highlight the word “significant” without explaining what it means.
Checklist: Questions to Ask About Any Study’s Sample Size
- How many total participants completed the study? A few dozen participants is much weaker evidence than several hundred or several thousand.
- Did the study report a power calculation, or explain how it chose its sample size?
- Was the study a pilot or preliminary study? Pilot studies are meant to test feasibility, not to prove an effect works.
- Has the finding been repeated in a larger, independent study?
- Does the study report the actual size of the effect, not just whether it was “statistically significant”?
- Were there dropouts? A study that started with 100 people but only 60 finished has an effective sample size closer to 60.
Small Studies vs. Larger Studies
Here’s the general pattern, though it isn’t a strict rule — study quality also depends on design, population, measurement methods, and whether results have been replicated.
- Sensitivity to chance: Small studies (roughly under 50 participants) are highly sensitive to chance — a few unusual participants can shift the results. Larger studies (several hundred or more) let individual variation average out.
- Risk of false positives: Small studies carry a higher risk, especially with low statistical power. Larger, properly powered studies carry a lower risk.
- Typical use: Small studies are often early or pilot research meant to generate hypotheses. Larger studies are used to confirm whether an effect is real and how large it is.
- How to treat the result: Treat a small study’s findings as a starting point, not a conclusion. Treat a larger study’s findings as stronger evidence, especially if the results have been replicated elsewhere.
A well-designed small study can still be useful as a first step toward larger research — the size alone doesn’t disqualify it, but it does mean the result needs more confirmation before it’s treated as settled.
Where This Fits in the Evidence Hierarchy
Sample size is one part of a bigger picture. A well-powered, well-designed randomized controlled trial generally sits near the top of the evidence hierarchy, while small, uncontrolled, or preliminary studies sit lower. For a full walkthrough of how study design affects evidence strength, see Understanding Study Design Types and Evaluating Evidence Quality, which covers sample size alongside consistency and clinical significance in more depth.
To check the actual enrollment numbers and design of a specific study, you can look up its registration on ClinicalTrials.gov, which lists planned and actual enrollment for registered trials, or search the published paper on PubMed to read the methods section directly.
Extra Caution Groups
Be especially cautious about small-study findings when:
- The study involves pregnant or breastfeeding people, children, or older adults, since these groups are often underrepresented and results from a general adult sample may not apply.
- The study is used to support claims about a specific health condition, since a small sample may not reflect how people with that condition typically respond.
- The result is being used to promote a product, since marketing materials may cite the most favorable small study while leaving out larger studies that found no effect.
Red Flags to Watch For
- A study is described only as “clinically proven” or “scientifically shown” without stating how many people were involved.
- Marketing content cites one small study while ignoring larger or more recent research on the same topic.
- A study reports statistical significance but never states the actual size of the effect.
- A “study” turns out to be a small pilot or a case series with no comparison group.
Evidence Limits
This guide explains general principles of study design and statistics used across clinical research. It does not evaluate any specific product, supplement, or treatment, and it is not a substitute for reading a study’s full methods section. Sample size requirements vary by field, by the size of the effect being studied, and by the type of statistical test used, so there is no single number that makes a study automatically “big enough.”
Next Step
Before trusting a study’s conclusion, look up the original study on ClinicalTrials.gov or PubMed and check the actual number of participants, whether the study has been replicated, and whether the effect size — not just the significance — is meaningful. For guidance on reading the rest of a study, see How to Read a Clinical Study.
FAQ
How many participants does a study need to be reliable?
There is no single number that applies to every study. The right sample size depends on the size of the effect being studied, how much natural variation exists, and the type of statistical test used. Researchers calculate this in advance using a power analysis rather than relying on a fixed cutoff.
Can a small study ever be trustworthy?
Yes, small studies can be useful as early, hypothesis-generating research, especially pilot studies designed to test feasibility. But their findings should be treated as preliminary until confirmed by larger, independent studies.
What does “statistically significant” actually mean?
It means the result is unlikely to have occurred by random chance alone, based on a specific statistical threshold researchers set in advance. It does not by itself tell you whether the effect is large enough to matter in real life.
Where can I check how many people were in a specific study?
You can look up a registered trial’s planned and actual enrollment on ClinicalTrials.gov, or find the published paper on PubMed and read the methods section, which usually states the total number of participants and how dropouts were handled.
Educational Disclaimer
This article is for general educational purposes only. It explains concepts in research design and statistics and is not medical advice. It does not diagnose any condition, recommend any treatment, or evaluate any specific product. Always talk with a qualified healthcare provider before starting, stopping, or changing any medication or treatment.