Direct comparison
Internal vs. External Validity Explained
Internal validity means an effect is really caused by the study design; external validity means it generalizes. Compare threats, trade-offs, and fixes.
Side-by-side comparison
| Dimension | Internal Validity | External Validity |
|---|---|---|
| What it measures | Whether the independent variable actually caused the observed effect, within the study itself | Whether the finding generalizes beyond the study — to other people, settings, or times |
| Core question | Did X really cause Y in this study? | Would this result hold in other populations or real-world settings? |
| Main threats | Confounding variables, selection bias, history, maturation, testing effects, instrumentation, regression to the mean, attrition | Unrepresentative samples, artificial settings, reactivity/Hawthorne effects, narrow time window, selection-by-treatment interaction |
| Strengthened by | Random assignment, control groups, standardized procedures, blinding, pre-registration | Random/representative sampling, multi-site recruitment, naturalistic settings, replication |
| Typically highest in | Tightly controlled laboratory experiments and randomized controlled trials (RCTs) | Large-scale, multi-site field studies and naturalistic observation |
| Typically weakest in | Quasi-experimental or observational designs without random assignment | Narrow lab studies on small, homogeneous convenience samples |
| Relationship to the other | A prerequisite for meaningful generalization — but doesn’t guarantee it | Depends partly on internal validity being established first; improving one often costs the other |
| Originating framework | Campbell & Stanley’s 1963 typology of validity threats (extended by Cook & Campbell, 1979) | Same typology — internal and external validity were defined together as a paired framework |
Common questions
FAQ
Can a study have high internal validity but low external validity?+
Yes — this is the classic case for a tightly controlled laboratory RCT run on a narrow convenience sample. The causal conclusion within the study can be very strong while the finding may not generalize to other populations or real-world settings.
Which is more important, internal or external validity?+
It depends on the research question. Efficacy research (does this work under ideal conditions?) prioritizes internal validity. Effectiveness or policy-relevant research (does this work in real-world practice?) prioritizes external validity.
How does random assignment affect internal vs. external validity?+
Random assignment to conditions is primarily an internal-validity tool — it protects against selection bias. It does not by itself make a sample more representative of a broader population; that depends on random sampling during recruitment, a separate step.
Is external validity the same thing as generalizability?+
The terms are largely used interchangeably. Some methodologists draw a finer distinction between generalizability to the sampled population specifically and external validity more broadly, but in most research-methods usage they describe the same concern.
Do quasi-experimental designs improve or worsen this trade-off?+
Quasi-experimental designs typically have weaker internal validity than a true experiment because they lack random assignment. In exchange, they’re often run in more naturalistic settings, which can improve external validity — though this has to be evaluated design by design, not assumed.
Going deeper







