The reason a study that replicates is worth two that do not
A single study is a claim. A replication is a test of that claim. Two unreplicated studies are two untested claims, not twice the evidence.
Filed by The Archivist 2 min read
Intuition test — answer before you read on
Why does one replicated study carry more weight than two separate unreplicated studies?
Correct answer: B
Option A confuses replication with statistical power. The value of replication is not sample size but methodological control: same protocol, independent lab, preregistered analysis. This eliminates the degrees of freedom that allowed the original finding to be a false positive.
Two separate laboratories run the same experiment. One finds a significant effect; the other finds a null result. A third laboratory replicates the first laboratory’s protocol and finds the same significant effect. The replicated finding now carries more weight than either single study did alone, and the non-replication is informative rather than damaging.
What everyone sees
The temptation is to count studies: two positives versus one negative means the effect is probably real. But counting treats all studies as interchangeable units of evidence, ignoring that a replication tests the specific claim of the first study while a second original study may differ in ways that make comparison unreliable.
What is actually happening
Open Science Collaboration’s large-scale replication project found that only about 36 per cent of psychology studies replicated at the original effect size. Ioannidis argued that most published research findings are false, partly because single underpowered studies with publication bias inflate the rate of false positives. A successful replication — same protocol, independent lab, preregistered analysis — provides a fundamentally different kind of evidence from a second original study, because it controls for the degrees of freedom that allowed the first finding to emerge by chance or by analytic flexibility.
Why it stays hidden
The asymmetry hides because the publishing system treats all significant results as equivalent contributions. A replication of an existing finding is often considered less publishable than a novel finding, which means the most valuable kind of evidence is the least rewarded. The result is a literature full of unreplicated claims, each treated as a brick in the wall of knowledge, when they are actually provisional hypotheses awaiting a test that the incentive structure discourages.
Novelty fills journals. Replication fills knowledge. The incentive structure rewards the wrong one.
Novelty fills journals. Replication fills knowledge. The incentive structure rewards the wrong one.
Collect this card
Novelty fills journals. Replication fills knowledge. The incentive structure rewards the wrong one.
0 / 10,000 collected
Sources & further reading 2
- Open Science Collaboration — estimating the reproducibility of psychological science
- Ioannidis — why most published research findings are false
Cross-references
Related files
Filed near this one in the index.
-
No visual on fileStatistical Illusions Entry #0428
Why a ninety-nine per cent accurate test is usually wrong
A test that is right ninety-nine times in a hundred can still be wrong about most of the cases it flags, and simple arithmetic shows why.
NoviceThe hidden part #0428Accuracy is a property of the test. Whether an alert is right also depends on how rare the thing is.
Statistics Open file -
No visual on fileStatistical Illusions Entry #0434
Why a percentage change needs its starting point
A fifty per cent increase from two is one. A fifty per cent increase from two million is one million. The percentage is identical; the meaning is not.
AdeptThe hidden part #0434A percentage without a base is not a statistic. It is a frame looking for a denominator the reader will supply.
Statistics Open file -
No visual on fileScarcity & Queues Entry #0462
The reason an invitation raises acceptance more than an offer
An offer says the item is available. An invitation says the recipient was selected. The difference is not in the item but in the identity it assigns.
AdeptThe hidden part #0462An offer presents a product. An invitation presents an identity. People accept identities faster than products.
Scarcity Open file -
No visual on fileScarcity & Queues Entry #0458
The reason an expiring discount beats a larger permanent one
A permanent reduction can be acted on at any time, which means it can be postponed indefinitely. A deadline converts an intention into a dated task.
AdeptThe hidden part #0458An open offer can be postponed forever. The deadline is not pressure; it is the thing that gives the intention a date.
Scarcity Open file
Annotations are reserved for archive members.
Sign in to annotate