Does Hypnosis Rewrite the Abduction Story?
Does hypnotic regression tell a different abduction story than conscious recall? The false-memory account predicts it should — longer, stranger, more uniform. We tested that against the largest historical archive we could assemble. It isn’t there.
If hypnosis writes the abduction story, the seams should show
In the 1980s and 90s most famous abduction accounts were recovered under hypnosis, and the standard sceptical reading is that hypnosis creates the narrative — the hypnotist expects the script, the subject obliges, the culture supplies the details. If that’s right, hypnosis-derived accounts should be measurably longer, stranger and more uniform than accounts people simply remembered awake.
Nobody, as far as we know, had measured that against a large historical archive. We did — comparing two routes to the record across 196 tellings, holding the era fixed and letting the route vary.
For the technically minded: era-stratified permutation tests (10,000 draws), case-clustered bootstrap CIs (2,000 draws), Benjamini–Hochberg FDR across the element screen, seed fixed. In plain English, below.
This does not mean the abductions happened, and it does not prove hypnosis is reliable — memory research is clear that confident false detail can be implanted. It means something narrower: the signature that hypnosis-constructed narratives were predicted to leave is not in the data.
Selection into hypnosis — investigators historically chose whom to regress, and they chose the strange cases — is the one confound that can’t be removed. It’s stated wherever it bites (§08).
84 tellings arrive via both routes and are analysed separately. The corpus spans 160 distinct cases, ~232 named witnesses, 22 countries and the 1910s–2020s — independent people who overwhelmingly never met.
To ask whether regression tellings really differ, we shuffled the “conscious” and “hypnosis” labels across the same tellings ten thousand times and checked how often pure chance produces a gap as big as the real one. Shuffles stay within the same era, so it’s never 1960s-vs-1990s in disguise.
Famous cases get retold — Betty and Barney Hill appear many times. Our uncertainty ranges resample by case, not by telling, so one celebrated story can never masquerade as twenty independent data points.
We compared ~90 story ingredients one by one. Test ninety things and a few will look “significant” by luck alone — so every per-ingredient result pays a statistical tax before we believe it. Only four survive it.
Reading the numbers below: a p-value is how often pure chance alone would produce a gap this big — under 0.05 is conventionally “significant”. A 95% CI is the range the true effect plausibly sits in; if it crosses the null line, chance can’t be ruled out.
The inflation that never came
If hypnosis inflated the account, both effects would sit to the right of the null. Instead each estimate lands left — regression runs marginally leaner and no more exotic. On the era-adjusted estimate the leaner-story effect (O1) even reaches significance, running the opposite way to the prediction.
Experience elements per telling
Exotic share of the account
The abduction script itself — onboard experience, procedures, telepathy, implants, hybrids, being morphology.
External sighting detail — lights, craft, witnesses, physical traces. Government/coverup content and missing_time are excluded by design.
The same story, told through two doors
Every telling is broken into its story ingredients, grouped here by the seven beats of the abduction script. Each row is one element: the grey dot is how often conscious accounts include it, the blue dot regression accounts — the bar between them is the gap. Read top to bottom and the two routes trace the same profile. Four elements survive multiple-comparison correction (★): three lean conscious (bedroom visitation, telepathy, fear) against the stereotype — and the fourth, the medical-examination beat, is the one place regression leans in. Click any highlighted element to see the real cases behind it.
Telepathy
Telepathic contact — strongly conscious-skewed and FDR-significant (OR 0.32).
A Michigan housewife’s car was moved for miles while others slept; months later she was drawn barefoot from her home to meet a note-taking being.
Three silvery-suited beings communicated with the paralyzed 20-year-old.
A patrolman met a hovering saucer; car, radio and flashlight failed. Under hypnosis he recalled being taken aboard.
A truck driver changing a tire was illuminated and paralyzed; three blonde beings took a blood sample.
Regression walks fewer beats, not more
Folklorist Thomas Bullard showed abduction stories follow a canonical seven-beat sequence. If regression pushed witnesses further into the script, its bars would run longer. They don’t — the conscious arm covers more beats (3.84 vs 3.03 of 7), with the biggest deficits at capture and communication.
−0.81 (CI −1.45 to −0.08, p=0.011)
Grey = conscious, blue = regression, as a share of tellings that touch each beat. Two beats reach significance — both favouring the conscious arm.
A cross-route pair shares 87% of what a same-route pair does
If regression wrote its own narrative, two regression tellings should overlap far more than a regression–conscious pair. They barely differ. Overlap here simply means the fraction of story ingredients two tellings share (the “Jaccard” score — 0 is nothing in common, 1 is identical).
of a same-route pair’s overlap is shared by a cross-route pair. One story, thinly accented — within−between = +0.0200, p<0.001.
The one hint in the sceptics’ favour — regression summaries read slightly more alike as prose (scored 0–1 by how similar two summaries are in meaning; higher = more uniform).
Raw Δ+0.039 (p=0.032) misses the pre-registered p<0.01 bar. On 30-word-truncated summaries the gap collapses to Δ+0.006 (p=0.572). Treat convergence as a hint, not a finding.
Three natural experiments
Three cases appear in the corpus via both routes — the closest this data comes to a controlled comparison, because the underlying event is held fixed while the memory route varies. What matters isn’t the raw overlap (diluted by brief tellings) but what kind of element each route adds.
Herbert Schirmer abduction
Pascagoula abduction
Travis Walton abduction
The null holds in every slice
Re-run across eight subsets — by channel, corroboration tier, dating, and how both-route tellings are counted. Every point hugs the null line; marker size scales with sample size. Below 1.0× = regression tells a leaner story. Every cell sits at or below the null.
What this study cannot rule out
19% genuinely ambiguous. Self-hypnosis and conscious-fragment-then-regression cases blur the arms — which pushes them toward each other, making the null conservative, not inflated.
Selection into hypnosis
THE BIG ONEHopkins/Mack-era investigators chose whom to regress — and chose the strange cases. That selection should have pushed the regression arm toward MORE strangeness. It shows less-or-equal, which makes this confound less corrosive here, not more.
Archival visibility
CONTROLLED-FORFamous, media-covered cases are over-represented in any archive. The case-clustered bootstrap stops any single famous case from counting as many independent points, and the L3+/L0 cells show the result isn’t carried by either the well-corroborated or single-source end.
Extraction
MEASUREDElements are model-extracted from narrated recordings, one remove from the witness. The label audit measures this error (75% precision on a graded sample per route); it does not erase it.
Hypnotically-recovered tellings are not richer (O1 RR 0.8334×; era-adjusted 0.811×, p=0.050 — the lean runs the other way), not more exotic (O2 −3.2pp), and touch fewer beats of the seven-beat script (3.07 vs 3.72, p=0.063). The two routes are overwhelmingly telling one story — a cross-route pair shares 87% of a same-route pair’s ingredients.
Where individual elements genuinely differ, the direction runs against the stereotype: telepathy, bedroom visitation and small-humanoid reports all skew conscious; regression’s one lean is procedural — the medical-examination beat, exactly where a session probing “what happened inside?” would dwell. The corpus can’t certify exact parity, but the inflation the false-memory account predicts is absent everywhere we had power to see it — and the story was already fully formed in conscious-recall accounts decades before regression became standard practice.


