What did Charles Honorton and Ray Hyman disagree about?
The Honorton–Hyman disagreement is one of the most substantive methodological exchanges in parapsychology’s history. It began as an adversarial debate, produced a landmark joint document, and left several specific questions genuinely open. Before going into detail: the studies in the ESP-Nexus library represent an unmeasured share of the published literature on this question, so treat what follows as a summary of what the library holds rather than a settled account of the field.
Charles Honorton was a parapsychology researcher profiled on ESP-Nexus at https://esp-nexus.org/scientists/charles-honorton/.
Experiments
The core of the Honorton–Hyman exchange centered on the ganzfeld database — a body of psi experiments accumulated through the 1970s and early 1980s using sensory attenuation procedures. Honorton published a meta-analysis of 28 direct-hit ganzfeld studies in the March 1985 issue of the Journal of Parapsychology, explicitly as a response to Hyman’s own critical appraisal of the same database published in the same issue. Both men analyzed the same corpus; they reached different conclusions about what it showed.
Following their joint communiqué, Honorton built the autoganzfeld system at the Psychophysical Research Laboratories (PRL) in Princeton — a computer-automated testing environment designed to satisfy the methodological standards he and Hyman had co-authored. The resulting series formed the basis of the 1994 Bem and Honorton Psychological Bulletin paper, which Hyman then critiqued in the same issue.
Methodology
The 1985 exchange foregrounded three categories of methodological dispute:
- Selective reporting. Hyman argued the published database likely omitted null-result studies, inflating the apparent effect. Honorton countered that a trim-and-fill analysis made an implausibly large file-drawer necessary to explain away the results.
- Multiple analysis. Hyman identified that many studies used several outcome indices without pre-specifying which was primary, making chance significant results more likely. Honorton’s 1985 meta-analysis addressed this by restricting to a single pre-chosen index — direct hits — across all 28 studies.
- Methodological flaws and their correlation with results. Hyman identified 12 potential flaws and argued their presence correlated with significance. Honorton disputed the coding and the magnitude of this relationship.
The joint communiqué was the practical resolution at the procedural level: Honorton and Hyman co-authored a set of specific design and reporting recommendations for future ganzfeld research — including pre-registration of protocols, automated target selection and recording, blind judging, and randomization checks. Honorton then built the autoganzfeld to implement those standards directly. The communiqué documented what both agreed the existing database showed, and separately documented where they continued to disagree — most critically, over whether a demonstrated effect, if real, constituted evidence for psi.
Data
The evidence rows in the ESP-Nexus library span different, non-comparable metrics and cannot be pooled into a single figure. What the rows show:
| Study | k / N | Metric | Direction |
|---|---|---|---|
| Honorton (1978) | 26 studies, 1,000 sessions | p = 8 × 10⁻⁹ (p only) | Positive |
| Honorton (1986) | 23 sessions (FT2 series) | Hit rate = 0.435, p = .04 | Positive |
| Honorton (1990) | 39 studies | ES = 0.29 (Cohen’s h), z = 7.57, p = 6.8 × 10⁻¹⁴ | Positive |
| Honorton (1989) | 62 investigators | ES = 0.033 (z/√N), z = 12.13 | Positive |
| Bem (1994) | 10 studies, 329 sessions | ES = 0.59 (proportion index π), hit rate = 0.32, z = 2.89, p = .002 | Positive |
All five rows in the library point positive. However, the corpus coverage for this question has not been measured, so this pattern reflects what the library holds rather than the full published record.
Skeptical critiques
What critics argue.
Hyman (1985) argued — before the joint communiqué — that Honorton’s meta-analysis could not be taken as evidence for psi because the database suffered from multiple undisclosed analysis paths: studies had used several outcome measures without pre-specifying which was primary, so significant results were more likely by chance than the reported p-values indicated. Hyman (1985) also argued that methodological flaw scores correlated positively with study outcome, meaning better-controlled studies tended to show smaller or null effects.
After the autoganzfeld series, Hyman (1994) shifted the critique to the reformed database. He argued that the hit rate in the autoganzfeld was partly an artifact of target-pool composition: dynamic targets (video clips) yielded substantially higher hit rates than static targets (photographs), and the mix of target types across studies was not uniform or pre-specified. He contended that target-type confounding, rather than a genuine anomalous signal, could account for the pattern Bem and Honorton reported. Hyman (1994) also raised the issue of optional stopping — whether the series was truly pre-specified as to stopping rule — and questioned whether the autoganzfeld results were genuinely independent of the earlier database Honorton had already analyzed.
What the experimental data show.
The autoganzfeld series used automated target selection and recording, blind judging, and pre-specified primary outcomes — the specific procedural requirements Hyman had co-authored in the joint communiqué. Utts (1991) noted, reviewing the broader ganzfeld database, that while Hyman’s flaw-correlation argument was the first to attempt quantification of the relationship between procedural problems and outcomes, the relationship he identified remained disputed as to magnitude and interpretation.
Analysis.
Honorton and Hyman co-authored both the problem statement and a partial resolution: they agreed the existing database showed an overall significant effect not reasonably explained by selective reporting or multiple analysis alone, and they agreed on the design standards needed for a decisive test. Where they continued to diverge was the inferential step — whether a replicated anomalous effect, even a methodologically clean one, should be interpreted as evidence for a genuine psi process or as an unidentified artifact. That question was not resolved in the communiqué and was still active in Hyman’s 1994 commentary. Milton and Wiseman (1999) later noted that meta-analytic investigation of variables Bem and Honorton identified as important in the PRL work indicated other experimenters had not replicated those moderator effects in the few areas where this had been attempted.
For the full spoke page on the communiqué and its methodological standards, see Honorton: Hyman–Honorton Joint Communiqué and Methodological Standards.
| Paper | Reported finding | Effect / significance | Basis |
|---|---|---|---|
| Bem et al. (1994), Psychological Bulletin [source] | Autoganzfeld primary pooled effect. | ES 0.59, z = 2.89, p = .002, hit rate 0.32 | 10 studies; N = 329 sessions; 240 participants |
| Honorton et al. (1990), Research in Parapsychology (RIP) conference proceedings | Combined Ganzfeld database. | ES 0.29, z = 7.57, p = 6.8 × 10−14 | 39 studies |
| Honorton et al. (1989), Journal of Parapsychology | Overall cumulation by investigator – 62 investigators. | ES 0.033, z = 12.13 | 62 studies |
| Honorton et al. (1986) | FT2 series – overall direct hits. | p = .04, hit rate 0.435 | N = 23 sessions |
| Honorton (1978), Psi and States of Awareness (conference proceedings) | Full ganzfeld corpus replication tally. | p = 8 × 10−9 | 26 studies; N = 1000 sessions; 500 participants |
References
- Bem, D. J., & Honorton, C. (1994). Does Psi Exist? Replicable Evidence for an Anomalous Process of Information Transfer. Psychological Bulletin, 115(1), 4–18. https://doi.org/10.1037/0033-2909.115.1.4
- Honorton, C., Berger, R. E., Varvoglis, M. P., Quant, M., Derr, P., Hansen, G. P., Schechter, E., & Ferrari, D. C. (1990). Psi Ganzfeld Experiments Using an Automated Testing System: An Update and Comparison with a Meta-Analysis of Earlier Studies. Research in Parapsychology (RIP) conference proceedings.
- Honorton, C., & Ferrari, D. C. (1989). “Future Telling”: A Meta-Analysis of Forced-Choice Precognition Experiments, 1935-1987. Journal of Parapsychology, 53, 281–308.
- Honorton, C., Barker, P., Varvoglis, M., Berger, R., & Schechter, E. (1986). First-Timers: An Exploration of Factors Affecting Initial Psi Ganzfeld Performance. Proceedings of Presented Papers, Parapsychological Association Annual Convention (Free-Response and Spontaneous-Case ESP Studies section), 28–32.
- Honorton, C. (1978). Psi and Internal Attention States: Information Retrieval in the Ganzfeld. Psi and States of Awareness (conference proceedings), 79–100.
More questions answered
- Pool the effect sizes across every ganzfeld study you hold and compare your number to the published Storm and Tressoldi meta-analyses—do they agree?
- Has the effect size for PK data increased in the last 50 years?
- Tell me about the clairvoyance work of the last 50 years
- What trends can you see in precognition research?
- What trends can you see in ESP research in the last 50 years?
- What are parapsychology's current arguments to justify that the phenomena are real and should be taken seriously?