For Researchers & Academics

A systematic analysis of first-hand accounts from the Hardy Archive

Built for researchers studying religious experience, mystical states, and transformative phenomena, with documented methodology, cross-model validation, and full transparency on limitations.

The dataset

What the archive contains

This analysis draws on individual first-hand accounts from the Alister Hardy Religious Experience Research Archive, now held at the University of Wales Trinity Saint David in Lampeter. The Hardy Archive is the largest systematic collection of first-hand religious experience accounts in existence, containing over accounts submitted by members of the public in response to Hardy's original 1969 appeal and subsequent solicitations.

The accounts analyzed here are those that meet our criteria for individual first-hand accounts, meaning accounts submitted by a single person describing their own direct experience, as opposed to letters commenting on the project, secondhand descriptions, or institutional submissions.

Individual first-hand accounts analyzed
Total accounts in the full Hardy Archive
5
Analytical dimensions, each cross-validated
2
Independent AI models used for cross-validation

The dataset spans five analytical dimensions, each applied systematically across all accounts:

Experience Categories
17 categories, divine presence, light, communication, vision, mystical union, and more
Phenomenological Qualities
9 qualities, noetic quality, joy/bliss, peace, ineffability, sacredness, fear/awe, love, passivity, timelessness
Triggers
11 categories, crisis, prayer, spontaneous, nature, bereavement, sleep/dream, worship, meditation, near-death
Tradition Framing
7 categories, Christian, syncretic, spiritual-not-religious, secular/agnostic, Buddhist, Jewish, Hindu/Eastern
Lasting Effects
8 categories, faith deepened, faith transformed, meaning restored, fear of death removed, ethical change, healing effect
Veridical Claims
Boolean flag, whether the account includes a claim of verified external accuracy (present in % of accounts)
Pipeline

How it was built

The analysis was conducted through a multi-stage pipeline. The full methodology is documented at methodology.html, including the complete tagging prompt text, inter-rater reliability statistics, and a description of all known limitations.

  • Source selection, Accounts from the Hardy Archive meeting the individual first-hand criteria described above. The full archive spans submissions from 1969 to the present; the accounts in this analysis represent the digitized and processed portion available through the RERC.
  • Quality filtering, AI-assisted quality review using Claude Sonnet; accounts lacking sufficient content for meaningful coding were excluded before tagging.
  • Tagging, Each account was analyzed across all five analytical dimensions using structured prompts with defined vocabulary and explicit coding criteria. Tags were applied by Claude Sonnet.
  • Cross-validation, Every account was independently analyzed by a second AI model using equivalent prompts. Disagreements between models were reviewed through an adjudication pass that re-reads the full account and issues a final determination with per-category reasoning. of accounts went through adjudication.
  • Validation and export, Automated schema validation confirmed structural integrity across all accounts. Aggregated statistics were generated for the public-facing website.

Unlike the dual broad/strict coding used in our companion site (storiesofawakening.org), the Hardy Archive analysis uses single-pass tagging per category. The accounts are shorter on average ( words per account) and less ambiguous in their coding, the written, reflective format of the Hardy accounts produces clearer categorical signals than transcribed spoken interviews.

Research questions

What this dataset can help answer

The dataset is suited to descriptive and correlational questions about the phenomenology, context, and aftermath of religious experience across a large, heterogeneous sample. Examples:

  • What experience types and phenomenological qualities co-occur most frequently, and which tend to appear in isolation?
  • Does the trigger of an experience predict the qualities reported? Do crisis-triggered experiences differ phenomenologically from prayer-triggered or spontaneous ones?
  • How do lasting effects distribute across tradition framings? Do Christian accounts, secular accounts, and syncretic accounts differ in their reported effects?
  • How does the prevalence of veridical claims vary across experience types?
  • What is the distribution of experiences across Hardy's original category system, and how does AI-assisted coding compare to prior hand-coded analyses of archive subsets?
  • How does the phenomenology of near-death experiences in this archive compare to NDE accounts in interview-based datasets?

The dataset is not suited to: causal claims about what produces religious experience, analysis of narrative structure or linguistic patterns (the available data is categorical, not textual), individual case study work (the archive is held by the RERC and access to original accounts requires separate arrangement), or generalization to non-Western, non-English-speaking populations.

Limitations

What this dataset cannot do

The following limitations are documented here in full and should be considered in any use of this data.

1
Voluntary response bias. The Hardy Archive consists of people who chose to respond to a public appeal. People who had experiences and chose not to report them, and people who had experiences that were primarily negative, are underrepresented. The sample skews toward experiences deemed meaningful and shareable.
2
Temporal and cultural span. Accounts span more than fifty years and primarily reflect British social and cultural contexts. Religious terminology, attitudes toward disclosure, and the meaning attached to experience vary across this period in ways that the current tagging schema does not fully capture.
3
Christian overrepresentation. % of accounts come from Christian or broadly Christian-influenced framings, reflecting the demographic of respondents to Hardy's original British appeal. This limits generalization to other cultural and religious contexts.
4
Retrospective reporting. Most accounts were written after the fact, sometimes many years after. Retrospective accounts are subject to memory consolidation, meaning-making influenced by subsequent experience, and the filtering effect of deciding what to include in a written submission.
5
AI tagging error. Despite cross-model validation and adjudication, AI tagging at this scale introduces coding error that cannot be fully quantified. The inter-rater statistics in the methodology document give an upper bound on per-category reliability. Individual tags should be treated as probabilistic, not definitive.
Access

Access and collaboration

This site provides aggregated statistics, prevalence rates and categorical breakdowns across all five dimensions. Individual account texts are held by the Religious Experience Research Centre (RERC) at the University of Wales Trinity Saint David, Lampeter. Researchers wishing to access original accounts should contact the RERC directly at uwtsd.ac.uk/rerc.

The tagging dataset and pipeline code are not yet publicly released, but collaboration inquiries are welcome. If you are working on research related to religious experience and would like to discuss access or collaboration, please reach out via the Contact page.

For citation purposes, please use:

Religious Experiences Archive (2026). A systematic analysis of 2,080 first-hand accounts from the Alister Hardy Religious Experience Research Archive. religiousexperiences.org
Explore

Start with the methodology or the data.