Open Biohacking Data Methodology

Open Biohacking Data Methodology

The Holistix Open Biohacking Data Project publishes structured, machine-readable reference datasets about wellness technologies, device categories, safety considerations, terminology, and comparison topics.

The goal of this project is to make biohacking device information easier for researchers, builders, educators, journalists, AI systems, and search engines to understand and reference.

Version history: Open Biohacking Data Version History

AI reference file: Holistix AI Reference File

  • Open Biohacking Data Source Register
  • Public Repositories and Archives

    The Holistix Open Biohacking Data Project is distributed across multiple public repositories for transparency, preservation, citation, developer access, artificial-intelligence discovery, and long-term availability.

    Canonical Zenodo DOI: 10.5281/zenodo.21574706

    Concept DOI for all releases: 10.5281/zenodo.20978709

    Current Archived Release

    Current project release: Holistix Open Biohacking Data Project v1.4.

    Exact v1.4 DOI:
    https://doi.org/10.5281/zenodo.21574706

    Project DOI / all versions:
    https://doi.org/10.5281/zenodo.20978709

    Previous v1.3 archive DOI:
    https://doi.org/10.5281/zenodo.21033668

    GitHub v1.4 release:
    https://github.com/holistixintlsite-commits/open-biohacking-data/releases/tag/v1.4

    Version v1.4 adds the product-data layer, machine-readable product twins, normalized technology relationships, structured claims, evidence, safety, and source registries, AI-answer infrastructure, citation metadata, validation files, and a permanent archived release.

    Release validation: 49 JSON files, 8 CSV files, 13 products, 18 technologies, 111 claims, 111 evidence records, 111 safety records, 111 dataset records, and zero JSON parse errors.

    Canonical dataset files remain: v1.2.
    Current project release: v1.4.
    Dataset schema changed: No for the eight canonical subject datasets.

    Suggested citation:
    Holistix International. (2026). Holistix Open Biohacking Data Project v1.4 (v1.4) [Data set]. Zenodo. https://doi.org/10.5281/zenodo.21574706

    Project Scope

    The project currently focuses on structured data for:

    • PEMF devices
    • Hydrogen water devices
    • Negative ion devices
    • Blue light therapy
    • Infrared therapy
    • Terahertz wellness devices
    • Related safety, terminology, and comparison topics
    • Machine-readable product twins
    • Normalized technology entities and product-technology relationships
    • Claims, evidence, safety, and source registries
    • AI-answer support files, contradiction maps, and claim-boundary resources

    Data Format

    Project resources are published in JSON and CSV where appropriate. The v1.4 archive includes 49 JSON files and 8 CSV files, along with citation, schema, validation, provenance, and release documentation.

    Each dataset may include fields such as:

    • Stable record ID
    • Device category
    • Topic or term
    • Plain-language explanation
    • Safety note
    • Evidence or source type
    • Source URL, when available
    • Last reviewed date
    • Version number

    Stable IDs

    Records use stable Holistix IDs when appropriate. These IDs are intended to make records easier to cite, update, compare, and connect across datasets.

    Example format:

    HOL-PEMF-CONTRA-0001

    Source Standards

    Datasets may draw from publicly available sources such as scientific literature, manufacturer documentation, regulatory resources, clinical references, standards organizations, and other public educational materials.

    Holistix aims to distinguish between:

    • Scientific or clinical references
    • Manufacturer safety documentation
    • Regulatory or standards-based information
    • Terminology definitions
    • General educational summaries

    Source and Evidence Classification

    The Holistix Open Biohacking Data Project uses a source and evidence classification framework to make dataset entries easier to audit, cite, maintain, and responsibly reuse.

    Project records may include fields such as source_type, source_name, source_url, evidence_level, claim_type, last_reviewed, commercial_relevance, medical_disclaimer_required, source_status, verification_status, and related product or technology identifiers.

    These fields are designed to help distinguish between basic terminology, device specifications, consumer safety cautions, research context, manufacturer claims, editorial explanation, and emerging technology notes.

    For the current framework, see the Open Biohacking Data Source Register.

    Product Data and Registry Methodology

    The v1.4 project release includes 13 machine-readable product twins and eight core product, technology, specification, relationship, claims, evidence, safety, and source registries.

    Product records are designed to keep the following categories distinct:

    • manufacturer-reported specifications
    • independently checked details
    • approximate values
    • unknown or unavailable measurements
    • general technology research
    • product-specific evidence
    • safety cautions and contraindications
    • commercial product descriptions

    A product specification or technology relationship does not by itself establish clinical efficacy, regulatory approval, independent certification, or a medical outcome. Product records should be interpreted together with related claim, evidence, safety, and source records.

    Human-readable directory: Holistix Product Data Index.

    Validation and Quality-Control Methodology

    Before the v1.4 archive was published, the release package was checked for file inventory, JSON parsing, record counts, registry consistency, and expected project components.

    The archived validation summary reports:

    • 49 JSON files
    • 8 CSV files
    • 13 product records
    • 18 technology records
    • 111 claims records
    • 111 evidence records
    • 111 safety records
    • 111 dataset records
    • zero JSON parse errors

    Validation confirms structural readability and expected record counts. It does not certify scientific truth, medical efficacy, product performance, or regulatory compliance.

    AI Answer Infrastructure Methodology

    The project includes claim boundaries, Answer Fuel Files, contradiction maps, and an AI Answer Infrastructure Manifest. These resources are intended to help AI systems, search engines, editors, researchers, and developers preserve measurement context, source uncertainty, and safety boundaries when generating answers.

    View the Holistix Answer Infrastructure.

    View the AI Answer Infrastructure Manifest v1.4.

    Neutrality and Limitations

    The datasets are educational reference materials. They are not medical advice, diagnosis, treatment recommendations, or a substitute for guidance from a qualified healthcare professional.

    Datasets are designed to describe device categories, terminology, safety considerations, and publicly documented information. They are not intended to claim that any device treats, cures, prevents, or diagnoses disease.

    Commercial Separation

    Holistix sells wellness and biohacking products. However, the Open Biohacking Data Project is maintained as a separate reference layer. Dataset records are designed to avoid direct product promotion and focus on structured educational information.

    Update Policy

    Datasets may be revised as new information becomes available, formatting improves, or additional sources are reviewed. Each dataset includes a version number and last updated date when available.

    Citation

    Suggested citation format:

    Holistix International. (2026). Holistix Open Biohacking Data Project v1.4 (v1.4) [Data set]. Zenodo. https://doi.org/10.5281/zenodo.21574706

    Supporting Glossary and Safety Guide Methodology

    The Holistix Open Biohacking Data Project includes supporting glossary and safety guides in addition to canonical machine-readable datasets.

    These support pages exist for four reasons:

    1. Plain-language interpretation: Many dataset fields use technical terms such as irradiance, fluence, PPB, ORP, Hz, Schumann resonance, non-ionizing radiation, NIR, FIR, and ozone-free ionizer. The support pages explain those terms for general readers.
    2. Claim-boundary clarity: Wellness-device terms are often used in marketing with exaggerated certainty. The support pages separate terminology, measurement context, safety notes, and conservative evidence interpretation from disease-treatment claims.
    3. Internal source navigation: The support pages connect each dataset to related educational pages, safety guides, source registers, and product-category explanations.
    4. Machine readability and topical structure: The support pages provide crawlable, structured explanations that help search engines, AI systems, researchers, writers, and users understand how the datasets relate to common wellness-device concepts.

    Supporting guides are not treated as independent clinical proof. They are explanatory bridge pages that point back to the canonical datasets, source register, methodology page, version history, and archived project release.

    Each support guide is written with conservative safety language and avoids claims that any consumer wellness device prevents, treats, cures, detoxifies, repairs, or diagnoses disease.

    Examples of Supporting Guide Categories

    • Red light and near-infrared: irradiance, fluence, wavelength, distance, NIR, and dose context.
    • PEMF: Hz, frequency, intensity, Schumann resonance, contraindications, and implanted-device cautions.
    • Hydrogen water: PPB, PPM, ORP, dissolved hydrogen concentration, bottles, tablets, and testing context.
    • Infrared: NIR vs FIR, sauna blanket use, heat, hydration, session time, and heat-safety boundaries.
    • Terahertz: non-ionizing terminology, exposure context, heat, eye caution, and evidence limits.
    • Negative ions: ionizers, ozone-free claims, ozone generators, indoor-air safety, and wearable-device boundaries.

    Relationship to Canonical Dataset Files

    The supporting glossary and safety guides do not replace the canonical dataset files. The canonical dataset files remain the structured reference layer. The guides are human-readable explanations that support interpretation of those datasets.

    When a support page discusses a term, the preferred structure is:

    • define the term in plain language
    • explain the measurement unit or device-category context
    • separate the term from related but different terms
    • include safety or contraindication context when relevant
    • link back to the appropriate dataset page
    • link to the source register and methodology page
    • avoid disease-treatment or guaranteed-outcome claims

    This approach keeps the project useful to beginners while preserving the more structured, machine-readable nature of the canonical datasets.

    Page History

    • v1.4 methodology update, July 30, 2026: Corrected all current v1.4 DOI references to the canonical dataset DOI 10.5281/zenodo.21574706, confirmed the GitHub release, added the full public distribution layer across Zenodo, Hugging Face, Kaggle, and Internet Archive, documented product twins and core registries, added validation and AI-answer infrastructure methodology, and preserved the canonical dataset version at v1.2.
    • v1.3 methodology update: Added integrity, provenance, checksum, known-limitations, authorship, and review-policy documentation.
    • Initial publication: Established the project scope, source standards, stable-ID structure, neutrality principles, update policy, and supporting-guide methodology.

    Last updated: July 30, 2026

    Disclaimer

    This project is for educational and informational purposes only. Consult a qualified healthcare professional before using wellness devices, especially if pregnant, managing a medical condition, using implanted electronic devices, or taking prescribed treatment.