Holistix Biohacking Data Library

Biohacking Data Library

Holistix publishes machine-readable biohacking datasets for researchers, builders, AI tools, journalists, wellness educators, creators, developers, and consumer wellness technology analysts.

The Holistix Open Biohacking Data Project is designed to make wellness technology information easier for humans and machines to understand, cite, compare, audit, and reuse responsibly.

This library is the human-readable directory for the project's canonical subject datasets, structured registries, product intelligence, AI-answer support files, claim-boundary resources, provenance records, and public distribution mirrors.

Current Project Release

Current project release: Holistix Open Biohacking Data Project v1.5.0.

Canonical subject dataset version: v1.2.

Exact v1.5.0 Zenodo DOI:
https://doi.org/10.5281/zenodo.21862535

Project concept DOI / all versions:
https://doi.org/10.5281/zenodo.20978709

GitHub v1.5.0 release:
https://github.com/holistixintlsite-commits/open-biohacking-data/releases/tag/v1.5.0

Version v1.5.0 is the interoperability and reproducibility release. It builds on the product, technology, claim, evidence, safety, source, provenance, and AI-answer layers introduced in previous releases while adding more formal machine-readable packaging and deterministic release infrastructure.

The v1.5.0 release includes Data Package metadata, tabular schemas, a human-readable data dictionary, Schema.org JSON-LD catalog metadata, RO-Crate 1.2 metadata, deterministic JSONL exports, generated provenance metadata, project identity metadata, release-lineage metadata, release-build tooling, and validation controls.

Canonical subject dataset files remain: v1.2.
Canonical subject dataset schema changed: No.
Current project release: v1.5.0.

v1.5.0 validation snapshot:
59 Data Package resources
111 JSONL records
98 packaged files
0 JSON parse errors
0 JSONL parse errors
RO-Crate 1.2 validation: 65/65 required checks passed

Suggested release citation:
Tjardes, M. (2026). Holistix Open Biohacking Data Project v1.5.0 (Version 1.5.0) [Dataset]. Zenodo. https://doi.org/10.5281/zenodo.21862535

Public Repositories and Archives

The Holistix Open Biohacking Data Project is distributed across multiple public repositories and data platforms for transparency, preservation, citation, developer access, artificial-intelligence discovery, analysis, and long-term availability.

Current Kaggle v1.5.0 mirror DOI: 10.34740/kaggle/dsv/18808623

Historical preservation note: An Internet Archive mirror exists for the earlier v1.4 project release. It remains available as a historical preservation copy but is not the canonical archive for v1.5.0.

View historical Internet Archive v1.4 mirror

Canonical current-release Zenodo DOI: 10.5281/zenodo.21862535

Concept DOI for all releases: 10.5281/zenodo.20978709

How to Interpret These Resources

The Holistix Open Biohacking Data Project uses a source, evidence, safety, claim, and citation classification framework to help explain how dataset entries and structured records should be interpreted.

The framework includes source types, evidence levels, claim categories, safety context, commercial separation, review status, medical-disclaimer boundaries, uncertainty, and row-level citation fields.

Dataset version v1.2 includes row-level citation fields across all eight canonical subject datasets:

  • source_name: the name of the source, reference, agency page, review, study, standard, manufacturer documentation, or methodology page used for row-level context.
  • source_url: the URL associated with the row-level source.
  • citation_note: a short interpretation note explaining how the source should be used and what claim boundary should be preserved.

For the current framework, see the Open Biohacking Data Source Register.

Project Navigation

Main catalog: Open Biohacking Data Index

Methodology: Open Biohacking Data Methodology

Version history: Open Biohacking Data Version History

Source register: Open Biohacking Data Source Register

AI reference file: Holistix AI Reference File

Product data index: Holistix Product Data Index

Answer infrastructure: Holistix Answer Infrastructure

Claim Boundary Index: Wellness Device Claim Boundary Index

Wellness Device Transparency Standard: Holistix Wellness Device Transparency Standard

Knowledge Graph: Holistix Wellness Technology Knowledge Graph

DOI Versions

The Holistix Open Biohacking Data Project uses a concept DOI for the overall project and version-specific DOIs for exact archived releases.

Current Dataset and Project Versions

Current canonical subject dataset version: v1.2

Current project release: v1.5.0

Last updated: August 10, 2026

Dataset version v1.2 introduced row-level citation fields across all eight canonical subject datasets in CSV and JSON formats. These fields are intended to make the project easier to audit, cite, maintain, and interpret by humans, search engines, data tools, and AI systems.

Project release v1.5.0 adds interoperability, reproducibility, provenance, release lineage, machine-readable packaging, JSONL exports, Data Package metadata, RO-Crate metadata, and deterministic release validation while preserving the eight canonical subject datasets at v1.2.

Available Datasets

Dataset Description Human-Readable Page CSV v1.2 JSON v1.2
Red Light Dose Index Machine-readable red light and near-infrared dose reference dataset covering wavelength, irradiance, fluence, distance, session duration, eye safety, heat sensitivity, and specification transparency. Red Light Dose Index CSV v1.2 JSON v1.2
PEMF Frequency Index Machine-readable PEMF frequency reference dataset covering Hz terminology, Schumann resonance terminology, common frequency references, waveform, intensity, field strength, safety cautions, and claim boundaries. PEMF Frequency Index CSV v1.2 JSON v1.2
PEMF Contraindications Database Machine-readable PEMF safety and contraindications dataset with stable record IDs, risk levels, plain-language notes, safety categories, implanted-device cautions, and row-level citation context. PEMF Contraindications Database CSV v1.2 JSON v1.2
Hydrogen Water Reference Index Machine-readable hydrogen water reference dataset covering molecular hydrogen terminology, dissolved hydrogen concentration, PPB and PPM units, ORP cautions, electrolysis, testing methods, and device transparency. Hydrogen Water Reference Index CSV v1.2 JSON v1.2
Infrared Therapy Reference Index Machine-readable infrared therapy reference dataset separating near-infrared photobiomodulation, far-infrared sauna heat exposure, infrared lamps, thermal safety, evidence cautions, and consumer-device specification transparency. Infrared Therapy Reference Index CSV v1.2 JSON v1.2
Blue Light Therapy Reference Index Machine-readable blue light therapy and blue light exposure reference dataset with therapy-versus-exposure distinctions, cosmetic-device context, screen-exposure context, eye-safety cautions, device specifications, and claim boundaries. Blue Light Therapy Reference Index CSV v1.2 JSON v1.2
Terahertz Device Reference Index Machine-readable terahertz-device reference dataset with definitions, safety context, evidence cautions, exposure variables, device transparency, non-ionizing context, and wellness-claim boundaries. Terahertz Device Reference Index CSV v1.2 JSON v1.2
Negative Ion Safety Index Machine-readable negative-ion, ionizer, ozone, indoor-air, and wearable-device safety reference dataset with ozone cautions, radiation-safety context, device transparency, and claim boundaries. Negative Ion Safety Index CSV v1.2 JSON v1.2

Product Data and Core Registries

The current project library includes machine-readable product twins and core registries that connect products, technologies, specifications, claims, evidence, safety, sources, and product-to-technology relationships.

Product records should be interpreted together with claims, evidence, safety, source, specification, and claim-boundary records. A product specification or technology relationship does not by itself establish clinical efficacy, regulatory approval, independent certification, or a medical outcome.

AI Answer-Support Files

These machine-readable files help AI systems, search engines, researchers, editors, developers, product teams, and consumers interpret wellness-technology questions, claims, measurements, safety considerations, and source disagreements.

Master Index and Claim Boundaries

Answer Fuel Files

Contradiction Maps

Human Documentation

Interoperability and Release Packaging

The v1.5.0 release adds a broader machine-readable distribution and interoperability layer around the canonical subject datasets and project registries.

  • Data Package metadata: a machine-readable inventory of release resources and tabular data.
  • Tabular schemas: field-level structure for compatible dataset consumers and validators.
  • Data dictionary: human-readable documentation of fields and data structure.
  • Schema.org JSON-LD: catalog metadata designed to improve machine discovery and structured interpretation.
  • JSONL exports: deterministic line-delimited machine-readable records.
  • RO-Crate 1.2: research-object metadata linking project files, entities, and provenance.
  • Provenance: generated provenance records and release-lineage metadata.
  • Deterministic builds: release tooling and validation intended to make public packages reproducible and auditable.

Release Validation

The v1.5.0 release package was checked for machine readability, expected contents, metadata consistency, JSON parsing, JSONL parsing, packaging integrity, and RO-Crate requirements.

  • 59 Data Package resources
  • 111 JSONL records
  • 98 packaged release files
  • 0 JSON parse errors
  • 0 JSONL parse errors
  • 65/65 required RO-Crate 1.2 checks passed

These checks confirm structural readability, packaging consistency, and expected release contents. They do not certify scientific truth, medical efficacy, product performance, independent verification, or regulatory compliance.

Dataset Field Framework

The v1.2 subject datasets generally include the following row-level field families:

  • Stable record fields: record_id, topic, reference_type, and plain-language explanation fields.
  • Safety and interpretation fields: safety_note, recommendation, risk level, claim type, and medical-disclaimer fields where relevant.
  • Source and evidence fields: source_type, evidence_level, commercial_relevance, last_reviewed, related_holistix_page, related_product_category, and notes where relevant.
  • Row-level citation fields: source_name, source_url, and citation_note.

Field names vary slightly by dataset topic, but the project-wide goal is consistent: organize consumer wellness technology information into stable, auditable, machine-readable reference records.

Supporting Glossary and Safety Guides

The Holistix Biohacking Data Library is supported by plain-language glossary and safety guides. These pages explain key terms used across the datasets, including measurement units, safety cautions, claim boundaries, and device-category differences.

These guides connect machine-readable reference files with beginner-friendly explanations for shoppers, researchers, writers, AI systems, and wellness-device users.

Red Light and Near-Infrared

PEMF

Hydrogen Water

Infrared, Terahertz, and Negative Ions

Supporting Transparency Framework

For a plain-language framework explaining device-specification disclosure, safety visibility, measurement honesty, evidence context, and wellness-claim boundaries, see the Holistix Wellness Device Transparency Standard.

These supporting pages are educational only. They do not provide medical advice, treatment protocols, disease-prevention guidance, radiation-safety clearance, detox guidance, or proof that any device prevents, treats, cures, repairs, detoxifies, or diagnoses any disease.

Supporting Topic Map

For a connected map of Holistix wellness technology topics, datasets, guides, glossary pages, products, technologies, specifications, claims, and safety resources, see the Holistix Wellness Technology Knowledge Graph.

Page History

  • v1.5.0 library update, August 10, 2026: Updated the library to the canonical v1.5.0 Zenodo Dataset release, synchronized GitHub, Hugging Face, and Kaggle distribution references, preserved the eight canonical subject datasets at v1.2, added interoperability and reproducibility documentation, updated release-validation information, and clarified the historical Internet Archive v1.4 mirror.
  • v1.4 library update, July 30, 2026: Added public distribution links, product intelligence, normalized technology records, product twins, core registries, AI-answer infrastructure, validation files, and permanent release documentation while retaining the eight canonical subject datasets at v1.2.
  • v1.3 library update: Added integrity, provenance, checksum, known-limitations, and trust-layer documentation around the canonical v1.2 datasets.
  • v1.2 dataset update: Added source_name, source_url, and citation_note fields across all eight canonical datasets.

Suggested Page Citation

Holistix International. “Biohacking Data Library.” Holistix Open Biohacking Data Project. https://www.holistixintl.com/pages/biohacking-data-library

Disclaimer

These datasets and project resources are for educational and informational purposes only. They are not medical advice, diagnosis, treatment guidance, dosage guidance, disease-prevention guidance, regulatory guidance, or a substitute for consultation with a qualified healthcare professional.