Age Ayurveda Nighantu

Open Ayurvedic Datasets

Open, machine-readable datasets of classical Ayurvedic pharmacology, phytochemical compounds, and botanical taxonomy. Published under Creative Commons Attribution 4.0 International (CC BY 4.0).

All datasets published by the Age Ayurveda Nighantu are generated directly from verified monograph records and classical texts, preventing drift between the reference text and tabular data. Every botanical taxon is resolved against global taxonomic authorities (GBIF, NCBI, Wikidata), and every pharmacological property conforms to standard Ayurvedic Pharmacopoeia of India (API) criteria.

Available Datasets

Ayurvedic Dravyaguna Pharmacological Dataset

Rasa, guna, virya, vipaka and prabhava for 245 monographs, parsed from the pages on this site so every value can be checked against the page it came from.

Coverage: 245 monographs

Direct Downloads: JSON CSV Parquet

Phytochemical Constituents & Co-occurrence Network

838 constituents named in the published monographs, 170 of them resolved to PubChem, with 21,942 co-occurrence edges computed from the pages themselves.

Coverage: 838 constituents · 21,942 co-occurrences

Direct Downloads: JSON CSV Parquet

Biomedical Literature & Classical Research Index

2,892 papers cited on this site, each with its PubMed ID or DOI, cross-referenced to the monographs that cite it.

Coverage: 2,892 papers

Direct Downloads: JSON CSV Parquet

Verified Botanical Taxonomy & Identifier Crosswalk

Botanical names resolved against GBIF, with NCBI taxonomy IDs and Wikidata QIDs where an exact match exists. 285 taxa, each carrying the match type so a fuzzy hit is never read as a fact.

Coverage: 285 resolved taxa

Direct Downloads: JSON Parquet

Nighantu Editorial Verification Ledger

What was checked, how, and what was thrown out: every verification run with its examined, published and rejected counts, computed from the run artifacts rather than stated.

Coverage: 0 published safety records, and the rejections behind them

Direct Downloads: JSON

Hugging Face & Zenodo Access

The consolidated corpus is also hosted in Apache Parquet format on the Hugging Face Hub and archived with a persistent Concept DOI on Zenodo.

  • Hugging Face Dataset: ageayurveda/nighantu (supports direct streaming with datasets.load_dataset("ageayurveda/nighantu"))
  • Zenodo Archive: Persistent research DOI archive under Concept DOI versioning.
  • Licensing: Creative Commons Attribution 4.0 International (CC BY 4.0). Derived in part from the Amidha Ayurveda Herb Database.

Citation & Academic Attribution

If you use these datasets in academic research, AI model training, or retrieval-augmented generation systems, please cite:

@dataset{age_ayurveda_nighantu_2026,
  author       = {Age Ayurveda Research Panel and Amidha Ayurveda},
  title        = {Age Ayurveda Nighantu: Open Classical Materia Medica and Dravyaguna Dataset},
  year         = 2026,
  publisher    = {Age Ayurveda / Zenodo / Hugging Face},
  license      = {CC BY 4.0},
  url          = {https://nighantu.ageayurveda.com/datasets/}
}