Open Ayurvedic Datasets
Open, machine-readable datasets of classical Ayurvedic pharmacology, phytochemical compounds, and botanical taxonomy. Published under Creative Commons Attribution 4.0 International (CC BY 4.0).
All datasets published by the Age Ayurveda Nighantu are generated directly from verified monograph records and classical texts, preventing drift between the reference text and tabular data. Every botanical taxon is resolved against global taxonomic authorities (GBIF, NCBI, Wikidata), and every pharmacological property conforms to standard Ayurvedic Pharmacopoeia of India (API) criteria.
Available Datasets
Ayurvedic Dravyaguna Pharmacological Dataset
Rasa, guna, virya, vipaka and prabhava for 245 monographs, parsed from the pages on this site so every value can be checked against the page it came from.
Coverage: 245 monographs
Phytochemical Constituents & Co-occurrence Network
838 constituents named in the published monographs, 170 of them resolved to PubChem, with 21,942 co-occurrence edges computed from the pages themselves.
Coverage: 838 constituents · 21,942 co-occurrences
Biomedical Literature & Classical Research Index
2,892 papers cited on this site, each with its PubMed ID or DOI, cross-referenced to the monographs that cite it.
Coverage: 2,892 papers
Verified Botanical Taxonomy & Identifier Crosswalk
Botanical names resolved against GBIF, with NCBI taxonomy IDs and Wikidata QIDs where an exact match exists. 285 taxa, each carrying the match type so a fuzzy hit is never read as a fact.
Coverage: 285 resolved taxa
Nighantu Editorial Verification Ledger
What was checked, how, and what was thrown out: every verification run with its examined, published and rejected counts, computed from the run artifacts rather than stated.
Coverage: 0 published safety records, and the rejections behind them
Hugging Face & Zenodo Access
The consolidated corpus is also hosted in Apache Parquet format on the Hugging Face Hub and archived with a persistent Concept DOI on Zenodo.
- Hugging Face Dataset:
ageayurveda/nighantu(supports direct streaming withdatasets.load_dataset("ageayurveda/nighantu")) - Zenodo Archive: Persistent research DOI archive under Concept DOI versioning.
- Licensing: Creative Commons Attribution 4.0 International (CC BY 4.0). Derived in part from the Amidha Ayurveda Herb Database.
Citation & Academic Attribution
If you use these datasets in academic research, AI model training, or retrieval-augmented generation systems, please cite:
@dataset{age_ayurveda_nighantu_2026,
author = {Age Ayurveda Research Panel and Amidha Ayurveda},
title = {Age Ayurveda Nighantu: Open Classical Materia Medica and Dravyaguna Dataset},
year = 2026,
publisher = {Age Ayurveda / Zenodo / Hugging Face},
license = {CC BY 4.0},
url = {https://nighantu.ageayurveda.com/datasets/}
}