datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
medical-ehr-training-data
Medical EHR Training Dataset
Training dataset for Medical EHR GEPA-optimized module.
Dataset Description
This dataset contains 382 medical EHR query examples for training DSPy GEPA optimization.
Dataset Structure
{
"query": "Show me diabetic patients",
"expected_strategy": "ENRICHMENT",
"expected_snomed_codes": ["73211009", "44054006"],
"expected_neo4j_count": 15,
"query_complexity": "simple",
"medical_category": "endocrine"
}
Splits… See the full description on the dataset page: https://huggingface.co/datasets/Fanoni/medical-ehr-training-data.ehr-cardiology-cohort-100
HipAAsynth Dataset
Summary
This dataset is a validation artifact generated by HipAAsynth.
HipAAsynth is a deterministic testing and validation service that simulates real-world variability to evaluate how healthcare systems perform under deployment conditions.
Description
This dataset represents a controlled cohort used for testing and benchmarking.
HipAAsynth generates cohorts to simulate how conditions present across:
patient populations
demographic… See the full description on the dataset page: https://huggingface.co/datasets/HipAAsynth/ehr-cardiology-cohort-100.ehr-longitudinal-cohort-100
HipAAsynth Dataset
Summary
This dataset is a validation artifact generated by HipAAsynth.
HipAAsynth is a deterministic testing and validation service that simulates real-world variability to evaluate how healthcare systems perform under deployment conditions.
Description
This dataset represents a controlled cohort used for testing and benchmarking.
HipAAsynth generates cohorts to simulate how conditions present across:
patient populations
demographic… See the full description on the dataset page: https://huggingface.co/datasets/HipAAsynth/ehr-longitudinal-cohort-100.ehr-rare-disease-cohort-100
HipAAsynth Dataset
Summary
This dataset is a validation artifact generated by HipAAsynth.
HipAAsynth is a deterministic testing and validation service that simulates real-world variability to evaluate how healthcare systems perform under deployment conditions.
Description
This dataset represents a controlled cohort used for testing and benchmarking.
HipAAsynth generates cohorts to simulate how conditions present across:
patient populations
demographic… See the full description on the dataset page: https://huggingface.co/datasets/HipAAsynth/ehr-rare-disease-cohort-100.thinking-data-100-synth
Dataset Card for thinking-data-100
This dataset has been created with distilabel.
Dataset Summary
This dataset contains a pipeline.yaml which can be used to reproduce the pipeline that generated it in distilabel using the distilabel CLI:
distilabel pipeline run --config "https://huggingface.co/datasets/ehristoforu/thinking-data-100/raw/main/pipeline.yaml"
or explore the configuration:
distilabel pipeline info --config… See the full description on the dataset page: https://huggingface.co/datasets/ehristoforu/thinking-data-100-synth.ehr-diabetes-cohort-100
HipAAsynth Dataset
Summary
This dataset is a validation artifact generated by HipAAsynth.
HipAAsynth is a deterministic testing and validation service that simulates real-world variability to evaluate how healthcare systems perform under deployment conditions.
Description
This dataset represents a controlled cohort used for testing and benchmarking.
HipAAsynth generates cohorts to simulate how conditions present across:
patient populations
demographic… See the full description on the dataset page: https://huggingface.co/datasets/HipAAsynth/ehr-diabetes-cohort-100.Med-ART_Clinical_Agent_EHR_Dataset
ART — Action-based Reasoning Tasks (Subset)
120-task stratified sample from the ART benchmark introduced in:
ART: Action-based Reasoning Task Benchmarking for Medical AI Agents
Ananya Mantravadi, Shivali Dalmia, Abhishek Mukherji
arXiv:2601.08988
ART is a programmatically generated clinical decision benchmark built on real FHIR patient data. It targets three dominant error categories in medical AI reasoning — retrieval failures, aggregation errors, and conditional logic… See the full description on the dataset page: https://huggingface.co/datasets/CentificAIResearch/Med-ART_Clinical_Agent_EHR_Dataset.dialogues
Dialogues-Data
In this dataset you will find conversations between users with an AI assistant who have agreed to share their conversations.
What can you use the dataset for?
You can use this dataset as you wish without breaking the rules outlined in the last section!
Examples use:
Train a neural network using OPEN SOURCE CODE
Use for your FREE dataset
For studying dialogues independently or in educational institutions
To compile statistics
Rules
If you use… See the full description on the dataset page: https://huggingface.co/datasets/ehristoforu/dialogues.
