datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
muslim-womens-interviews
Muslim Women's Interviews (MWI)
De-identified interviews with 25 Muslim women in the United States, Canada and
France — 585 question/answer exchanges, about 144,000 words of interview text.
Released with participant consent for the express purpose of shaping how large language
models represent Muslim women.
The dataset
Two views of the same 585 question/answer exchanges:
interviews/MWI-001.md … MWI-025.md — one file per interview. Readable on
GitHub, importable… See the full description on the dataset page: https://huggingface.co/datasets/InOurWords/muslim-womens-interviews.women-health-mini
Dataset Overview
This dataset was developed using a robust and structured process, forming the foundation for fine-tuning our language model. The dataset generation began by scraping reputable health-related websites and collecting high-quality, open-source e-books and PDFs that focus on women's health. These diverse sources were curated to create a rich and varied instruction dataset, ensuring that the final dataset covered a wide range of topics relevant to women's health. The… See the full description on the dataset page: https://huggingface.co/datasets/altaidevorg/women-health-mini.SHSP_Snake_TB_Women_NepQA
Nepali ShareGPT Health & Social Statistics Dataset
A Nepali-language, ShareGPT-formatted, single-turn instruction-following dataset of 453 question–answer pairs derived from official Nepali government health and demographic statistical publications. Every question is a fact-based question in Nepali (Devanagari script), and every answer is a short, fact-grounded response extracted directly from the source statistical tables — no hallucinated or generated numbers.
This README… See the full description on the dataset page: https://huggingface.co/datasets/sabin1234/SHSP_Snake_TB_Women_NepQA.womens_health_evidence_qa
Womens Health Evidence QA
Adaption-enhanced training dataset for evidence-based nutrition AI
Produced for the Adaption AutoScientist Challenge 2026 using Adaptive Data by Adaption.
Dataset Summary
2,000 training pairs with enhanced prompts, completions, and reasoning traces produced by Adaptive Data by Adaption. This is the dataset that trained the winning model at 69% win rate against the GPT-OSS-120B baseline.
Content covers evidence-based nutritional and… See the full description on the dataset page: https://huggingface.co/datasets/parissharpe/womens_health_evidence_qa.womens-health-benchmark-ground-truth
Womens Health Benchmark Ground Truth
A curated evaluation dataset for assessing large language models (LLMs) in womens health-related tasks. The dataset consists of model stumps paired with expert-written justifications describing observed errors.
Research Focus
This dataset supports structured evaluation of LLM behavior in clinically relevant womens health contexts, with emphasis on safety, reasoning quality, and evidence alignment.
Methodology
The dataset was… See the full description on the dataset page: https://huggingface.co/datasets/therubricai/womens-health-benchmark-ground-truth.womens-health-fertility-safety-qa
Dataset Card for WomenHealth-Ferility-safety-guideline
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
This dataset contains clinically grounded question–answer pairs focused on fertility treatments and IVF (In Vitro Fertilization), with a strong emphasis on patient safety, procedure awareness, and emotional reassurance.
The dataset is designed to address… See the full description on the dataset page: https://huggingface.co/datasets/Khyatimirani/womens-health-fertility-safety-qa.women_health_10k_insights
Women's Health 10k Insights
Women's Health 10k Insights is a dataset of medical cases selected for their relevance to women's health and processed into concise, practice-oriented tips for clinicians and medical AI agents.
The dataset is intended to help surface uncommon presentations, diagnostic pitfalls, misleading test results, and other lessons that may reduce avoidable clinical mistakes. It is an educational and decision-support resource—not a diagnostic system or a… See the full description on the dataset page: https://huggingface.co/datasets/FremyCompany/women_health_10k_insights.women-health-ferility-natural-try-QAwomen_llama3
