datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
needle-1M-bench-mvp
needle-1M-bench
A long-context faithfulness benchmark on real arxiv papers.
The standard 1M-context needle-retrieval bench for dense scientific text — the kind production users actually feed in. Maintained by Muhammad Awais (@drawais_ai).
This dataset ships test items + leaderboard scores. The benchmark is centrally scored to ensure consistent methodology and prevent self-reporting drift — see How to get your model on the leaderboard below. Source paper text is NOT redistributed… See the full description on the dataset page: https://huggingface.co/datasets/drawais/needle-1M-bench-mvp.india-medical-value-travel-mvp
India Medical Value Travel (MVT) Platform – MVP Dataset
A comprehensive, structured JSON dataset for building an AI-powered Medical Value Travel platform connecting international patients with Indian hospitals.
Overview
India is a global leader in medical tourism due to 60–80% lower treatment costs vs US/UK, world-class hospital chains, and government support through initiatives like "Heal in India" and e-Medical Visa. This dataset provides the complete data foundation… See the full description on the dataset page: https://huggingface.co/datasets/Dhanush008/india-medical-value-travel-mvp.
