CoolFace
5 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01FreedomIntelligence /medical-o1-verifiable-problem Introduction This dataset features open-ended medical problems designed to improve LLMs' medical reasoning. Each entry includes a open-ended question and a ground-truth answer based on challenging medical exams. The verifiable answers enable checking LLM outputs, refining their reasoning processes. For details, see our paper and GitHub repository. Citation If you find our data useful, please consider citing our work! @misc{chen2024huatuogpto1medicalcomplexreasoning… See the full description on the dataset page: https://huggingface.co/datasets/FreedomIntelligence/medical-o1-verifiable-problem.textquestion-answering10K<n<100K124 likes778 downloads2y agoHugging Face02devvrit /polaris_filtered_nemotron_medium_math_verifiable Polaris Filtered Nemotron Medium Sympy Verifiable (v2) This dataset is a curated subset of reasoning data from nvidia/Nemotron-Math-v2, specifically filtered for mathematical verifiability (verified using math verify-based equivalence), not having tool-reliance (TIR), and decontamination against the POLAIRS (POLARIS-Project/Polaris-Dataset-53K) dataset. Dataset Summary Total Original Samples: 2,424,392 Final Kept Samples: 357,790 (14.8%) Target Reasoning Length: 4k-8k… See the full description on the dataset page: https://huggingface.co/datasets/devvrit/polaris_filtered_nemotron_medium_math_verifiable.texttext-generation100K<n<1M0 likes53 downloads9mo agoHugging Face03devvrit /polaris_filtered_nemotron_medium_sympy_verifiable Polaris Filtered Nemotron Medium Sympy Verifiable This dataset is a curated subset of reasoning data from nvidia/Nemotron-Math-v2, specifically filtered for mathematical verifiability (verified using sympy-based equivalence), not having tool-reliance (TIR), and decontamination against the POLAIRS (POLARIS-Project/Polaris-Dataset-53K) dataset. Dataset Summary Total Original Samples: 2,500,820 Final Kept Samples: 263,123 (10.5%) Target Reasoning Length: Optimized for… See the full description on the dataset page: https://huggingface.co/datasets/devvrit/polaris_filtered_nemotron_medium_sympy_verifiable.texttext-generation100K<n<1M0 likes45 downloads9mo agoHugging Face04ZombitX64 /Medical-o1-verifiable-problem-Thai Introduction This dataset features open-ended medical problems designed to improve LLMs' medical reasoning. Each entry includes a open-ended question and a ground-truth answer based on challenging medical exams. The verifiable answers enable checking LLM outputs, refining their reasoning processes. For details, see our paper and GitHub repository. Citation If you find our data useful, please consider citing our work! @misc{chen2024huatuogpto1medicalcomplexreasoning… See the full description on the dataset page: https://huggingface.co/datasets/ZombitX64/Medical-o1-verifiable-problem-Thai.textquestion-answering10K<n<100K0 likes29 downloads1y agoHugging Face05ilijalichkovski /medical-o1-verifiable-problem-mk Dataset Card for Dataset Name This is a preview of a Macedonian translation of the medical-o1-verifiable-problem dataset by Freedom Intelligence. Note that this preview currently contains 1068 rows. Dataset Details Dataset Structure Each example consists of a question and a verifiable answer. Dataset Creation For methodological details regarding the creation of the original dataset, please refer to the original paper. Machine translation was… See the full description on the dataset page: https://huggingface.co/datasets/ilijalichkovski/medical-o1-verifiable-problem-mk.textquestion-answering1K<n<10K0 likes15 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.