datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
f1-strategy-faithfulness
F1 & Weather Strategy Faithfulness Benchmark
Companion data for the paper "Precision Is Not Faithfulness: Coverage-Aware Evaluation
of Grounded Generation with a Complete Oracle." Each instance ships a structured
complete oracle — the full, enumerable set of checkable facts that a good explanation
should cover — which is what lets the metric measure recall (coverage) alongside
precision (faithfulness), unlike open-domain settings.
Contents
f1/instances.jsonl… See the full description on the dataset page: https://huggingface.co/datasets/jsantillana/f1-strategy-faithfulness.medical-o1-reasoning-SFT-it_f10_incremental
News
[2025/03/08] We open sourced the medical reasoning dataset for SFT translated into italian language.
Dataset Description
This dataset will be used to fine-tuning a distiled model to generate an italian medical LLM designed for advanced medical reasoning.
We used the "facebook/nllb-200-distilled-600M" model to translate the "FreedomIntelligence/medical-o1-reasoning-SFT" dataset, from English to Italian
Citation
We would like to thank the authors of the… See the full description on the dataset page: https://huggingface.co/datasets/eugrug-60/medical-o1-reasoning-SFT-it_f10_incremental.
