datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mm-long-storytelling-bench
MM Long Storytelling Bench — v3
⚠️ The 756 model-drafted questions have been WITHDRAWN from this
dataset's splits (2026-08-05). They were drafted by a model that is also
an evaluation target, which makes them circular as a measurement
instrument. They are kept in full, with the reasoning, under
data/v3/archive/ — nothing was deleted.
The splits currently hold 6 worked examples (status: "example"),
which document the required format and are not a benchmark. Do not use
this… See the full description on the dataset page: https://huggingface.co/datasets/luoojason/mm-long-storytelling-bench.SemiEvol
Dataset Card for Dataset Name
The SemiEvol dataset is part of the broader work on semi-supervised fine-tuning for Large Language Models (LLMs). The dataset includes labeled and unlabeled data splits designed to enhance the reasoning capabilities of LLMs through a bi-level knowledge propagation and selection framework, as proposed in the paper SemiEvol: Semi-supervised Fine-tuning for LLM Adaptation.
Dataset Details
Dataset Sources [optional]… See the full description on the dataset page: https://huggingface.co/datasets/luojunyu/SemiEvol.
