datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
GenRef-wds
GenRef-1M
We provide 1M high-quality triplets of the form (flawed image, high-quality image, reflection) collected across
multiple domains using our scalable pipeline from [1]. We used this dataset to train our reflection tuning model.
To know the details of the dataset creation pipeline, please refer to Section 3.2 of [1].
Project Page: https://diffusion-cot.github.io/reflection2perfection
Dataset loading
We provide the dataset in the webdataset format for fast… See the full description on the dataset page: https://huggingface.co/datasets/diffusion-cot/GenRef-wds.GenRef-CoT
GenRef-CoT
We provide 227K high-quality CoT reflections which were used to train our Qwen-based reflection generation model in ReflectionFlow [1]. To
know the details of the dataset creation pipeline, please refer to Section 3.2 of [1].
Dataset loading
We provide the dataset in the webdataset format for fast dataloading and streaming. We recommend downloading
the repository locally for faster I/O:
from huggingface_hub import snapshot_download
local_dir =… See the full description on the dataset page: https://huggingface.co/datasets/diffusion-cot/GenRef-CoT.VideoVista-CoTs
VideoVista-CoTs
This repository contains VideoVista-CoTs, used in Uni-MoE-2.0 training.
This dataset samples a portion of data from LLaVA-Video-178K, SEED-Bench-R1, SR-91K, and STAR, and uses our automatic Video QA generation framework to perform multi-step reasoning annotations for filtered complex questions.
The automatic video QA generation codes and our VideoVista series are presented in VideoVista Family
Citation
If you find VideoVista-CulturalLingo useful for your… See the full description on the dataset page: https://huggingface.co/datasets/HIT-TMG/VideoVista-CoTs.cot-migration
Model Card for Model ID
Model Details
Model Description
Developed by: [More Information Needed]
Funded by [optional]: [More Information Needed]
Shared by [optional]: [More Information Needed]
Model type: [More Information Needed]
Language(s) (NLP): [More Information Needed]
License: [More Information Needed]
Finetuned from model [optional]: [More Information Needed]
Model Sources [optional]
Repository: [More Information Needed]
Paper… See the full description on the dataset page: https://huggingface.co/datasets/Danqi7/cot-migration.pusht_96_norm4_visual_nomarker_stopreq_candidate_shuffle_cot_aligned100k
PushT 96 Norm4 Visual Nomarker Stop-Required Candidate-Shuffle CoT Aligned 100k
This repository contains the exact PushT CoT dataset used for the 2026-06-04 ordered SFT CoT run.
Archive:
pusht_96_norm4_visual_nomarker_stopreq_candidate_shuffle_cot_aligned100k_20260604_ordered.tar.gz
Contents after extraction:
data/train/: 100000 CoT JSONL records in 8 gzip shards.
data/test/: 100 CoT JSONL records in 1 gzip file.
metadata/train_alignment_manifest.jsonl… See the full description on the dataset page: https://huggingface.co/datasets/novastar113/pusht_96_norm4_visual_nomarker_stopreq_candidate_shuffle_cot_aligned100k.map_train_data_cot_300_v7_HoldObject_images
