datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dancegrpo-t2av
MiniMax H3 FL2VA First-Frame Dataset
27,815 FLUX-generated reference images paired with text prompts, built as
first-frame (image) conditions for MiniMax H3 FL2VA (text+image to
audio-video) RL training.
Dataset recipe
Prompts (prompts.txt, 27,815 lines): English video captions from
ConsisID-preview-Data,
as filtered and released by DanceGRPO
(assets/consist-id.txt).
Images (images/{index:06d}.jpg): each prompt rendered offline with
FLUX.1-dev on 8 GPUs — 400x640… See the full description on the dataset page: https://huggingface.co/datasets/zyfenghit/dancegrpo-t2av.mle-playbooksmle-skills-kaggle-writeups-clean
