CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Vchitect /Vchitect_T2V_DataVerse Vchitect-T2V-Dataverse Vchitect Team1  1Shanghai Artificial Intelligence Laboratory  Paper | Project Page | Data Overview The Vchitect-T2V-Dataverse is the core dataset used to train our text-to-video diffusion model, Vchitect-2.0: Parallel Transformer for Scaling Up Video Diffusion Models. It comprises 14 million high-quality videos collected from the Internet, each paired with detailed textual… See the full description on the dataset page: https://huggingface.co/datasets/Vchitect/Vchitect_T2V_DataVerse.texttext-to-video1M<n<10M11 likes58k downloads1y agoHugging Face02Central-Cat /vbvr-latent-cache-832x832x33f-t2v-only VBVR Latent Cache (832×832 × 33f, Wan2.2-TI2V-5B VAE + UMT5-XXL) Pre-encoded latent cache for the Video-Reason/VBVR-Dataset geometric / logical reasoning video corpus, prepared for Equilibrium Matching (EqM) post-training of Wan-AI/Wan2.2-TI2V-5B-Diffusers on AWS Trainium2. This is a working cache, not a primary dataset. It exists to skip the ~5 s/sample VAE+T5 encode cost during training. The original videos + prompts live in the upstream VBVR-Dataset repo. Source →… See the full description on the dataset page: https://huggingface.co/datasets/Central-Cat/vbvr-latent-cache-832x832x33f-t2v-only.text-to-video0 likes7.3k downloads5mo agoHugging Face03omnimem /Wan2.1-T2V-1.3B_vidprom_81x480x832_40step_5cfg_5.0shift_4t0 likes1.5k downloads5mo agoHugging Face04RaphaelLiu /EvalCrafter_T2V_Dataset EvalCrafter Text-to-Video (ECTV) Dataset 🎥📊 Code · Project Page · Huggingface Leaderboard · Paper@ArXiv · Prompt list Welcome to the ECTV dataset! This repository contains around 10000 videos generated by various methods using the Prompt list. These videos have been evaluated using the innovative EvalCrafter framework, which assesses generative models across visual, content, and motion qualities using 17 objective metrics and subjective user opinions. Dataset Details 📚… See the full description on the dataset page: https://huggingface.co/datasets/RaphaelLiu/EvalCrafter_T2V_Dataset.textn<1K10 likes1.4k downloads3y agoHugging Face05NilanE /Vchitect_T2V_DataVerse_256p_8fps_wdshttps://huggingface.co/datasets/Vchitect/Vchitect_T2V_DataVerse resampled to 256p. Intended for training https://github.com/NilanEkanayake/TiTok-Video text100K<n<1M0 likes963 downloads1y agoHugging Face06samuelt0207 /Wan2.2-T2V-Activations-FP40 likes784 downloads10mo agoHugging Face07FastVideo /wan_t2v_distillation_datasettext1K<n<10K2 likes566 downloads1y agoHugging Face08ViBe-T2V-Bench /ViBeThis repository contains the data presented in ViBe: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models. ViBe Annotations and Category Labels Order of Model Annotations The annotations for the models in metadata.csv are organized in the following order: animatelcm zeroscopeV2_XL show1 mora animatelightning animatemotionadapter magictime zeroscopeV2_576w ms1.7b hotshotxl videotext-to-videon<1K0 likes310 downloads2y agoHugging Face09stepfun-ai /Step-Video-T2V-EvalThis dataset contains the data of the paper Step-Video-T2V Technical Report: The Practice, Challenges, and Future of Video Foundation Model. Code: https://github.com/stepfun-ai/Step-Video-T2V Project page: https://yuewen.cn/videos videotext-to-videon<1K1 likes296 downloads2y agoHugging Face10uEval /t2v-sampled-videos0 likes244 downloads1y agoHugging Face11wlsaidhi /crush-smol_processed_t2vtabularn<1K0 likes210 downloads1y agoHugging Face12hffordata /t2v-distill-benchmark T2V Distillation Benchmark A benchmark comparison of Text-to-Video distillation/acceleration models based on Wan2.1-14B. Models Compared Model Steps Source FastVideo CausalWan2.2 8 FastVideo/CausalWan2.2-I2V-A14B-Preview-Diffusers Krea Realtime-Video 4 krea/krea-realtime-video LightX2V CausVid 9 lightx2v/Wan2.1-T2V-14B-CausVid NVlabs rCM 14B 4 worstcoder/rcm-Wan Helios-Distilled 3 (pyramid 2-2-2) BestWishYsh/Helios-Distilled… See the full description on the dataset page: https://huggingface.co/datasets/hffordata/t2v-distill-benchmark.videotext-to-videon<1K0 likes202 downloads1d agoHugging Face13mteb /FLARE-1k-Unified-T2VAaudio1K<n<10K0 likes183 downloads2mo agoHugging Face14shriyasudhakar /ChinaOpen1k-T2V ChinaOpen-1k T2V retrieval Multilingual video retrieval built from ChinaOpen-1k: 1,092 Bilibili videos with human-written Chinese captions and their English translations. Both language configs share the same videos, so the two subsets form a controlled comparison. Built by scripts/data/chinaopen_retrieval/create_data.py in mteb. Please cite the original dataset: @inproceedings{chen2023chinaopen, title = {ChinaOpen: A Dataset for Open-world Multimodal Learning}, author =… See the full description on the dataset page: https://huggingface.co/datasets/shriyasudhakar/ChinaOpen1k-T2V.textvideo-text-to-text1K<n<10K0 likes178 downloads16d agoHugging Face15mteb /FLARE-1k-Audio-T2VAaudio1K<n<10K0 likes161 downloads2mo agoHugging Face16Yi30 /vbench-wan22-t2v-dense93-xpu VBench dense93 — Wan2.2 T2V videos (Intel XPU / vLLM-Omni) 93 text-to-video generations for the VBench "dense93" prompt set, produced with Wan2.2-T2V-A14B (Diffusers) served by vLLM-Omni on Intel XPU (oneAPI, 720x1280 @ 16fps, 81 frames, 40 denoise steps, guidance 4.0/3.0, boundary 0.875, flow shift 5.0, seed 42, MXFP8 linear + cache-dit + Sage V3 hybrid attention with SDPA fallback on blocks 33,34,38,39). One video per prompt; file name = <prompt>-0.mp4 (prompt list:… See the full description on the dataset page: https://huggingface.co/datasets/Yi30/vbench-wan22-t2v-dense93-xpu.videon<1K0 likes159 downloads9d agoHugging Face17samuelt0207 /Wan2.2-T2V-Activations-INT40 likes136 downloads10mo agoHugging Face18Kaiyue /T2V-CompBench-Videosimagevideo-classification10K<n<100K1 likes131 downloads11mo agoHugging Face19StefanFalkok /Wan_2.2_T2V_10steps_GGUF0 likes130 downloads11mo agoHugging Face20CinematicT2vData /cinepile-t2v-split_scenes_single_shot_uniform Important Columns for Captioning Caption_t2v_style: Expressive and long caption generated by Gemini Flash 2.5 for the extracted shot. Caption_t2v_style_short: Short caption generated by Gemini Flash 2.5 for the extracted shot. Avg-Aesthetic-Score-Laion-Aesthetics: Average (over frames) aesthetic score of the extracted shot from Laion Aesthetics. Frame-Aesthetic-Scores-Laion-Aesthetics: Aesthetic scores of each frame of the extracted shot from Laion Aesthetics.… See the full description on the dataset page: https://huggingface.co/datasets/CinematicT2vData/cinepile-t2v-split_scenes_single_shot_uniform.tabular10K<n<100K0 likes94 downloads1y agoHugging Face21StefanFalkok /Wan_2.2_T2V_10steps0 likes91 downloads11mo agoHugging Face22qgfvadfuvads /t2v_data_v2 DenseDPO T2V Broad-Pair Dataset (v2) Cross-model text-to-video (T2V) generation pairs for training video reward models (RM) and DPO-style preference learning. The HF Dataset Viewer renders each row as prompt + two videos side-by-side. Generation task All videos are generated T2V from a shared text prompt. For every pair, both videos share the same prompt, so the primary comparison axis is the model identity itself. Plan-A tier structure Models are grouped into… See the full description on the dataset page: https://huggingface.co/datasets/qgfvadfuvads/t2v_data_v2.texttext-to-video1K<n<10K0 likes88 downloads4mo agoHugging Face23Wissam42 /FLARE-1k-Audio-T2VAaudio1K<n<10K0 likes87 downloads2mo agoHugging Face24George870 /Wan21_CausVid_14B_T2V_lora_rank32.safetensors5 likes86 downloads1y agoHugging Face25qgfvadfuvads /azm-archive-20260909-t2v-needlabel-videos t2v_data_needlabel_videos.tar Backup of an existing dataset archive, preserving its original bytes. File: t2v_data_needlabel_videos.tar Size: 4,675,983,360 bytes SHA256: 8672488cef23b55e2a70d3279327a38ddbbd8e2809661600d7c707d59b9e1e18 Verify the downloaded archive with sha256sum -c SHA256SUMS. textn<1K0 likes85 downloads12d agoHugging Face26Wissam42 /FLARE-1k-Unified-T2VAaudio1K<n<10K0 likes79 downloads2mo agoHugging Face27mmfm-trust /T2Vtext1K<n<10K1 likes74 downloads1y agoHugging Face28mteb /FLARE-1k-Vision-T2Vtext1K<n<10K0 likes72 downloads2mo agoHugging Face29minzh23 /mixkit-t2vtabular1K<n<10K0 likes66 downloads3mo agoHugging Face30ApacheOne /Info_Wan_Video_2.2_T2V-A14B Model Index by Creator 423748 Page Model Base Model Full Model Page Archive Link wan2.2,t2v,low,zzzyixuan. Wan Video 2.2 T2V-A14B View View Version Links Model Version Base Model Version Link wan2.2,t2v,low,zzzyixuan. v1.0 Wan Video 2.2 T2V-A14B View Aaron_PP Page Model Base Model Full Model Page Archive Link NSFW WAN 2.2 T2V Bunny girl, red patent leather tights, black high stockings, red high heels Wan Video 2.2… See the full description on the dataset page: https://huggingface.co/datasets/ApacheOne/Info_Wan_Video_2.2_T2V-A14B.textn<1K7 likes59 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.