WenhaoWang/VidProM
Summary This is the dataset proposed in our paper VidProM: A Million-scale Real Prompt-Gallery Dataset for Text-to-Video Diffusion Models (NeurIPS 2024). VidProM is the first dataset featuring 1.67 million unique text-to-video prompts and 6.69 million videos generated from 4 different state-of-the-art diffusion models. It inspires many exciting new research areas, such as Text-to-Video Prompt Engineering, Efficient Video Generation, Fake Video Detection, and Video Copy… See the full description on the dataset page: https://huggingface.co/datasets/WenhaoWang/VidProM.
8311k
1version https://git-lfs.github.com/spec/v12oid sha256:7668b87803b67638fa899b69f3a2a298ea13daffe508ce625e21937bff958b6e3size 3827346744 