i2v
Datasets
All datasets matching “i2v”TIP-I2V
Summary
This is the dataset proposed in our paper TIP-I2V: A Million-Scale Real Text and Image Prompt Dataset for Image-to-Video Generation.
Project page | Paper
TIP-I2V is the first dataset comprising over 1.70 million unique user-provided text and image prompts. Besides the prompts, TIP-I2V also includes videos generated by five state-of-the-art image-to-video models (Pika, Stable Video Diffusion, Open-Sora, I2VGen-XL, and CogVideoX-5B). The TIP-I2V contributes to the development… See the full description on the dataset page: https://huggingface.co/datasets/WenhaoWang/TIP-I2V.TIP-I2V
News
🌟 Downloaded 10,000+ times on Hugging Face after one month of release.
✨ Ranked Top 1 in the Hugging Face Dataset Trending List for the visual generation community (image-to-video, text-to-video, text-to-image, and image-to-image) on November 10, 2024.
Summary
This is the dataset proposed in our paper TIP-I2V: A Million-Scale Real Text and Image Prompt Dataset for Image-to-Video Generation.
TIP-I2V is the first dataset comprising over 1.70 million unique… See the full description on the dataset page: https://huggingface.co/datasets/tipi2v/TIP-I2V.I2VEdit-pretrained-videosVBench-I2V_sampled_videoWan2.2-I2V-Activations-INT4I2V-CompBench
I2V-CompBench
A compositional image-to-video (I2V) generation benchmark spanning 7 evaluation dimensions, with first-frame images derived from TIP-I2V and refined text prompts produced by a dual VLM/LLM pipeline.
⚠️ License: CC BY-NC 4.0 (inherits from TIP-I2V). Non-commercial use only.
📦 Versions
This repository hosts two parallel snapshots of the same benchmark. Pick the layout that fits your tooling.
Version
Path
Questions
Layout
Best for
v2 ⭐… See the full description on the dataset page: https://huggingface.co/datasets/YiningZ2002/I2V-CompBench.
