wang-sj16/TI2V-Bench
This is the official repository of TI2V-Bench dataset proposed in the paper MotiF: Making Text Count in Image Animation with Motion Focal Loss. TI2V-Bench comprises text-image pairs from 22 diverse scenarios featuring a variety of objects and scenes. Each scenario includes 3 to 5 images with similar content presented in different styles, alongside multiple distinct prompts designed to animate these images and generate varied outputs. The benchmark includes a total of 320 image-text pairs… See the full description on the dataset page: https://huggingface.co/datasets/wang-sj16/TI2V-Bench.
This is the official repository of TI2V-Bench dataset proposed in the paper MotiF: Making Text Count in Image Animation with Motion Focal Loss.
TI2V-Bench comprises text-image pairs from 22 diverse scenarios featuring a variety of objects and scenes. Each scenario includes 3 to 5 images with similar content presented in different styles, alongside multiple distinct prompts designed to animate these images and generate varied outputs. The benchmark includes a total of 320 image-text pairs, consisting of 88 unique images and 133 unique prompts.
Check `prompt.csv` for the image path and corresponding prompts.
You can also download the dataset from Google Drive through this link.
Check more details about this work in our project page.
