t2i
Datasets
All datasets matching “t2i”fine-t2i
Fine-T2I: An Open, Large-Scale, and Diverse Dataset for High-Quality T2I Fine-Tuning [arxiv]
by Xu Ma, Yitian Zhang,
Qihua Dong, Yun Fu
Northeastern Univeristy
Please see our [Dataset Explore] to view detailed samples (loading is slow, be patient).
🆕 What's New
[2026.02.20]: Fine-T2I reaches the #1 spot among Hugging Face Datasets Trending list ⭐️⭐️⭐️
[2026.02.16]: Fine-T2I tops the Hugging Face Datasets Trending list, reaching the #2 spot and #1… See the full description on the dataset page: https://huggingface.co/datasets/ma-xu/fine-t2i.T2I-CoReBench-Images
T2I-CoReBench-Images
📖 Overview
T2I-CoReBench-Images is the companion image dataset of T2I-CoReBench. It contains images generated using 1,080 challenging prompts, covering both composition and reasoning scenarios undere real-world complexities.
This dataset is designed to evaluate how well current Text-to-Image (T2I) models can not only paint (produce visually consistent outputs) but also think (perform reasoning over causal chains, object relations, and logical… See the full description on the dataset page: https://huggingface.co/datasets/lioooox/T2I-CoReBench-Images.T2IScoreScore
Dataset Card for Text-to-Image ScoreScore (T2IScoreScore or TS2)
This dataset exists as part of the T2IScoreScore metaevaluation for assessing the faithfulness and consistency of text-to-image model prompt-image evaluation metrics.
Necessary code for utilizing the resource is present at github.com/michaelsaxon/T2IScoreScore
Dataset Details
Dataset Description
This is a test set of 165 "target prompts" which each have between 5 and 76 generated images of… See the full description on the dataset page: https://huggingface.co/datasets/saxon/T2IScoreScore.GPT4O_Image_T2Ispright
Dataset Description
SPRIGHT (SPatially RIGHT) is the first spatially focused, large scale vision-language dataset. It was built by re-captioning
∼6 million images from 4 widely-used datasets:
CC12M
Segment Anything
COCO Validation
LAION Aesthetics
This repository contains the re-captioned data from CC12M and Segment Anything, while the COCO data is present here. We do not release images from LAION, as the parent images are currently private.
Below are some illustrative examples… See the full description on the dataset page: https://huggingface.co/datasets/SPRIGHT-T2I/spright.revision-generator
REVISION: Rendering Tools Enable Spatial Fidelity in Vision-Language Models (ECCV 2024)
This is the official dataset of the REVISION framework with all the corresponding assets (i.e. objects, backgrounds, and floors).
⚒️ Requirements
REVISION requires blenderproc. Simply install it with pip:
pip install blenderproc
👁️ Single Test Run
To generate a two-object reference image deterministically on your own, you may invoke one of the 4 blenderproc scripts… See the full description on the dataset page: https://huggingface.co/datasets/revision-t2i/revision-generator.
