CoolFace
20 results

t2i

ma-xu /fine-t2i Fine-T2I: An Open, Large-Scale, and Diverse Dataset for High-Quality T2I Fine-Tuning [arxiv] by Xu Ma, Yitian Zhang, Qihua Dong, Yun Fu Northeastern Univeristy Please see our [Dataset Explore] to view detailed samples (loading is slow, be patient). 🆕 What's New [2026.02.20]: Fine-T2I reaches the #1 spot among Hugging Face Datasets Trending list ⭐️⭐️⭐️ [2026.02.16]: Fine-T2I tops the Hugging Face Datasets Trending list, reaching the #2 spot and #1… See the full description on the dataset page: https://huggingface.co/datasets/ma-xu/fine-t2i.imageimage-to-text100K<n<1M120 likes21k downloads7mo agoHugging Facelioooox /T2I-CoReBench-Images T2I-CoReBench-Images 📖 Overview T2I-CoReBench-Images is the companion image dataset of T2I-CoReBench. It contains images generated using 1,080 challenging prompts, covering both composition and reasoning scenarios undere real-world complexities. This dataset is designed to evaluate how well current Text-to-Image (T2I) models can not only paint (produce visually consistent outputs) but also think (perform reasoning over causal chains, object relations, and logical… See the full description on the dataset page: https://huggingface.co/datasets/lioooox/T2I-CoReBench-Images.imagetext-to-image10K<n<100K5 likes8.2k downloads7mo agoHugging Facesaxon /T2IScoreScore Dataset Card for Text-to-Image ScoreScore (T2IScoreScore or TS2) This dataset exists as part of the T2IScoreScore metaevaluation for assessing the faithfulness and consistency of text-to-image model prompt-image evaluation metrics. Necessary code for utilizing the resource is present at github.com/michaelsaxon/T2IScoreScore Dataset Details Dataset Description This is a test set of 165 "target prompts" which each have between 5 and 76 generated images of… See the full description on the dataset page: https://huggingface.co/datasets/saxon/T2IScoreScore.imagetext-to-image1K<n<10K8 likes4k downloads2y agoHugging Faceyufan /GPT4O_Image_T2Iimage10K<n<100K2 likes2.3k downloads1y agoHugging FaceSPRIGHT-T2I /spright Dataset Description SPRIGHT (SPatially RIGHT) is the first spatially focused, large scale vision-language dataset. It was built by re-captioning ∼6 million images from 4 widely-used datasets: CC12M Segment Anything COCO Validation LAION Aesthetics This repository contains the re-captioned data from CC12M and Segment Anything, while the COCO data is present here. We do not release images from LAION, as the parent images are currently private. Below are some illustrative examples… See the full description on the dataset page: https://huggingface.co/datasets/SPRIGHT-T2I/spright.1M<n<10M33 likes2k downloads2y agoHugging Facerevision-t2i /revision-generator REVISION: Rendering Tools Enable Spatial Fidelity in Vision-Language Models (ECCV 2024) This is the official dataset of the REVISION framework with all the corresponding assets (i.e. objects, backgrounds, and floors). ⚒️ Requirements REVISION requires blenderproc. Simply install it with pip: pip install blenderproc 👁️ Single Test Run To generate a two-object reference image deterministically on your own, you may invoke one of the 4 blenderproc scripts… See the full description on the dataset page: https://huggingface.co/datasets/revision-t2i/revision-generator.text-to-image4 likes1.8k downloads2y agoHugging Face