Perle-ai/T2I-benchmark
T2I Benchmark: Frontier Text-to-Image Models on Image Description Prompts Accompanying dataset for Benchmarking Frontier Text-to-Image Models on the Image Description Prompts (Perle AI). Four frontier text-to-image systems are compared on the 48 hardest prompts in the DataSeeds.AI Sample Dataset (DSD). Every prompt is a verbatim, human-written description of a real photograph — no prompt engineering. Every generated image is graded against a per-prompt weighted rubric by an… See the full description on the dataset page: https://huggingface.co/datasets/Perle-ai/T2I-benchmark.
07
Add T2I benchmark: 48 hardest DSD prompts, 4 models, full rubric grades
initial commit
