datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
surya-ocr-500-image-to-textsurya-ocr-1K-image-to-texttext-to-image-diffusiondb-2M
DiffusionDB text-to-image subset
A cleaned, safety-filtered image-prompt dataset for training a text-to-image
model, built from DiffusionDB.
Built on Hugging Face Jobs directly from poloclub/diffusiondb. It covers
part_id 1-20 (20,000 source images) before filtering. The same content is
also kept on the 20k-subset branch.
Load it with:
load_dataset("whosouravsharma/text-to-image-diffusiondb-2M")
Note on the repo name: despite "2M" in the name, this is a small slice of… See the full description on the dataset page: https://huggingface.co/datasets/whosouravsharma/text-to-image-diffusiondb-2M.Detonate_Text_To_Imagecode-image-to-text
Code Snippet Image → Text
A multimodal dataset for fine-tuning vision-language models (VLMs) on the task of
transcribing an image of a code snippet back into its source text — syntax-aware OCR.
Each example pairs a syntax-highlighted PNG of code with the exact code text that
produced it. It spans 8 programming languages and deliberately mixes two capture types:
block — a complete function / unit (6–45 lines).
fragment — a contiguous partial view (3–14 lines) that may start or… See the full description on the dataset page: https://huggingface.co/datasets/anisiraj/code-image-to-text.text_to_imageghibli-TextToImagediagram_image_to_text
Dataset Card for "diagram_image_to_text"
More Information needed
trending-text-to-image
CivitAI Improved Prompts Dataset
This dataset contains trending AI-generated images from CivitAI with Flux-improved prompts for better generation results.
Dataset Format (JSONL)
Each line contains a JSON object with:
id: Original image ID from CivitAI
improved_prompt: Flux-enhanced version of the prompt
category: Automatically determined theme category
All original CivitAI metadata including:
Original prompt and negative prompt
Model information
Image URL and… See the full description on the dataset page: https://huggingface.co/datasets/k-mktr/trending-text-to-image.dior_text_to_imagetext-to-image-2M
text-to-image-2M: A High-Quality, Diverse Text-to-Image Training Dataset
Overview
text-to-image-2M is a curated text-image pair dataset designed for fine-tuning text-to-image models. The dataset consists of approximately 2 million samples, carefully selected and enhanced to meet the high demands of text-to-image model training. The motivation behind creating this dataset stems from the observation that datasets with over 1 million samples tend to produce better… See the full description on the dataset page: https://huggingface.co/datasets/rivisia/text-to-image-2M.Chemistry_text_to_image
Dataset Card for "Chemistry_text_to_image"
More Information needed
diffusion.4.text_to_image
Dataset Card for "diffusion.4.text_to_image"
More Information needed
aid-text-to-imageimage-description_text_to_image_BASE64Chemistry_text_to_image_BASE64visdrone-text-to-image16xModdedMinecraft-TextToImage
Minecraft 16x Text-to-Image Dataset (Captioned)
Description
This dataset contains over 1 million Minecraft textures in 16x16 resolution. It has been specifically processed and captioned for training generative AI models (Text-to-Image).
Each image is paired with a descriptive natural language caption derived from the original file labels, enabling AI models to learn the relationship between Minecraft concepts (blocks, items, tools) and their pixel-art representation.… See the full description on the dataset page: https://huggingface.co/datasets/NathMen12/16xModdedMinecraft-TextToImage.mm_diagram_image_to_textdiagram_image_to_text_BASE64carla_image_to_text_datasetTREC-2023-Image-to-Text
Dataset Card for "TREC-2023-Image-to-Text"
More Information needed
winogroud_text_to_image
Dataset Card for "winogroud_text_to_image"
More Information needed
audio-text-embed-to-imagesrick_and_morty_text_to_image
Dataset Card for "rick_and_morty_text_to_image"
More Information needed
5508nailset_diffusion.4.text_to_image
Dataset Card for "5508nailset_diffusion.4.text_to_image"
More Information needed
diffusion.4.text_to_image.book
Dataset Card for "diffusion.4.text_to_image.book"
More Information needed
ocr-image-to-textpinterest-multimodal-text-to-imageText_to_image_dataset
