datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Text_to_Image
Dataset Card
Dataset in ImagenHub.
Citation
Please kindly cite our paper if you use our code, data, models or results:
@article{ku2023imagenhub,
title={ImagenHub: Standardizing the evaluation of conditional image generation models},
author={Max Ku and Tianle Li and Kai Zhang and Yujie Lu and Xingyu Fu and Wenwen Zhuang and Wenhu Chen},
journal={arXiv preprint arXiv:2310.01596},
year={2023}
}
text-to-image-diffusiondb-2M
DiffusionDB text-to-image subset
A cleaned, safety-filtered image-prompt dataset for training a text-to-image
model, built from DiffusionDB.
Built on Hugging Face Jobs directly from poloclub/diffusiondb. It covers
part_id 1-20 (20,000 source images) before filtering. The same content is
also kept on the 20k-subset branch.
Load it with:
load_dataset("whosouravsharma/text-to-image-diffusiondb-2M")
Note on the repo name: despite "2M" in the name, this is a small slice of… See the full description on the dataset page: https://huggingface.co/datasets/whosouravsharma/text-to-image-diffusiondb-2M.Detonate_Text_To_Imagecode-image-to-text
Code Snippet Image → Text
A multimodal dataset for fine-tuning vision-language models (VLMs) on the task of
transcribing an image of a code snippet back into its source text — syntax-aware OCR.
Each example pairs a syntax-highlighted PNG of code with the exact code text that
produced it. It spans 8 programming languages and deliberately mixes two capture types:
block — a complete function / unit (6–45 lines).
fragment — a contiguous partial view (3–14 lines) that may start or… See the full description on the dataset page: https://huggingface.co/datasets/anisiraj/code-image-to-text.text_to_imagefashion_text_to_image
annotations_creators:
- machine-generated
language:
- en
language_creators:
- other
multilinguality:
- monolingual
pretty_name: "Fashion captions"
size_categories:
- n<100K
tags: []
task_categories:
- text-to-image
task_ids: []
Dataset Card for [Dataset Name]
Dataset Summary
[More Information Needed]
Supported Tasks and Leaderboards
[More Information Needed]
Languages
[More Information Needed]
Dataset Structure… See the full description on the dataset page: https://huggingface.co/datasets/duyngtr16061999/fashion_text_to_image.diagram_image_to_text
Dataset Card for "diagram_image_to_text"
More Information needed
dior_text_to_imageChemistry_text_to_image
Dataset Card for "Chemistry_text_to_image"
More Information needed
TREC-2023-Text-to-Image
Dataset Card for "TREC-2023-Text-to-Image"
More Information needed
diffusion.4.text_to_image
Dataset Card for "diffusion.4.text_to_image"
More Information needed
aid-text-to-imageimage-description_text_to_image_BASE64Chemistry_text_to_image_BASE64visdrone-text-to-imageimage-to-textImagetotextMGLdataset_info:
features:
- name: image
dtype: image
- name: text
dtype: string
splits:
- name: train
mm_diagram_image_to_textdiagram_image_to_text_BASE64carla_image_to_text_datasetTREC-2023-Image-to-Text
Dataset Card for "TREC-2023-Image-to-Text"
More Information needed
winogroud_text_to_image
Dataset Card for "winogroud_text_to_image"
More Information needed
text-to-image-partialrick_and_morty_text_to_image
Dataset Card for "rick_and_morty_text_to_image"
More Information needed
Text_to_ImageText to Image
5508nailset_diffusion.4.text_to_image
Dataset Card for "5508nailset_diffusion.4.text_to_image"
More Information needed
diffusion.4.text_to_image.book
Dataset Card for "diffusion.4.text_to_image.book"
More Information needed
pinterest-multimodal-text-to-imagetext-to-image-checkpoint-downloadsimage-to-text-checkpoint-downloads
