datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
text-to-image-diffusiondb-2M
DiffusionDB text-to-image subset
A cleaned, safety-filtered image-prompt dataset for training a text-to-image
model, built from DiffusionDB.
Built on Hugging Face Jobs directly from poloclub/diffusiondb. It covers
part_id 1-20 (20,000 source images) before filtering. The same content is
also kept on the 20k-subset branch.
Load it with:
load_dataset("whosouravsharma/text-to-image-diffusiondb-2M")
Note on the repo name: despite "2M" in the name, this is a small slice of… See the full description on the dataset page: https://huggingface.co/datasets/whosouravsharma/text-to-image-diffusiondb-2M.Detonate_Text_To_Imagecode-image-to-text
Code Snippet Image → Text
A multimodal dataset for fine-tuning vision-language models (VLMs) on the task of
transcribing an image of a code snippet back into its source text — syntax-aware OCR.
Each example pairs a syntax-highlighted PNG of code with the exact code text that
produced it. It spans 8 programming languages and deliberately mixes two capture types:
block — a complete function / unit (6–45 lines).
fragment — a contiguous partial view (3–14 lines) that may start or… See the full description on the dataset page: https://huggingface.co/datasets/anisiraj/code-image-to-text.text_to_imagediagram_image_to_text
Dataset Card for "diagram_image_to_text"
More Information needed
dior_text_to_imageChemistry_text_to_image
Dataset Card for "Chemistry_text_to_image"
More Information needed
diffusion.4.text_to_image
Dataset Card for "diffusion.4.text_to_image"
More Information needed
aid-text-to-imageimage-description_text_to_image_BASE64Chemistry_text_to_image_BASE64visdrone-text-to-imageImagetotextMGLdataset_info:
features:
- name: image
dtype: image
- name: text
dtype: string
splits:
- name: train
mm_diagram_image_to_textdiagram_image_to_text_BASE64carla_image_to_text_datasetTREC-2023-Image-to-Text
Dataset Card for "TREC-2023-Image-to-Text"
More Information needed
winogroud_text_to_image
Dataset Card for "winogroud_text_to_image"
More Information needed
rick_and_morty_text_to_image
Dataset Card for "rick_and_morty_text_to_image"
More Information needed
5508nailset_diffusion.4.text_to_image
Dataset Card for "5508nailset_diffusion.4.text_to_image"
More Information needed
diffusion.4.text_to_image.book
Dataset Card for "diffusion.4.text_to_image.book"
More Information needed
pinterest-multimodal-text-to-imagetext-to-image-10kolam-text-to-imageTest-Image-To-Text
Dataset Card for Test-Image-To-Text
This dataset has been created with Argilla. As shown in the sections below, this dataset can be loaded into your Argilla server as explained in Load with Argilla, or used directly with the datasets library in Load with datasets.
Using this dataset with Argilla
To load with Argilla, you'll just need to install Argilla as pip install argilla --upgrade and then use the following code:
import argilla as rg
ds =… See the full description on the dataset page: https://huggingface.co/datasets/HafssaRabah/Test-Image-To-Text.text-to-image-2M_HUkolam-text-to-image-2Image-To-Text-Validation-Datasetminecraft-text-to-imagetext-to-image-1k
