datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
text-2-image-human-preferences-2m
Text-to-image human preferences: 2M votes across 30 models
This dataset contains the complete voting record behind the
Datapoint Image Bench
leaderboard: 2,161,160 validated pairwise votes — exactly 10 for each of
216,116 image pairs. The votes compare 30 text-to-image models in a complete
round-robin on 500 prompts, judged by annotators from over 200 countries.
Every vote includes the annotator's trust score at the time the vote was
cast.
Built on the Datapoint annotation… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-2-image-human-preferences-2m.SMK-image-text
Dataset Card for “SMK Image-Text (Danish/English)”
Summary
Each row corresponds to an SMK collection object. It contains:
Raw image bytes (image_bytes) plus thumbnails and basic image stats (width/height/size, entropy, contrast, etc.).
Object metadata in Danish and English: titles, object names, artists/creators, production dates, techniques, materials, inscriptions, labels, documentation references.
Rights information: public_domain flag and rights text per… See the full description on the dataset page: https://huggingface.co/datasets/V4ldeLund/SMK-image-text.SMK-image-text-synthimage_query_text_gme7bparallel-image-text-dataset-builder
parallel-image-text-dataset-builder (sample)
A small representative sample from the
parallel-image-text-dataset-builder
pipeline: it ingests image-text pairs, removes near-duplicates with
perceptual-hash (dhash) LSH-style bucketing, filters weak pairs by CLIP
image-text similarity, and writes fixed-size WebDataset-style tar shards.
Contents
shard-00002.tar - one WebDataset-style shard (536 samples). Each sample is
two members sharing a key: {key}.jpg (image) and… See the full description on the dataset page: https://huggingface.co/datasets/narinzar/parallel-image-text-dataset-builder.vg_actions_spatial_for_graphormer_processed_with_text_image_graphscoco_val2017_100_text_image_poseText_to_ImageTrain Demo Datasets
image-to-text-checkpoint-downloadsvg_actions_spatial_for_graphormer_processed_with_text_action_image_graphsmultimodal-image-text-retrieval-artifacts
