datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mscoco_train_2014_openai_clip-vit-base-patch32_image_caption_retrieval_pairs_2022-09-01ImageCaptions-7M-Translations-Arabicvg-captions-graphs-processed-image-graphsImageCaptions-7M-Translationsversion https://git-lfs.github.com/spec/v1
oid sha256:835f3f7d88a86e05a882c6a6b6333da6ab874776385f85473798769d767c2fca
size 27
pixiv-image-with-caption
Dataset Card for Pixiv Daily Trending Illusions Dataset
Note, this dataset contains copyright issue, and is displayed for fun personal project only. Do not use it.
Dataset Summary
This dataset comprises 949 images scrapped from Pixiv's daily trend, specifically curated to include only illustrations that are illusions and suitable for all ages. Each image in the dataset is accompanied by a caption generated by the LLaVa model, providing a descriptive or interpretive text… See the full description on the dataset page: https://huggingface.co/datasets/Xiao215/pixiv-image-with-caption.testImageCaptions-7M-Translations-Arabic-subset-150000image_caption_regularization
Regularization Image Caption Dataset
Number of Images: 1976
Source
This is a subset of tomg-group-umd/pixelprose, converted to .csv format.
Files
people.csv: 1976 images with captions that contain one of these terms: ['person', 'people', 'man', 'men', 'woman', 'women']
pt-paligemma-multilingual-imagecaptionssafety-image-captions-1
