datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
nordjylland-news-image-captioning
Dataset Card for "nordjylland-news-image-captioning"
Dataset Summary
This dataset is a collection of image-caption pairs from the Danish newspaper TV2 Nord.
Supported Tasks and Leaderboards
Image captioning is the intended task for this dataset. No leaderboard is active at this point.
Languages
The dataset is available in Danish (da).
Dataset Structure
An example from the dataset looks as follows.
{
"file_name": "1.jpg",
"caption":… See the full description on the dataset page: https://huggingface.co/datasets/alexandrainst/nordjylland-news-image-captioning.Chinese_Children_Image_Captioning_Dataset_Split0
CODP-1200:Children Oral Description of Picture(Chinese-Child-Captions)
CODP-1200: An AIGC based benchmark for assisting in child language acquisition
数据集介绍
目前已知最大的儿童图像描述数据集,children image captioning
共有1200张图片
每张图片对应五个中文描述,每两张图片为一组
描述文字600*5=3000
如果使用CODP-1200数据集,请引用以下文章
@article{LENG2024102627,
title = {CODP-1200: An AIGC based benchmark for assisting in child language acquisition},
journal = {Displays},
volume = {82},
pages = {102627},
year =… See the full description on the dataset page: https://huggingface.co/datasets/svjack/Chinese_Children_Image_Captioning_Dataset_Split0.IntraOral_Gingivitis_Image_Captioning
A DENTAL INTRAORAL IMAGE DATASET OF GINGIVITIS FOR IMAGE CAPTIONING
Dataset Description
This dataset is a copy of A Dental IntraOral Image Dataset of Gingivitis for Image Captioning which is shared with the license CC BY 4.0.
This dataset contains 1,096 samples organized across multiple splits.
The dataset includes image data.
Splits
train: 732 samples
test: 182 samples
validation: 182 samples
Dataset Creation
This dataset was created using… See the full description on the dataset page: https://huggingface.co/datasets/ekacare/IntraOral_Gingivitis_Image_Captioning.image-captioning-turkish
Türkçe Image Captioning Veri Seti
Bu veri seti BLIP3o modelinin pretrain eğitiminde kullanılan BLIP3o-Pretrain-Long-Caption ve BLIP3o-Pretrain-Short-Caption veri setlerinin Türkçeye çevirilmiş bir alt parçasıdır. Orijinal veri setinin oluşturulması ile ilgili detaylı bilgiye BLIP-3o makalesi üzerinden ulaşabilirsiniz.
Veri seti Image-to-Text modellerinin eğitilmesinde veya ince ayar sürecinde kullanılabilir. Veri seti, orijinal veri setinin lisansı olan Apache 2.0 altında… See the full description on the dataset page: https://huggingface.co/datasets/ituperceptron/image-captioning-turkish.ImageCaptioning_EN-HUImageCaptioning_SmallParquetsChinese_Children_Image_Captioning_Dataset_Split1
CODP-1200:Children Oral Description of Picture(Chinese-Child-Captions)
CODP-1200: An AIGC based benchmark for assisting in child language acquisition
数据集介绍
目前已知最大的儿童图像描述数据集,children image captioning
共有1200张图片
每张图片对应五个中文描述,每两张图片为一组
描述文字600*5=3000
如果使用CODP-1200数据集,请引用以下文章
@article{LENG2024102627,
title = {CODP-1200: An AIGC based benchmark for assisting in child language acquisition},
journal = {Displays},
volume = {82},
pages = {102627},
year =… See the full description on the dataset page: https://huggingface.co/datasets/svjack/Chinese_Children_Image_Captioning_Dataset_Split1.COCO-Image-Captioningru_image_captioningmy_image_captioning_datasetucf101-captioned-mappednordjylland-news-image-captioningimageCaptioningimage-captioning-idPersian-Image-Captioning
Dataset Card for "Persian-Image-Captioning"
More Information needed
X-ray_Image_captioningImageCaptioning_SmallParquets_oldimage-captioning
Scene Description Dataset
A comprehensive dataset of anime-style images with detailed scene descriptions and tags. This dataset contains high-quality annotations for image understanding and scene analysis tasks.
Dataset Description
This dataset consists of anime-style images paired with detailed textual descriptions and comprehensive tag annotations. Each entry includes:
Image: High-resolution anime-style artwork
Tags: Extensive tag annotations covering character… See the full description on the dataset page: https://huggingface.co/datasets/alex43219/image-captioning.ImageCaptioningimage-captioning-FACAD-baseimage-captioning-retrievalA dataset of curated images used for captioning using LLMs and retrieval based on caption(text based not embedding based) using a go server.
ucf101-captionedsample-image-captioningfkr30k-image-captioning-dataset
Dataset Card for "fkr30k-image-captioning-dataset"
More Information needed
image-captioning-openimages-subsetkhmer-historical-heritage-image-captioningHCMUS-Vietnamese-Image-captioning-for-visually-impaired
Vietnamese Image Captioning for Visually Impaired
Bộ dữ liệu Image Captioning tiếng Việt chuyên dụng hỗ trợ người khiếm thị di chuyển an toàn trong môi trường Giao thông và Trong nhà.
Cấu trúc hiển thị (Dataset Viewer)
file_name: Mã định danh ảnh gốc (ID).
image: Hình ảnh thực tế (được tự động ánh xạ từ file_name).
caption: Nội dung mô tả súc tích và chỉ dẫn an toàn.
Thống kê dữ liệu
Tổng số ảnh gốc: 8,000 ảnh.
Tổng số mẫu huấn luyện: 40,000 mẫu.
Tỷ lệ phân… See the full description on the dataset page: https://huggingface.co/datasets/pqthinh232/HCMUS-Vietnamese-Image-captioning-for-visually-impaired.Arabic-Image-Captioning_100M
Arabic Image Captioning Dataset (100M Sample)
The first large-scale Arabic multimodal dataset.
This groundbreaking dataset contains 100 million Arabic image captions, representing the first comprehensive Arabic multimodal resource of this scale and quality. Generated using our Mutarjim translation model, this dataset addresses the critical gap in Arabic multimodal AI resources and enables researchers to develop sophisticated Arabic vision-language systems for the first time.… See the full description on the dataset page: https://huggingface.co/datasets/Misraj/Arabic-Image-Captioning_100M.image-captioning-FACAD-smalltest
