CoolFace
Datasetpublic

Tran1312/Image-Caption

BLIP3o Long-Caption 100K Image-Text Subset This dataset is a locally reorganized subset of BLIP3o/BLIP3o-Pretrain-Long-Caption, containing approximately 100,000 image-text pairs selected from the original BLIP3o long-caption pretraining dataset [1]. The original BLIP3o long-caption collection contains approximately 27 million images, each paired with a long caption of roughly 120 tokens generated using Qwen2.5-VL-7B-Instruct [1]. The BLIP3-o project was introduced as part of a… See the full description on the dataset page: https://huggingface.co/datasets/Tran1312/Image-Caption.

sourceHugging Faceupdated 22h agoView on Hugging Face
0likes26downloads
settings

This repository belongs to Tran1312 on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameImage-Caption
visibilitypublic
licencenot set
gatedno
ownerTran1312
Account settings
Tran1312/Image-Caption · CoolFace