CoolFace
Datasetpublic

Tran1312/Image-Caption

BLIP3o Long-Caption 100K Image-Text Subset This dataset is a locally reorganized subset of BLIP3o/BLIP3o-Pretrain-Long-Caption, containing approximately 100,000 image-text pairs selected from the original BLIP3o long-caption pretraining dataset [1]. The original BLIP3o long-caption collection contains approximately 27 million images, each paired with a long caption of roughly 120 tokens generated using Qwen2.5-VL-7B-Instruct [1]. The BLIP3-o project was introduced as part of a… See the full description on the dataset page: https://huggingface.co/datasets/Tran1312/Image-Caption.

sourceHugging Faceupdated 16h agoView on Hugging Face
0likes26downloads
4 commits on main
b6423cc16h ago

Create README.md

Tran1312
c0633dc16h ago

Upload Image.tar.xz

Tran1312
60eeace17h ago

Upload 2 files

Tran1312
9a3229617h ago

initial commit

Tran1312