CoolFace
Datasetpublic

damerajee/Hindi-LLaVA-CC3M-Pretrain-595K

LLaVA Visual Instruct CC3M 595K Pretrain Dataset Card Dataset details Dataset type: LLaVA Visual Instruct CC3M Pretrain 595K is a subset of CC-3M dataset, filtered with a more balanced concept coverage distribution. Captions are also associated with BLIP synthetic caption for reference. It is constructed for the pretraining stage for feature alignment in visual instruction tuning. We aim to build large multimodal towards GPT-4 vision/language capability. Dataset… See the full description on the dataset page: https://huggingface.co/datasets/damerajee/Hindi-LLaVA-CC3M-Pretrain-595K.

sourceHugging Faceotherupdated 2y agoView on Hugging Face
0likes16downloads
7 commits on main
e7881af2y ago

Update README.md

damerajee
79488d22y ago

Update README.md

damerajee
744309a2y ago

Update README.md

damerajee
f14c0fb2y ago

Update README.md

damerajee
b4b703f2y ago

Update README.md

damerajee
3ac98f32y ago

Upload dataset

damerajee
247ec9b2y ago

initial commit

damerajee