CoolFace
Datasetpublic

liuhaotian/LLaVA-CC3M-Pretrain-595K

LLaVA Visual Instruct CC3M 595K Pretrain Dataset Card Dataset details Dataset type: LLaVA Visual Instruct CC3M Pretrain 595K is a subset of CC-3M dataset, filtered with a more balanced concept coverage distribution. Captions are also associated with BLIP synthetic caption for reference. It is constructed for the pretraining stage for feature alignment in visual instruction tuning. We aim to build large multimodal towards GPT-4 vision/language capability. Dataset… See the full description on the dataset page: https://huggingface.co/datasets/liuhaotian/LLaVA-CC3M-Pretrain-595K.

sourceHugging Faceotherupdated 3y agoView on Hugging Face
180likes599downloads
settings

This repository belongs to liuhaotian on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameLLaVA-CC3M-Pretrain-595K
visibilitypublic
licenceother
gatedno
ownerliuhaotian
Account settings