CoolFace
11 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01pppop7 /GQA Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage | 📚 Documentation | 🤗 Huggingface Datasets This Dataset This is a formatted version of GQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models. @inproceedings{hudson2019gqa, title={Gqa: A new dataset for real-world visual reasoning and compositional question… See the full description on the dataset page: https://huggingface.co/datasets/pppop7/GQA.image10M<n<100M0 likes910 downloads9mo agoHugging Face02pppppppppp2 /planeperturbed Dataset Card for "planeperturbed" More Information needed image1K<n<10K1 likes306 downloads3y agoHugging Face03pppop7 /OCR-VQA Dataset Card for "OCR-VQA" More Information needed image100K<n<1M0 likes268 downloads9mo agoHugging Face04pppop7 /textvqa Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage | 📚 Documentation | 🤗 Huggingface Datasets This Dataset This is a formatted version of TextVQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models. @inproceedings{singh2019towards, title={Towards vqa models that can read}, author={Singh, Amanpreet and Natarajan… See the full description on the dataset page: https://huggingface.co/datasets/pppop7/textvqa.image10K<n<100K0 likes234 downloads9mo agoHugging Face05PPPPPeter /artaimage10K<n<100K0 likes78 downloads1y agoHugging Face06pppop7 /VisualGenome_VG_100K_1_and_2 image0 likes20 downloads9mo agoHugging Face07ppporridge /DiT-STimage10K<n<100K0 likes16 downloads1y agoHugging Face08pppop7 /coco_train_2017image0 likes11 downloads9mo agoHugging Face09pppop7 /LLaVA-Pretrain LLaVA-Pretrain Dataset Pretraining data for LLaVA (Large Language and Vision Assistant). Description This dataset contains the pretraining data used in LLaVA training, including: blip_laion_cc_sbu_558k.json - Annotation file with 558K image-caption pairs images/ - Corresponding images Usage from huggingface_hub import snapshot_download # Download the dataset snapshot_download( repo_id="pppop7/LLaVA-Pretrain", repo_type="dataset"… See the full description on the dataset page: https://huggingface.co/datasets/pppop7/LLaVA-Pretrain.imageimage-to-text100K<n<1M0 likes8 downloads9mo agoHugging Face10ppppppps /eeeimagen<1K0 likes6 downloads3y agoHugging Face11pppp1432 /frau-federkiel-pinsimagen<1K0 likes23h agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.