datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
PROVE
Trust but Verify: Programmatic VLM Evaluation in the Wild
Viraj Prabhu, Senthil Purushwalkam, An Yan, Caiming Xiong, Ran Xu
Explorer
| Paper
| Quickstart
Vision-Language Models (VLMs) often generate plausible but incorrect responses to visual queries. However, reliably quantifying the effect of such hallucinations in free-form responses to open-ended queries is challenging as it requires visually verifying each claim within the response. We propose Programmatic VLM… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/PROVE.CogAlign
Dataset Card for CogAlign
Dataset Description
Citation
Dataset Description
CogAlign is a post-training strategy for Vision Language Models (VLMs) aimed at enhancing their visual arithmetic capabilities. This repository presents the training data for CogAlign, a synthetic dataset containing 64,000 examples designed to facilitate this post-training process.
CogAlign is inspired by Piaget's theory of cognitive development and focuses on improving a VLM's understanding of… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/CogAlign.
