CoolFace
Datasetpublic

txsgbb/VL-RewardBench

Dataset Card for VLRewardBench Project Page: https://vl-rewardbench.github.io Dataset Summary VLRewardBench is a comprehensive benchmark designed to evaluate vision-language generative reward models (VL-GenRMs) across visual perception, hallucination detection, and reasoning tasks. The benchmark contains 1,250 high-quality examples specifically curated to probe model limitations. Dataset Structure Each instance consists of multimodal queries… See the full description on the dataset page: https://huggingface.co/datasets/txsgbb/VL-RewardBench.

sourceHugging Facemitupdated 5mo agoView on Hugging Face
0likes33downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
txsgbb/VL-RewardBench · CoolFace