CoolFace
Datasetpublic

txsgbb/VL-RewardBench

Dataset Card for VLRewardBench Project Page: https://vl-rewardbench.github.io Dataset Summary VLRewardBench is a comprehensive benchmark designed to evaluate vision-language generative reward models (VL-GenRMs) across visual perception, hallucination detection, and reasoning tasks. The benchmark contains 1,250 high-quality examples specifically curated to probe model limitations. Dataset Structure Each instance consists of multimodal queries… See the full description on the dataset page: https://huggingface.co/datasets/txsgbb/VL-RewardBench.

sourceHugging Facemitupdated 5mo agoView on Hugging Face
0likes33downloads
fileinference_results.zip81.1 MBdownload

txsgbb/VL-RewardBench · main · files are served by the source, never re-hosted here