datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
multimodal-rewardbench-2Paper: https://arxiv.org/abs/2512.16899
Multimodal RewardBench 2 (MMRB2). Processed from https://github.com/facebookresearch/MMRB2 .
If you find this useful, please cite with following bibtex:
@article{hu2025multimodalrewardbench2,
title={Multimodal RewardBench 2: Evaluating Omni Reward Models for Interleaved Text and Image},
author={Hu, Yushi and Askari-Hemmat, Reyhane and Hall, Melissa and Dinan, Emily and Zettlemoyer, Luke and Ghazvininejad, Marjan},
journal={arXiv preprint… See the full description on the dataset page: https://huggingface.co/datasets/rl-research/multimodal-rewardbench-2.easyr1-103k-4MP-stage-three-temp-1_7-RL-ui-vision-jedi-show-ui-desktop-gta-dense-rewardeasyr1-103k-4MP-stage-three-temp-1_7-RL-ui-vision-jedi-show-ui-desktop-gta-0p2-zero-dense-reward
