CoolFace
Datasetpublic

zlab-princeton/Vero-2.5M-unfiltered

Vero-2.5M-unfiltered [!Note] This repository contains the full unfiltered dataset used to construct Vero-600k and Vero-1.6M, before question and answer filtering. Note that task categories are not balanced in this dataset. Vero is a fully open reinforcement learning (RL) recipe for training and evaluating multi-task visual reasoning with vision-language models. This repository contains the Vero-2.5M-unfiltered dataset, a curation of 2.5M reinforcement learning samples… See the full description on the dataset page: https://huggingface.co/datasets/zlab-princeton/Vero-2.5M-unfiltered.

sourceHugging Faceupdated 4mo agoView on Hugging Face
1likes2.4kdownloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
zlab-princeton/Vero-2.5M-unfiltered · CoolFace