omlab/VLM-R1
This repository contains the dataset used in the paper VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model. Code: https://github.com/om-ai-lab/VLM-R1
18328
Add link to paper, github repo, visual grounding task category (#2)
Upload 2 files
Upload qwen2_5_vl_full_sft.yaml
Upload 2 files
Upload lisa_test.zip
Upload qwen2_5_vl_full_sft.yaml
Upload qwen2_5_vl_full_sft.yaml
Upload gui_multi-image.zip
Upload rec_jsons_internvl.zip
add sft yaml
add sft
add refgta json
coco
Upload 2 files
initial commit
