datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
qvq-r1-RLViet-qvq-r1
Description
This dataset is a Vietnamese translation of the ahmedheakl/qvq-r1, intended for training and evaluating multimodal Vision–Language Models (VLMs) on visual reasoning tasks involving document-style images such as receipts, forms, invoices.,
Each example includes:
An input image containing text (typically scanned documents),
A conversation simulating a user question and an assistant’s step-by-step reasoning leading to the answer,
A Vietnamese version of the full… See the full description on the dataset page: https://huggingface.co/datasets/5CD-AI/Viet-qvq-r1.qvq-r1virgo_qvqbo16_acc_0_3
