datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
VLM_img2end_newVerMultiThis repository contains the data presented in LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.
Project page: https://forjadeforest.github.io/LMM-R1-ProjectPage
VLM_img2endVLM-SFTVLMBench_datasetVLMEval-MovieChat1kVLM-150M
Dataset Card for VLM-150M
VLM-150M is a large-scale image-text dataset that has been recaptioned using an SFT-enhanced Qwen2VL model to enhance the alignment and detail of textual descriptions.
Dataset Sources
Repository: [https://zxwei.site/hqclip/)
Usage Guide
See https://github.com/w1oves/hqclip/blob/main/README.md#dataset-usage-guide.
animetimm-Danbooru-VLMsmall-publaynet-wds
Small PubLayNet (WebDataset)
This dataset consists in the first WebDataset shards of PubLayNet from http://storage.googleapis.com/nvdata-publaynet
It is mostly used to test the WebDataset integration within the Hugging Face ecosystem.
vlm3r_sample_10kvlm_tunneleval_benchmark_vlm_hallucollect_vlm_v3_1_rl-0521-backupGoPro_Deblur_Spkhub_dataAgent_document_1
