datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
directv-zocalos-5fps
Dataset Card for "directv-zocalos-5fps"
More Information needed
heb-directfit-1mGeoQA-8K-direct-synthesizingThis dataset supports the unsupervised post-training of multi-modal large language models (MLLMs) as described in the paper Unsupervised Post-Training for Multi-Modal LLM Reasoning via GRPO. It's designed to enable continual self-improvement without external supervision, using a self-rewarding mechanism based on majority voting over multiple sampled responses. The dataset is used to improve the reasoning ability of MLLMs, as demonstrated by significant improvements on benchmarks like MathVista… See the full description on the dataset page: https://huggingface.co/datasets/WaltonFuture/GeoQA-8K-direct-synthesizing.geometry3k-direct-synthesizingThis dataset is used in the paper Unsupervised Post-Training for Multi-Modal LLM Reasoning via GRPO to improve the reasoning abilities of multi-modal large language models (MLLMs). It contains image-text pairs, where each pair consists of an image, a problem described in text, and the corresponding answer. The dataset is designed for unsupervised post-training of MLLMs.
🐙 GitHub Repo: waltonfuture/MM-UPT
📜 Paper (arXiv): Unsupervised Post-Training for Multi-Modal LLM Reasoning via GRPO… See the full description on the dataset page: https://huggingface.co/datasets/WaltonFuture/geometry3k-direct-synthesizing.G-Directed-Graph
Dataset Card
Add more information here
This dataset was produced with DataDreamer 🤖💤. The synthetic dataset card can be found here.
banners-directv_peach
Dataset Card for "banners-directv_peach"
More Information needed
heb-direct-rendervisualpuzzles-direct-gemini25provisualpuzzles-direct-gemini3proObject_direction_1MMR1-direct-synthesizingThis dataset is used in the paper Unsupervised Post-Training for Multi-Modal LLM Reasoning via GRPO to evaluate the performance of an unsupervised post-training framework (MM-UPT) for multi-modal LLMs. The dataset contains image-text pairs where the text represents a problem or question, and the corresponding answer. It is used to demonstrate the ability of MM-UPT to improve the reasoning capabilities of the Qwen2.5-VL-7B model without relying on any external supervised data.
🐙 GitHub Repo:… See the full description on the dataset page: https://huggingface.co/datasets/WaltonFuture/MMR1-direct-synthesizing.arc-agi-transduction100k-direct-ftvstar_direct_attributes_seal_zoomdirectv-zocalos-2fps
Dataset Card for "directv-zocalos-2fps"
More Information needed
direction_horizontal_tacThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "Unitree_G1_Inspire",
"total_episodes": 1,
"total_frames": 775,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/eunjuri/direction_horizontal_tac.directv-zocalos-agosto-5fps
Dataset Card for "directv-zocalos-agosto-5fps"
More Information needed
directv-zocalos_1.0fps_03-07-2023_05-07-2023
Dataset Card for "directv-zocalos_1.0fps_03-07-2023_05-07-2023"
More Information needed
direct_difficulty_final_reviewdirectv-zocalos-agosto-5fps_vectors
Dataset Card for "directv-zocalos-agosto-5fps_vectors"
More Information needed
puzzlevqa-direct-gemini3prodirectv-zocalos-new-test-1fps
Dataset Card for "directv-zocalos-new-test-1fps"
More Information needed
directv-zocalos_1.0fps_03-08-2023_05-08-2023
Dataset Card for "directv-zocalos_1.0fps_03-08-2023_05-08-2023"
More Information needed
mmiq-direct-gemini25prodirectv-zocalos
Dataset Card for "directv-zocalos"
More Information needed
arc-agi-ft-direct-v1self-imagine-direct-inferenceocr-output-Directive017-1761352144
Document OCR using LightOnOCR-0.9B-32k-1025
This dataset contains OCR results from images in stckmn/ocr-input-Directive017-1761352140 using LightOnOCR, a fast and compact 1B OCR model.
Processing Details
Source Dataset: stckmn/ocr-input-Directive017-1761352140
Model: lightonai/LightOnOCR-0.9B-32k-1025
Vocabulary Size: 32k tokens
Number of Samples: 10
Processing Time: 1.7 min
Processing Date: 2025-10-25 00:32 UTC
Configuration
Image Column: image
Output… See the full description on the dataset page: https://huggingface.co/datasets/stckmn/ocr-output-Directive017-1761352144.directv-zocalos_1.0fps_21-08-2023_24-08-2023
Dataset Card for "directv-zocalos_1.0fps_21-08-2023_24-08-2023"
More Information needed
ocr-output-Directive017-1761354526
Document OCR using NuMarkdown-8B-Thinking
This dataset contains markdown-formatted OCR results from images in stckmn/ocr-input-Directive017-1761354522 using NuMarkdown-8B-Thinking.
Processing Details
Source Dataset: stckmn/ocr-input-Directive017-1761354522
Model: numind/NuMarkdown-8B-Thinking
Number of Samples: 21
Processing Time: 3.8 minutes
Processing Date: 2025-10-25 01:17 UTC
Configuration
Image Column: image
Output Column: markdown
Dataset Split:… See the full description on the dataset page: https://huggingface.co/datasets/stckmn/ocr-output-Directive017-1761354526.bilsem-batch-puzzlevqa-direct-gemini-2.5-pro
