datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
blenderbench-direct-results
BlenderBench Direct reproduction artifacts
This dataset preserves artifacts and provenance for an independent, community-run reproduction of the public BlenderBench task set. It is not an official Blender Foundation product, official BlenderBench submission, or leaderboard result.
Experiment
Dataset: DietCoke4671/BlenderBench revision 203e4d325e9438ca55b29bdfc4f6a90842d74e68
Attribution: DietCoke4671 and contributors, CC BY 4.0
Generation model: gpt-6-astra
Codex… See the full description on the dataset page: https://huggingface.co/datasets/michaelgold/blenderbench-direct-results.NDBC_Wave_Direction_Spectrumdirectv-zocalos-5fps
Dataset Card for "directv-zocalos-5fps"
More Information needed
heb-directfit-1mDirectional_Guidance
Directional Guidance
This dataset provides a benchmark for evaluating Vision-Language Models (VLMs) in their ability to guide users to adjust an image to better answer a relevant question.
Dataset Description
The Directional Guidance dataset focuses on Visual Question Answering (VQA) tasks where a model needs to evaluate visual information sufficiency and guide the user on where to reposition the camera if the image lacks necessary details. This dataset addresses a unique… See the full description on the dataset page: https://huggingface.co/datasets/LeoLee7/Directional_Guidance.directionan_sign_detectionorigami-direct-tiny
Origami Direct Crease Pattern Dataset
A multiview image dataset for training models to predict complete origami crease patterns from 3D visualizations.
Task
Given 14 camera views of a folded origami shape, predict the complete crease pattern as a FOLD JSON (vertices, edges, mountain/valley assignments).
Dataset Structure
Each example contains:
Field
Type
Description
id
string
Unique sample ID (e.g., grid4_4c_0000)
images
list[string]
14 PNG paths —… See the full description on the dataset page: https://huggingface.co/datasets/Origametry/origami-direct-tiny.GeoQA-8K-direct-synthesizingThis dataset supports the unsupervised post-training of multi-modal large language models (MLLMs) as described in the paper Unsupervised Post-Training for Multi-Modal LLM Reasoning via GRPO. It's designed to enable continual self-improvement without external supervision, using a self-rewarding mechanism based on majority voting over multiple sampled responses. The dataset is used to improve the reasoning ability of MLLMs, as demonstrated by significant improvements on benchmarks like MathVista… See the full description on the dataset page: https://huggingface.co/datasets/WaltonFuture/GeoQA-8K-direct-synthesizing.geometry3k-direct-synthesizingThis dataset is used in the paper Unsupervised Post-Training for Multi-Modal LLM Reasoning via GRPO to improve the reasoning abilities of multi-modal large language models (MLLMs). It contains image-text pairs, where each pair consists of an image, a problem described in text, and the corresponding answer. The dataset is designed for unsupervised post-training of MLLMs.
🐙 GitHub Repo: waltonfuture/MM-UPT
📜 Paper (arXiv): Unsupervised Post-Training for Multi-Modal LLM Reasoning via GRPO… See the full description on the dataset page: https://huggingface.co/datasets/WaltonFuture/geometry3k-direct-synthesizing.G-Directed-Graph
Dataset Card
Add more information here
This dataset was produced with DataDreamer 🤖💤. The synthetic dataset card can be found here.
banners-directv_peach
Dataset Card for "banners-directv_peach"
More Information needed
heb-direct-rendervisualpuzzles-direct-gemini25provisualpuzzles-direct-gemini3proObject_direction_1MMR1-direct-synthesizingThis dataset is used in the paper Unsupervised Post-Training for Multi-Modal LLM Reasoning via GRPO to evaluate the performance of an unsupervised post-training framework (MM-UPT) for multi-modal LLMs. The dataset contains image-text pairs where the text represents a problem or question, and the corresponding answer. It is used to demonstrate the ability of MM-UPT to improve the reasoning capabilities of the Qwen2.5-VL-7B model without relying on any external supervised data.
🐙 GitHub Repo:… See the full description on the dataset page: https://huggingface.co/datasets/WaltonFuture/MMR1-direct-synthesizing.arc-agi-transduction100k-direct-ftvstar_direct_attributes_seal_zoomdirectv-zocalos-2fps
Dataset Card for "directv-zocalos-2fps"
More Information needed
direction_horizontal_tacThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "Unitree_G1_Inspire",
"total_episodes": 1,
"total_frames": 775,
"total_tasks": 1,
"total_videos": 0,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/eunjuri/direction_horizontal_tac.directv-zocalos-agosto-5fps
Dataset Card for "directv-zocalos-agosto-5fps"
More Information needed
directv-zocalos_1.0fps_03-07-2023_05-07-2023
Dataset Card for "directv-zocalos_1.0fps_03-07-2023_05-07-2023"
More Information needed
direct_difficulty_final_reviewdirectv-zocalos-agosto-5fps_vectors
Dataset Card for "directv-zocalos-agosto-5fps_vectors"
More Information needed
puzzlevqa-direct-gemini3prodirectv-zocalos-new-test-1fps
Dataset Card for "directv-zocalos-new-test-1fps"
More Information needed
directv-zocalos_1.0fps_03-08-2023_05-08-2023
Dataset Card for "directv-zocalos_1.0fps_03-08-2023_05-08-2023"
More Information needed
mmiq-direct-gemini25prodirectv-zocalos
Dataset Card for "directv-zocalos"
More Information needed
arc-agi-ft-direct-v1
