datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
4c_differential_cameras_compareThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "koch_follower",
"total_episodes": 5,
"total_frames": 1791,
"total_tasks": 1,
"total_videos": 10,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:5"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ethanCSL/4c_differential_cameras_compare.omnidocbench-render-compare-parquet
OmniDocBench Render-and-Compare — Parquet Edition
Parquet-shard repackaging of
gt-free-ocr-metrics/omnidocbench-render-compare.
Overview
The pipeline processes each page of OmniDocBench
through a Qwen3.5-122B-A10B OCR model, renders the structured output back
to a PNG via HTML (reconstructed), and compares it against the original
page scan (masked_original) using reference-free visual metrics.
Five OCR extraction variants are provided, each targeting a different subset
of… See the full description on the dataset page: https://huggingface.co/datasets/gt-free-ocr-metrics/omnidocbench-render-compare-parquet.compare_value_kind_3816auditkit-testrun-compare
auditkit-testrun-compare
Built using AuditKIT — evaluate any model on any dataset and any task.
Method
compare_models
Model
—
Artifact
compare
Published
2026-09-02 06:50 UTC
Usage
from datasets import load_dataset
ds = load_dataset("ram-lexsi/auditkit-testrun-compare")
compare-offlinegrpo-runpod-payload-public
compare-offline-grpo
Сравнение 6 методов оффлайн RL/RFT для дистилляции ризонинга.
Полный план — в CLAUDE.md. Этот файл — короткий обзор для быстрого старта.
Quickstart
# 1) deps
python -m venv .venv && source .venv/bin/activate
pip install -r requirements.txt
# 2) secrets
cp .env.example .env
# .env уже содержит OPENROUTER_API_KEY (см. CLAUDE.md). Заполни HF_TOKEN, WANDB_API_KEY.
# 3) eval-сеты + промпты для тренировки
python scripts/download_eval_sets.py
python… See the full description on the dataset page: https://huggingface.co/datasets/AlexWortega/compare-offlinegrpo-runpod-payload-public.CompareBench
CompareBench
CompareBench is a benchmark for evaluating visual comparison reasoning in vision-language models (VLMs),a fundamental yet understudied skill. CompareBench consists of 1,200 QA pairs across four tasks:
Quantity (600)
Geometric (200)
Spatial (100)
Temporal (300)
It is derived from two auxiliary datasets we constructed: TallyBench and OmniCaps.
Related Datasets
OmniCaps
TallyBench
Code
👉 CompareBench on GitHub
curriculum_1_compare_value_kind_drawing_3816evalap-compare-open-weight-models-31th-83
Compare Open Weight Models 31th (ID: 83)
Comparing open weight models
Overview
This dataset contains 44 experiments
from the EvalAP evaluation platform.
Datasets: Assistant IA - QA, MFS_questions_v01
Models evaluated: Groq/Llama-3-Groq-8B-Tool-Use, Qwen/Qwen3-VL-4B-Instruct, Qwen/Qwen3-VL-8B-Thinking, meta-llama/Llama-3.1-8B-Instruct, mistral-medium-2508, mistralai/Magistral-Small-2509, mistralai/Mistral-Small-3.2-24B-Instruct-2506… See the full description on the dataset page: https://huggingface.co/datasets/AgentPublic/evalap-compare-open-weight-models-31th-83.compare-count-enhanced-sftcurriculum_1_compare_value_kind_drawing_3816_v3compare_18-19_terminate_mergeThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "rby1",
"total_episodes": 428,
"total_frames": 98127,
"total_tasks": 2,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 15,
"splits": {
"train": "0:428"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/rainbowrobotics/compare_18-19_terminate_merge.agilex_roll_dice_and_compareThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "arx5_bimanual",
"total_episodes": 13,
"total_frames": 10829,
"total_tasks": 1,
"total_videos": 39,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 25,
"splits": {
"train": "0:13"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/villekuosmanen/agilex_roll_dice_and_compare.compare_count_4000curriculum_1_compare_count_drawing_4000so101_new_3cam_red_cube_black_pen_compare_v1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 20,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"
],
"shape": [
6… See the full description on the dataset page: https://huggingface.co/datasets/L7-Robotics/so101_new_3cam_red_cube_black_pen_compare_v1.voa-test-compare-semambaAudios are enhanced by https://github.com/RoyChao19477/SEMamba
Generated by https://github.com/RustedBytes/audio-parquet-merger
sfm-cpt-reasoning-comparecompare-models
Compare Models (Teeny-Tiny Castle)
This dataset is part of a tutorial tied to the Teeny-Tiny Castle, an open-source repository containing educational tools for AI Ethics and Safety research.
How to Use
from datasets import load_dataset
dataset = load_dataset("AiresPucrs/compare-models", split = 'train')
compare_testdo_as_i_do_jaw_compare_jun20This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": null,
"total_episodes": 1,
"total_frames": 2,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 1,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/thewisp/do_as_i_do_jaw_compare_jun20.self_compare_alpacaeval_related_8maxturns_complete_805_gpt3.5_cleanedsfm-cpt-reasoning-compare-pairedmedia-cloud-comparetrainset_political_party_big_compareself_compare_alpacaeval_repeated_5turns_diff_gpt4_responsessmolified-ocr-data-extract-and-compare
🤏 smolified-ocr-data-extract-and-compare
Intelligence, Distilled.
This is a synthetic training corpus generated by the Smolify Foundry.
It was used to train the corresponding model titou4ng/smolified-ocr-data-extract-and-compare.
📦 Asset Details
Origin: Smolify Foundry (Job ID: 790dd5fa)
Records: 21835
Type: Synthetic Instruction Tuning Data
⚖️ License & Ownership
This dataset is a sovereign asset owned by titou4ng.
Generated via Smolify.ai.
africa-owid-electricity-generation-from-solar-and-wind-compared-to-coal
Electricity Generation From Solar And Wind Compared To Coal | Africa (Our World in Data) | Africa (Electric Sheep Africa metadata inventory)
Size category: 1K<n<10K - Formats: parquet - Sector: energy - Engineered by Electric Sheep Africa
TL;DR
This dataset is part of the Electric Sheep Africa catalog on Hugging Face. It is indexed for African data discovery with standardized metadata, loading guidance, provenance notes, and analyst-oriented context.… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-owid-electricity-generation-from-solar-and-wind-compared-to-coal.africa-uganda-general-uace-performance-in-2023-compared-to-2022-1b6fbef8
General Uace Performance in 2023 Compared to 2022 | Africa (Uganda Bureau of Statistics)
7 rows - 1 Africa country/area - 2022-2025 - source table - Engineered by Electric Sheep Africa
TL;DR
This dataset contains 7 rows from Uganda Bureau of Statistics, covering General Uace Performance in 2023 Compared to 2022. It is published as ML-ready Parquet with consistent Hugging Face metadata, source provenance, and analysis-friendly loading examples.… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepafrica/africa-uganda-general-uace-performance-in-2023-compared-to-2022-1b6fbef8.asia-owid-electricity-generation-from-solar-and-wind-compared-to-coal
Electricity Generation From Solar And Wind Compared To Coal | Asia (Our World in Data)
🌏 2,151 observations · 49 Asia countries · 1965–2025 · Repackaged by Electric Sheep Asia
TL;DR
This dataset contains 2,151 observations of Electricity Generation From Solar And Wind Compared To Coal data across 49 Asia countries, spanning 1965–2025.
About the source
Source: Our World in Data
Publisher: Our World in Data
License: cc-by-4.0
Topic: Electricity… See the full description on the dataset page: https://huggingface.co/datasets/electricsheepasia/asia-owid-electricity-generation-from-solar-and-wind-compared-to-coal.vidore_v3_finance_en_english_compare-contrast_Chart
