datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
self_repair_gripper_dagger
self_repair_gripper_dagger
Robot self-repair, DAgger rollouts with operator corrections on the same task as self_repair_gripper_bc.
Real-robot bimanual manipulation data collected on a YAM arm pair, released as part
of the Flex-π project. Stored in LeRobot v2.1 format with synchronized RGB and
metric depth from three cameras.
At a glance
Episodes
2,154
Frames
609,385
Duration
~5.6 h @ 30 fps
Tasks
1
Robot
yam (bimanual)
Cameras
cam_high… See the full description on the dataset page: https://huggingface.co/datasets/flex-pi/self_repair_gripper_dagger.ex-repairci-repair-bench
CI-REPAIR-BENCH
Overview
CI-REPAIR-BENCH is a benchmark dataset for research on Continuous Integration (CI) failures and automated repair in Python repositories.
The dataset contains 567 CI failure instances collected from 105 real-world GitHub repositories, all written in Python.Each instance captures a CI workflow failure, its logs, the corresponding code diff, and repository-level metadata.
Dataset Statistics
Programming language: Python
Number… See the full description on the dataset page: https://huggingface.co/datasets/ci-benchmark-user/ci-repair-bench.self_repair_gripper_bc
self_repair_gripper_bc
Robot self-repair, human teleoperation (BC): install a gripper into an empty holder, drive a screw with a screwdriver, then clear the table.
Real-robot bimanual manipulation data collected on a YAM arm pair, released as part
of the Flex-π project. Stored in LeRobot v2.1 format with synchronized RGB and
metric depth from three cameras.
At a glance
Episodes
802
Frames
1,278,804
Duration
~11.8 h @ 30 fps
Tasks
1
Robot
yam… See the full description on the dataset page: https://huggingface.co/datasets/flex-pi/self_repair_gripper_bc.self-repair-gripper-v2.1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam_bimanual",
"total_episodes": 97,
"total_frames": 157583,
"total_tasks": 1,
"total_videos": 291,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:97"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/self-repair-gripper-v2.1.self-repair-gripper-v2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam_bimanual",
"total_episodes": 97,
"total_frames": 157583,
"total_tasks": 1,
"total_videos": 291,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:97"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/self-repair-gripper-v2.self-repair-gripper-v2.4This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam_bimanual",
"total_episodes": 97,
"total_frames": 157583,
"total_tasks": 1,
"total_videos": 291,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:97"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/self-repair-gripper-v2.4.self-repair-gripperThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam_bimanual",
"total_episodes": 130,
"total_frames": 267566,
"total_tasks": 1,
"total_videos": 390,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:130"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/self-repair-gripper.VeriLoop-Structural-Repair-Verified
VLR-StructuralRepair v1.0.0 — non-regressive repair of real semantic defects
Evidence-convergent supervision for function-level semantic repair under a
hidden set of protected obligations. A candidate is positive only when it
preserves every already-satisfied obligation and strictly repairs at least
one. Aggregate improvement that breaks a protected obligation is a negative,
however far the total failure count drops.
The previous generation of this dataset… See the full description on the dataset page: https://huggingface.co/datasets/tsinghua-sigs-robot-lab/VeriLoop-Structural-Repair-Verified.self-repair-gripper-v2.2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam_bimanual",
"total_episodes": 97,
"total_frames": 157583,
"total_tasks": 1,
"total_videos": 291,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:97"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/self-repair-gripper-v2.2.self-repair-gripper-dagger-r1-v1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "yam_bimanual",
"total_episodes": 741,
"total_frames": 159186,
"total_tasks": 1,
"total_videos": 2223,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:741"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/self-repair-gripper-dagger-r1-v1.grab_block_20260909_160043_repaired_plus2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"
],
"shape": [
6… See the full description on the dataset page: https://huggingface.co/datasets/Hailey-5-2026/grab_block_20260909_160043_repaired_plus2.moss-voice-identity-repairs
MOSS voice-acting v2 -- repaired takes
For each voice profile, every take whose ECAPA speaker similarity to the voice's reference fell
below 0.40, regenerated with that voice's identity LoRA (see
laion/moss-voice-identity-loras) merged at scale 1.0 on top of the identical condition
adapters at the identical lambdas.
Nothing here replaces anything. The original takes are untouched and remain part of the
corpus; low-similarity takes are kept deliberately, because they are useful… See the full description on the dataset page: https://huggingface.co/datasets/laion/moss-voice-identity-repairs.lca-ci-builds-repair
🏟️ Long Code Arena (CI builds repair)
This is the benchmark for CI builds repair task as part of the
🏟️ Long Code Arena benchmark.
🛠️ Task. Given the logs of a failed GitHub Actions workflow and the corresponding repository snapshot,
repair the repository contents in order to make the workflow pass.
All the data is collected from repositories published under permissive licenses (MIT, Apache-2.0, BSD-3-Clause, and BSD-2-Clause). The datapoints can be removed upon request.
To… See the full description on the dataset page: https://huggingface.co/datasets/JetBrains-Research/lca-ci-builds-repair.appliancedb-error-codes-repair-database
ApplianceDB: Home Appliance Error Codes & Ranked Repairs
Full dataset: appliancedb.dataengineered.io · $99 one-time (Repair Intelligence Snapshot: commercial licence + SQLite and Parquet builds; the same rows as this sample) → Buy on Stripe · the same sample on Kaggle
Relational database mapping 438 home-appliance error codes across 13 brands and 26 (brand, appliance-type) pairs to 288 ranked repair procedures with DIY difficulty tiers. Every code is identified by its… See the full description on the dataset page: https://huggingface.co/datasets/Ichlibitiche/appliancedb-error-codes-repair-database.polaris-53k-repaired
POLARIS-53K, label-repaired
49,289 of the 53,291 rows in
POLARIS-Project/Polaris-Dataset-53K,
with 4,580 stored answers corrected and 4,002 rows removed as unrepairable.
Measurements on the source set put its bad-label rate at roughly 15.9%
[14.3, 17.6] (two independent detectors agreeing on a 2,000-row sample).
Mislabelled rows are not uniformly distributed: they concentrate in the problems
models fail, which is exactly where a difficulty-calibration pipeline looks.… See the full description on the dataset page: https://huggingface.co/datasets/joanvelja/polaris-53k-repaired.mechanicdb-obd2-repair-sample
🔧 MechanicDB — OBD-II Diagnostic & Repair Database (Free Sample)
Full dataset: mechanicdb.dataengineered.io · $49 Standard (SAE) · $149 OEM Complete, one-time → Buy Standard · Buy OEM Complete · the same sample on Kaggle
The free developer sample of MechanicDB: an automotive dataset mapping OBD-II
Diagnostic Trouble Codes (DTCs) to ranked repair procedures with DIY
difficulty ratings, aftermarket parts-cost ranges (USD), labor-hour estimates,
and step-by-step instructions.
90… See the full description on the dataset page: https://huggingface.co/datasets/Ichlibitiche/mechanicdb-obd2-repair-sample.tabrepair-science-repair-under-shift
TabRepair Science: Repair Under Shift
TabRepair Science is a finite authored benchmark for a deceptively hard
question: does better tabular cell repair produce better downstream models
under distribution shift?
The 3,648-row pilot spans three structural generator families, missingness and
present-value contamination, four test regimes, eight repair representations,
and five downstream learners. A separate eight-world sensitivity layer tests a
damage-aware v2 candidate without… See the full description on the dataset page: https://huggingface.co/datasets/haidang2405/tabrepair-science-repair-under-shift.satd-repayment-context
SATD Repayment Context Dataset
Extends the SATD Repayment replication package (Python: 58,722 rows,
Java: 97,347 rows) with per-SATD repository context, all anchored at
parent(deleted_in_commit) (the commit right before the fix), so no information
from the repayment itself leaks into the context. Total package size: ~1.56 GB.
Contents
data/
python_final.parquet -- main SATD table (Python), 58,722 rows, 130 MB
java_final.parquet -- main SATD table… See the full description on the dataset page: https://huggingface.co/datasets/ngducloc1112002/satd-repayment-context.indist-tool-v0-pool-v2-gpt55-1k_issue_rewritten_prompt-v3-v5_swesmith_metadata_repairedcontext-repair-benchmark
ThoughtDAG Context Repair Benchmark
What happens after one wrong assumption enters a long LLM conversation?
This dataset turns context editing into a measurable intervention. Each synthetic case starts with a clean fact, introduces a false update, lets the error propagate through one to three downstream turns, and then asks the same final question under five graph conditions:
clean
polluted
source_prune
subgraph_prune
recompute_descendants
The central question is not only… See the full description on the dataset page: https://huggingface.co/datasets/thoughtdag/context-repair-benchmark.code_ujb_repairSysMLv2_Repair_with_SLMs
SysMLv2 Repair with SLMs
Dataset used in "Automated Semantic Fault Localization in SysML v2: A Human-in-the-Loop Framework Using Knowledge-Graph Augmented LLMs", presented at INCOSE International Symposium 2026.
Dataset Structure
This dataset provides two configurations:
default: Contains train/validation/test splits used for fine-tuning small models. Samples exceeding 2048 tokens have been removed.
full: Contains complete dataset
Task
Given SysML v2 code… See the full description on the dataset page: https://huggingface.co/datasets/rohhaiil/SysMLv2_Repair_with_SLMs.tb21-eval-qwen35-action-only-20k-infra-repaired-c164-max32k-timeout2x
qwen35-action-only-20k — Terminal-Bench 2.1
Noncanonical Terminal-Bench 2.1 evaluation of violetxi/qwen35-4b-offline-echo-action-only-20k-tacc through the served
model ID qwen35-action-only-20k with Terminus-2.
Noncanonical run: timeout_multiplier=2 instead of 1.0; repair concurrency=164 exceeds 30. Do not compare this score directly with canonical TB2.1 leaderboard runs.
Result
Recorded trials: 445
Tasks / attempts: 89 × 5
Errored trials scored as zero: 250… See the full description on the dataset page: https://huggingface.co/datasets/violetxi/tb21-eval-qwen35-action-only-20k-infra-repaired-c164-max32k-timeout2x.eval_repair_menuThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101_follower",
"total_episodes": 1,
"total_frames": 526,
"total_tasks": 1,
"total_videos": 2,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ankithreddy/eval_repair_menu.eval_repair_rendaThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101_follower",
"total_episodes": 1,
"total_frames": 451,
"total_tasks": 1,
"total_videos": 2,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ankithreddy/eval_repair_renda.repair-aware-ibn
Repair-Aware IBN Intent-Conflict Benchmark
Labelled pairs of network intents for conflict detection and conflict resolution in
Intent-Based Networking. Each conflicting pair records not only that the two intents
conflict, but which predicate relation, if removed, resolves the conflict. That second
label is what the dataset exists for: it makes it possible to measure whether a detector
has learned what resolves a conflict rather than only what one looks like.
Companion artefact… See the full description on the dataset page: https://huggingface.co/datasets/shahoismael/repair-aware-ibn.eval_repair_2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "so101_follower",
"total_episodes": 1,
"total_frames": 1098,
"total_tasks": 1,
"total_videos": 2,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ankithreddy/eval_repair_2.xarm-n_pick_repairedThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "u850",
"total_episodes": 48,
"total_frames": 29949,
"total_tasks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:48"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path": "videos/{video_key}/chunk-{chunk_index:03d}/file-{file_index:03d}.mp4",
"features": {… See the full description on the dataset page: https://huggingface.co/datasets/kasiv008/xarm-n_pick_repaired.bibletts-asante-twi-repaired
BibleTTS Asante Twi — Repaired Transcripts
The Asante Twi transcripts released with BibleTTS have had the
characters ɛ (U+025B) and ɔ (U+0254) stripped out. This dataset restores them.
Audio is not included. This is a drop-in replacement for the .txt files that ship with the
BibleTTS Asante Twi package, matched by clip ID.
The problem
Both are Twi vowels, and both are required by the orthography. Measured across the released
Asante Twi transcripts:
Character… See the full description on the dataset page: https://huggingface.co/datasets/danieldzikunuofmarvel/bibletts-asante-twi-repaired.
