CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01flex-pi /self_repair_gripper_dagger self_repair_gripper_dagger Robot self-repair, DAgger rollouts with operator corrections on the same task as self_repair_gripper_bc. Real-robot bimanual manipulation data collected on a YAM arm pair, released as part of the Flex-π project. Stored in LeRobot v2.1 format with synchronized RGB and metric depth from three cameras. At a glance Episodes 2,154 Frames 609,385 Duration ~5.6 h @ 30 fps Tasks 1 Robot yam (bimanual) Cameras cam_high… See the full description on the dataset page: https://huggingface.co/datasets/flex-pi/self_repair_gripper_dagger.tabularrobotics100K<n<1M0 likes563 downloads1mo agoHugging Face02nus-yam /ex-repairtabular1M<n<10M3 likes508 downloads3y agoHugging Face03ci-benchmark-user /ci-repair-bench CI-REPAIR-BENCH Overview CI-REPAIR-BENCH is a benchmark dataset for research on Continuous Integration (CI) failures and automated repair in Python repositories. The dataset contains 567 CI failure instances collected from 105 real-world GitHub repositories, all written in Python.Each instance captures a CI workflow failure, its logs, the corresponding code diff, and repository-level metadata. Dataset Statistics Programming language: Python Number… See the full description on the dataset page: https://huggingface.co/datasets/ci-benchmark-user/ci-repair-bench.tabularn<1K1 likes436 downloads10h agoHugging Face04flex-pi /self_repair_gripper_bc self_repair_gripper_bc Robot self-repair, human teleoperation (BC): install a gripper into an empty holder, drive a screw with a screwdriver, then clear the table. Real-robot bimanual manipulation data collected on a YAM arm pair, released as part of the Flex-π project. Stored in LeRobot v2.1 format with synchronized RGB and metric depth from three cameras. At a glance Episodes 802 Frames 1,278,804 Duration ~11.8 h @ 30 fps Tasks 1 Robot yam… See the full description on the dataset page: https://huggingface.co/datasets/flex-pi/self_repair_gripper_bc.tabularrobotics1M<n<10M0 likes395 downloads1mo agoHugging Face05YOLO2431 /self-repair-gripper-v2.1This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "yam_bimanual", "total_episodes": 97, "total_frames": 157583, "total_tasks": 1, "total_videos": 291, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:97" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/self-repair-gripper-v2.1.tabularrobotics100K<n<1M0 likes393 downloads2mo agoHugging Face06YOLO2431 /self-repair-gripper-v2This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "yam_bimanual", "total_episodes": 97, "total_frames": 157583, "total_tasks": 1, "total_videos": 291, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:97" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/self-repair-gripper-v2.tabularrobotics100K<n<1M0 likes375 downloads3mo agoHugging Face07YOLO2431 /self-repair-gripper-v2.4This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "yam_bimanual", "total_episodes": 97, "total_frames": 157583, "total_tasks": 1, "total_videos": 291, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:97" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/self-repair-gripper-v2.4.tabularrobotics1M<n<10M0 likes370 downloads2mo agoHugging Face08YOLO2431 /self-repair-gripperThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "yam_bimanual", "total_episodes": 130, "total_frames": 267566, "total_tasks": 1, "total_videos": 390, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:130" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/self-repair-gripper.tabularrobotics100K<n<1M0 likes255 downloads3mo agoHugging Face09tsinghua-sigs-robot-lab /VeriLoop-Structural-Repair-Verified VLR-StructuralRepair v1.0.0 — non-regressive repair of real semantic defects Evidence-convergent supervision for function-level semantic repair under a hidden set of protected obligations. A candidate is positive only when it preserves every already-satisfied obligation and strictly repairs at least one. Aggregate improvement that breaks a protected obligation is a negative, however far the total failure count drops. The previous generation of this dataset… See the full description on the dataset page: https://huggingface.co/datasets/tsinghua-sigs-robot-lab/VeriLoop-Structural-Repair-Verified.tabulartext-generation10K<n<100K0 likes248 downloads27d agoHugging Face10YOLO2431 /self-repair-gripper-v2.2This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "yam_bimanual", "total_episodes": 97, "total_frames": 157583, "total_tasks": 1, "total_videos": 291, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:97" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/self-repair-gripper-v2.2.tabularrobotics100K<n<1M0 likes245 downloads2mo agoHugging Face11YOLO2431 /self-repair-gripper-dagger-r1-v1This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "yam_bimanual", "total_episodes": 741, "total_frames": 159186, "total_tasks": 1, "total_videos": 2223, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:741" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/YOLO2431/self-repair-gripper-dagger-r1-v1.tabularrobotics100K<n<1M0 likes176 downloads2mo agoHugging Face12Hailey-5-2026 /grab_block_20260909_160043_repaired_plus2This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "names": [ "shoulder_pan.pos", "shoulder_lift.pos", "elbow_flex.pos", "wrist_flex.pos", "wrist_roll.pos", "gripper.pos" ], "shape": [ 6… See the full description on the dataset page: https://huggingface.co/datasets/Hailey-5-2026/grab_block_20260909_160043_repaired_plus2.tabularrobotics1K<n<10K0 likes156 downloads13d agoHugging Face13laion /moss-voice-identity-repairs MOSS voice-acting v2 -- repaired takes For each voice profile, every take whose ECAPA speaker similarity to the voice's reference fell below 0.40, regenerated with that voice's identity LoRA (see laion/moss-voice-identity-loras) merged at scale 1.0 on top of the identical condition adapters at the identical lambdas. Nothing here replaces anything. The original takes are untouched and remain part of the corpus; low-similarity takes are kept deliberately, because they are useful… See the full description on the dataset page: https://huggingface.co/datasets/laion/moss-voice-identity-repairs.tabulartext-to-speech1M<n<10M0 likes151 downloads13d agoHugging Face14JetBrains-Research /lca-ci-builds-repair 🏟️ Long Code Arena (CI builds repair) This is the benchmark for CI builds repair task as part of the 🏟️ Long Code Arena benchmark. 🛠️ Task. Given the logs of a failed GitHub Actions workflow and the corresponding repository snapshot, repair the repository contents in order to make the workflow pass. All the data is collected from repositories published under permissive licenses (MIT, Apache-2.0, BSD-3-Clause, and BSD-2-Clause). The datapoints can be removed upon request. To… See the full description on the dataset page: https://huggingface.co/datasets/JetBrains-Research/lca-ci-builds-repair.tabularn<1K3 likes145 downloads2y agoHugging Face15Ichlibitiche /appliancedb-error-codes-repair-database ApplianceDB: Home Appliance Error Codes & Ranked Repairs Full dataset: appliancedb.dataengineered.io · $99 one-time (Repair Intelligence Snapshot: commercial licence + SQLite and Parquet builds; the same rows as this sample) → Buy on Stripe · the same sample on Kaggle Relational database mapping 438 home-appliance error codes across 13 brands and 26 (brand, appliance-type) pairs to 288 ranked repair procedures with DIY difficulty tiers. Every code is identified by its… See the full description on the dataset page: https://huggingface.co/datasets/Ichlibitiche/appliancedb-error-codes-repair-database.tabular1K<n<10K0 likes141 downloads3d agoHugging Face16joanvelja /polaris-53k-repaired POLARIS-53K, label-repaired 49,289 of the 53,291 rows in POLARIS-Project/Polaris-Dataset-53K, with 4,580 stored answers corrected and 4,002 rows removed as unrepairable. Measurements on the source set put its bad-label rate at roughly 15.9% [14.3, 17.6] (two independent detectors agreeing on a 2,000-row sample). Mislabelled rows are not uniformly distributed: they concentrate in the problems models fail, which is exactly where a difficulty-calibration pipeline looks.… See the full description on the dataset page: https://huggingface.co/datasets/joanvelja/polaris-53k-repaired.tabulartext-generation10K<n<100K0 likes136 downloads16d agoHugging Face17Ichlibitiche /mechanicdb-obd2-repair-sample 🔧 MechanicDB — OBD-II Diagnostic & Repair Database (Free Sample) Full dataset: mechanicdb.dataengineered.io · $49 Standard (SAE) · $149 OEM Complete, one-time → Buy Standard · Buy OEM Complete · the same sample on Kaggle The free developer sample of MechanicDB: an automotive dataset mapping OBD-II Diagnostic Trouble Codes (DTCs) to ranked repair procedures with DIY difficulty ratings, aftermarket parts-cost ranges (USD), labor-hour estimates, and step-by-step instructions. 90… See the full description on the dataset page: https://huggingface.co/datasets/Ichlibitiche/mechanicdb-obd2-repair-sample.tabular1K<n<10K0 likes133 downloads3d agoHugging Face18haidang2405 /tabrepair-science-repair-under-shift TabRepair Science: Repair Under Shift TabRepair Science is a finite authored benchmark for a deceptively hard question: does better tabular cell repair produce better downstream models under distribution shift? The 3,648-row pilot spans three structural generator families, missingness and present-value contamination, four test regimes, eight repair representations, and five downstream learners. A separate eight-world sensitivity layer tests a damage-aware v2 candidate without… See the full description on the dataset page: https://huggingface.co/datasets/haidang2405/tabrepair-science-repair-under-shift.tabulartabular-regression100K<n<1M0 likes127 downloads25d agoHugging Face19ngducloc1112002 /satd-repayment-context SATD Repayment Context Dataset Extends the SATD Repayment replication package (Python: 58,722 rows, Java: 97,347 rows) with per-SATD repository context, all anchored at parent(deleted_in_commit) (the commit right before the fix), so no information from the repayment itself leaks into the context. Total package size: ~1.56 GB. Contents data/ python_final.parquet -- main SATD table (Python), 58,722 rows, 130 MB java_final.parquet -- main SATD table… See the full description on the dataset page: https://huggingface.co/datasets/ngducloc1112002/satd-repayment-context.tabular100K<n<1M0 likes104 downloads2mo agoHugging Face20VmaxRL /indist-tool-v0-pool-v2-gpt55-1k_issue_rewritten_prompt-v3-v5_swesmith_metadata_repairedtabularn<1K0 likes101 downloads4mo agoHugging Face21thoughtdag /context-repair-benchmark ThoughtDAG Context Repair Benchmark What happens after one wrong assumption enters a long LLM conversation? This dataset turns context editing into a measurable intervention. Each synthetic case starts with a clean fact, introduces a false update, lets the error propagate through one to three downstream turns, and then asks the same final question under five graph conditions: clean polluted source_prune subgraph_prune recompute_descendants The central question is not only… See the full description on the dataset page: https://huggingface.co/datasets/thoughtdag/context-repair-benchmark.tabular1K<n<10K0 likes87 downloads1mo agoHugging Face22ZHENGRAN /code_ujb_repairtabularn<1K0 likes75 downloads3y agoHugging Face23rohhaiil /SysMLv2_Repair_with_SLMs SysMLv2 Repair with SLMs Dataset used in "Automated Semantic Fault Localization in SysML v2: A Human-in-the-Loop Framework Using Knowledge-Graph Augmented LLMs", presented at INCOSE International Symposium 2026. Dataset Structure This dataset provides two configurations: default: Contains train/validation/test splits used for fine-tuning small models. Samples exceeding 2048 tokens have been removed. full: Contains complete dataset Task Given SysML v2 code… See the full description on the dataset page: https://huggingface.co/datasets/rohhaiil/SysMLv2_Repair_with_SLMs.tabular10K<n<100K0 likes55 downloads5mo agoHugging Face24violetxi /tb21-eval-qwen35-action-only-20k-infra-repaired-c164-max32k-timeout2x qwen35-action-only-20k — Terminal-Bench 2.1 Noncanonical Terminal-Bench 2.1 evaluation of violetxi/qwen35-4b-offline-echo-action-only-20k-tacc through the served model ID qwen35-action-only-20k with Terminus-2. Noncanonical run: timeout_multiplier=2 instead of 1.0; repair concurrency=164 exceeds 30. Do not compare this score directly with canonical TB2.1 leaderboard runs. Result Recorded trials: 445 Tasks / attempts: 89 × 5 Errored trials scored as zero: 250… See the full description on the dataset page: https://huggingface.co/datasets/violetxi/tb21-eval-qwen35-action-only-20k-infra-repaired-c164-max32k-timeout2x.tabularreinforcement-learningn<1K0 likes54 downloads2mo agoHugging Face25ankithreddy /eval_repair_menuThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "so101_follower", "total_episodes": 1, "total_frames": 526, "total_tasks": 1, "total_videos": 2, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:1" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ankithreddy/eval_repair_menu.tabularroboticsn<1K0 likes52 downloads1y agoHugging Face26ankithreddy /eval_repair_rendaThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "so101_follower", "total_episodes": 1, "total_frames": 451, "total_tasks": 1, "total_videos": 2, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:1" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ankithreddy/eval_repair_renda.tabularroboticsn<1K0 likes50 downloads1y agoHugging Face27shahoismael /repair-aware-ibn Repair-Aware IBN Intent-Conflict Benchmark Labelled pairs of network intents for conflict detection and conflict resolution in Intent-Based Networking. Each conflicting pair records not only that the two intents conflict, but which predicate relation, if removed, resolves the conflict. That second label is what the dataset exists for: it makes it possible to measure whether a detector has learned what resolves a conflict rather than only what one looks like. Companion artefact… See the full description on the dataset page: https://huggingface.co/datasets/shahoismael/repair-aware-ibn.tabulartext-classificationn<1K0 likes39 downloads4d agoHugging Face28ankithreddy /eval_repair_2This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "so101_follower", "total_episodes": 1, "total_frames": 1098, "total_tasks": 1, "total_videos": 2, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:1" }, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/ankithreddy/eval_repair_2.tabularrobotics1K<n<10K0 likes38 downloads1y agoHugging Face29kasiv008 /xarm-n_pick_repairedThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "u850", "total_episodes": 48, "total_frames": 29949, "total_tasks": 1, "chunks_size": 1000, "fps": 20, "splits": { "train": "0:48" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path": "videos/{video_key}/chunk-{chunk_index:03d}/file-{file_index:03d}.mp4", "features": {… See the full description on the dataset page: https://huggingface.co/datasets/kasiv008/xarm-n_pick_repaired.tabularrobotics10K<n<100K0 likes35 downloads11mo agoHugging Face30danieldzikunuofmarvel /bibletts-asante-twi-repaired BibleTTS Asante Twi — Repaired Transcripts The Asante Twi transcripts released with BibleTTS have had the characters ɛ (U+025B) and ɔ (U+0254) stripped out. This dataset restores them. Audio is not included. This is a drop-in replacement for the .txt files that ship with the BibleTTS Asante Twi package, matched by clip ID. The problem Both are Twi vowels, and both are required by the orthography. Measured across the released Asante Twi transcripts: Character… See the full description on the dataset page: https://huggingface.co/datasets/danieldzikunuofmarvel/bibletts-asante-twi-repaired.tabularautomatic-speech-recognition10K<n<100K0 likes33 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.