HaptalAI/robotics-failure-benchmark
Haptal Robotics Failure Benchmark v1.0 The first public benchmark for robot training data annotation quality and failure detection in manipulation episodes. What this is A held-out test set of 600 robot episodes across 6 failure classes generated from real LeRobot trajectories with physics-based failure injection. The test set is fixed. Anyone can evaluate their annotation pipeline against it and get a comparable score. Why it exists No standardized… See the full description on the dataset page: https://huggingface.co/datasets/HaptalAI/robotics-failure-benchmark.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face