datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
gr1_arms_waist-CuttingboardToPangr1_arms_waist-CuttingboardToCardboardBoxgr1_arms_waist-PlaceMilkToMicrowavegr1_arms_waist-WineToCabinetgr1_arms_waist-TrayToTieredShelfgr1_arms_waist-TrayToPlateus-airport-wait-times
US Airport Security & Immigration Wait Times
Minute-resolution TSA security checkpoint wait times for 30 US airports, plus
hourly CBP immigration hall wait times for arriving international passengers.
Airports publish their current wait time and then overwrite it. Nobody keeps the
history. This dataset is that history: a continuous archive collected by polling
each airport's public feed roughly once a minute.
Collection began 1 April 2026 with the New York, Philadelphia and… See the full description on the dataset page: https://huggingface.co/datasets/digitalhen/us-airport-wait-times.gr1_arms_waist-TrayToPotgr1_arms_waist-PlateToCardboardBoxgr1_arms_waist-PlateToPangr1_arms_waist-CupToDrawergr1_arms_waist-PlaceBottleToCabinetgr1_arms_waist-PlateToBowlgr1_arms_waist-PotatoToMicrowaveswerl-tmax-15k-solvable-gpt-5-6-terra
swerl-tmax-15k hardened, post-validation-filter (dataset 3 of 3)
Which tasks in hamishivi/swerl-tmax-15k can a strong model actually solve? Every
task was attempted twice as a full agentic episode — real sandbox, real bash,
real verifier — and a task is verified when at least one attempt earned reward.
The last of three artifacts that exist to be compared by task_id:
original — hamishivi/swerl-tmax-15k, unchanged — 14,601 tasks
hardened, pre-validation-filter —… See the full description on the dataset page: https://huggingface.co/datasets/wAI-org/swerl-tmax-15k-solvable-gpt-5-6-terra.gr1_arms_waist-PlacematToTieredShelfgr1_arms_waist-PlateToPlateswerl-tmax-15k-hardened-prefilter
swerl-tmax-15k hardened, pre-validation-filter (dataset 2 of 3)
The middle artifact of three, which exist to be compared against each other by
task_id:
original — hamishivi/swerl-tmax-15k, unchanged.
hardened, pre-validation-filter — this dataset.
hardened, post-validation-filter — wAI-org/swerl-tmax-15k-solvable-gpt-5-6-terra, a strict subset of this one — 7,015 tasks.
What is in it
Tasks where a patch was actually applied, plus tasks originally labelled
CLEAN… See the full description on the dataset page: https://huggingface.co/datasets/wAI-org/swerl-tmax-15k-hardened-prefilter.gr1_arms_waist-CuttingboardToTieredBasketSSPO-data
SSPO
Beyond Outcome Rewards: Step-Level Self-Distilled Policy Optimization for Deep Search Agents
📄 arXiv •
💻 Code •
🤗 Dataset
🌟Overview
Deep search agents operate over trajectories spanning dozens of information-seeking steps, but standard reinforcement learning provides only a single outcome reward for the entire trajectory. This sparse signal makes it difficult to determine which intermediate reasoning and tool-use actions should be reinforced… See the full description on the dataset page: https://huggingface.co/datasets/WaitHZ/SSPO-data.g1-inspire-red-ball-success-waist
success-only filtered copy of g1-inspire-red-ball
76/85 episodes kept (dropped failed: [2, 4, 34, 43, 51, 62, 66, 75, 80]); re-indexed contiguously; numeric stats recomputed.
license: other
task_categories:
- robotics
tags:
- LeRobot
- GR00T
- GR00T-N1.7
- humanoid
- unitree-g1
- inspire-hand
- tactile
- manipulation
pretty_name: G1 + Inspire "place the red ball in the box" (GR00T N1.7, G1_INSPIRE)
G1 + Inspire — "place the red ball in the box"… See the full description on the dataset page: https://huggingface.co/datasets/birbirll/g1-inspire-red-ball-success-waist.gr1_arms_waist-PlacematToBasketwaiting_pick_and_place2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"arm_shoulder_pan.pos",
"arm_shoulder_lift.pos",
"arm_elbow_flex.pos",
"arm_wrist_flex.pos",
"arm_wrist_roll.pos",
"arm_gripper.pos",
"x.vel"… See the full description on the dataset page: https://huggingface.co/datasets/youcutmebadimstillwaitingforyou/waiting_pick_and_place2.legal-privilege-log-document-basis-waiver-risk-v0.1What this dataset does
You receive
doc description
date
author
recipients
privilege basis
redaction choice
context
waiver flags
You decide
coherent
or
incoherent
Daily use
privilege log QC
waiver risk detection
disclosure challenge prep
wait_then_stuffed_animal_pick_placeThis dataset was created using my fork of LeRobot.
Joint calibration
Joint calibration for the featured SO-101 is on Github
Dataset Structure
meta/info.json:
spain-nhs-waiting-lists-2025-12
ES·pera — Spain NHS Waiting Lists, December 2025
This dataset contains the December 2025 interannual extract published by ES·pera from Spain's official SISLE-SNS waiting-list data. It covers national and autonomous-community figures, selected first-consultation specialties, surgical specialties and procedures, with corresponding December 2024 comparison fields where published.
ES·pera is an independent Spanish public-data project that integrates and normalises official public… See the full description on the dataset page: https://huggingface.co/datasets/esperaorg/spain-nhs-waiting-lists-2025-12.gr1_arms_waist-CuttingboardToBasketfinal_arabic_summariesWAIC_20260624_175732This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"shoulder_pan.pos",
"shoulder_lift.pos",
"elbow_flex.pos",
"wrist_flex.pos",
"wrist_roll.pos",
"gripper.pos"
],
"shape": [
6… See the full description on the dataset page: https://huggingface.co/datasets/dragon-95/WAIC_20260624_175732.GeneralThought-195K-pruned-keep-0.5-wait
