CoolFace
Modelpublic

Spa-Bench/spa-bench-model-vla-0-epoch-12

sourceHugging Faceotherupdated 10d agoView on Hugging Face
0likes10downloads
Model Card

VLA-0 — Spa-Bench Epoch 12

This checkpoint is released as an anonymous supplementary artifact for a paper under double-blind review. It was evaluated on a physical SO-101 robot in Spa-Bench.

Model details

FieldValue
CheckpointEnd of epoch 12; 77,246 optimizer updates
InputsMiddle RGB, wrist RGB, six joint positions, and a text instruction
Action representationSix joint targets encoded as text with 1,000 bins
Action horizon8
AdaptationFull Qwen backbone; no LoRA or QLoRA
OptimizerAdamW; learning rate 5e-6; weight decay 0.01

Training-data documentation: `spa-bench-training-teleoperation-1200`.

Physical evaluation

0/72 familiar physical rollouts succeeded. Evaluation stopped before the withheld-composition protocol, so this is a partial result. These are physical rollout results, not simulation metrics.

Rollouts: `spa-bench-eval-rollouts-vla-0-partial`.

Limitations and safety

Training used a motion-trimmed projection of the 1,200-episode release. Deployment requires the VLA-0 inference stack; this is not a drop-in LeRobot policy. Robot policies can move hardware unexpectedly. Use conservative limits, an accessible emergency stop, a clear workspace, and direct supervision. Do not deploy unattended or in safety-critical settings.

Double-blind release note

Author, institution, source-repository, and archival citation details are intentionally omitted during review. They will be restored in the archival release.