Myungkyu/rldx_1_rmbench_taco_wodemo_gemini_b128_60k
0136
rldx1rmbenchtacowodemogeminib128_60k
Low-level policy for RMBench (9 simulated tabletop tasks), fine-tuned from RLWRLD/RLDX-1-PT on `Myungkyu/RMBench-taco-wodemo-gemini` — demonstrations with dense subtask labels from the task-specific context (offline annotation).
- Architecture: RLDX-1-PT (video length 4, three camera views + one keyframe slot)
- Optimizer batch 128, 60000 steps, final checkpoint
- Inputs: head + left/right wrist images, proprioception, the current subtask text; the keyframe slot takes a retrieved past frame when the label calls for one
Configs reference the base backbone / tokenizer by hub id or by the training site's local path — point them at your local copies before loading.
