Myungkyu/RoboDojo-taco-gemini
Viewer • Updated • 3.2k • 142
Partial run of the RoboDojo TACO low-level policy: RLWRLD/RLDX-1-PT fine-tuned on Myungkyu/RoboDojo-taco-gemini
(8 long-horizon bimanual tasks, 100 demonstrations each, dense subtask labels from the task-specific context). The training was preempted at
step 30000 of the planned 60000 (2026-09-12); the checkpoints are kept for resumption and analysis.
checkpoint-30000/: complete checkpoint — weights + global_step30000/ DeepSpeed optimizer states + RNG states + latest (resume with the RLDX-1 trainer)checkpoint-20000/, checkpoint-25000/: weights only (config, index, safetensors shards, experiment_cfg/, processor/)SHA256SUMS of its weight filesConfigs reference the base backbone / tokenizer by hub id or by the training site's local path — point them at your local copies before loading.