File size: 1,043 Bytes
be88555 b9f5f20 be88555 b9f5f20 be88555 b9f5f20 be88555 4d6a0cd be88555 b9f5f20 be88555 b9f5f20 be88555 b9f5f20 be88555 | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 | ---
license: other
library_name: pytorch
tags:
- robotics
- reinforcement-learning
- rainbow-dqn
- speed-control
- metarobobench
---
# MetaRoboBench Speed-Control Heads
Offboarding archive of Rainbow DQN, CQL, and candidate-value heads trained for chunk-level Fast/Normal/Slow selection on top of Pi05 checkpoint `21650`.
The primary model is `rainbow/shared_all50_final/models/shared_all50.pt`:
- 50-task shared Rainbow head
- 100,491 decisions and 99,962 updates
- 1,941/2,500 successes = 77.64%
- Fast/Normal/Slow chunk shares = 38.6/32.3/29.1%
- Mean speed switches = about 4.23 per episode
The accelerated shared checkpoint is a distinct earlier checkpoint, not a duplicate. Per-task Rainbow, CQL, and candidate-value models are retained as baselines and negative-result artifacts.
Code reference:
- Repository: https://github.com/MMMM3202/MetaRoboBench
- Branch: `codex/vlm-rainbow-shared-20260821`
- Commit: `5b36567f0c6174ad82a3d4c669134f7b28acc945`
See `MODEL_MANIFEST.md` for the full inventory and result caveats.
|