harrywang01 commited on
Commit
c88c3dc
·
verified ·
1 Parent(s): 288ad2f

CAMP RMBench policies (inference weights only)

Browse files
README.md ADDED
@@ -0,0 +1,64 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: mit
3
+ tags:
4
+ - robotics
5
+ - imitation-learning
6
+ - diffusion-policy
7
+ - memory
8
+ - rmbench
9
+ - robotwin
10
+ ---
11
+
12
+ # CAMP on RMBench
13
+
14
+ Policies for the paper **"Remember what you did: learning behavioral memory for robot manipulation"**
15
+ ([CAMP](https://robo-camp.github.io/), code: https://github.com/ucsdarclab/CAMP), trained on the
16
+ [RMBench](https://github.com/robotwin-Platform/rmbench) (RoboTwin 2.0, Aloha-AgileX) benchmark from the 50 released
17
+ demonstrations per task. Evaluated with RMBench's own protocol (`demo_clean`, seeds from 100000 validated by the
18
+ scripted expert, per-task step limits, 100 episodes).
19
+
20
+ | task | success (100 episodes) | folder |
21
+ |---|---|---|
22
+ | rearrange_blocks | 100 / 100 | `rearrange_blocks/` |
23
+ | blocks_ranking_try | 100 / 100 | `blocks_ranking_try/` |
24
+ | put_back_block | 100 / 100 | `put_back_block/` |
25
+ | battery_try | 97 / 100 | `battery_try/` |
26
+
27
+ ## Files
28
+
29
+ Each task folder holds only inference weights (no optimizer, scheduler or training bookkeeping):
30
+
31
+ - `policy.ckpt` — CAMP policy (Diffusion Policy conditioned on the compressed action memory), EMA weights and the
32
+ resolved training config.
33
+ - `memory/best_model.pt` — the Stage-1 action-memory LSTM (weights + architecture args) the policy was trained with.
34
+ - `memory/normalizer.pt` — its input normaliser.
35
+
36
+ ## Recipe (all tasks)
37
+
38
+ - Stage 1: memory LSTM pretrained on the 50 demos to reconstruct its past actions (DCT heads), head camera
39
+ 96x128 + 14-D joint state, hidden 128, action subsampling 4.
40
+ - Stage 2: Diffusion Policy (head camera 240x320, 14-D joint targets, `n_obs_steps=1`, 8-step action chunks) conditioned on the
41
+ memory through a 32-D projection; memory frozen for 400 epochs, then jointly finetuned (200 epochs; put_back_block 600).
42
+ The checkpoint reported per task is the best one over evaluated epochs.
43
+
44
+ ## Usage
45
+
46
+ ```bash
47
+ # inside the CAMP + RoboTwin evaluation image (see scripts/rmbench/eval in the CAMP repo)
48
+ python scripts/rmbench/eval/rmbench_eval.py eval --task rearrange_blocks --ckpt policy --episodes 100 \
49
+ --ckpt_root <this repo>/stage2 --stage1_root <this repo>/stage1
50
+ ```
51
+ where `stage2/<task>/checkpoints/policy.ckpt` and `stage1/<task>/{best_model.pt,normalizer.pt}` point at the files
52
+ of this repo (symlink or copy). The policy adapter is `scripts/rmbench/eval/policy_CAMP` and follows RMBench's
53
+ `get_model / eval / reset_model` interface.
54
+
55
+ ## Citation
56
+
57
+ ```bibtex
58
+ @article{wang2026rememberdidlearningbehavioral,
59
+ title = {Remember What You Did: Learning Behavioral Memory for Robot Manipulation},
60
+ author = {Wang, Kuancheng and Yeom, Hyunsoo and Cao, Yifan and Zhi, Huanyu and Shinde, Ishan and Yip, Michael C.},
61
+ journal = {arXiv preprint arXiv:2606.21188},
62
+ year = {2026}
63
+ }
64
+ ```
battery_try/memory/best_model.pt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:db86f9fc2b07b2eddfbd796cf72a0c9994b7dba69a84988ddbc5cbd4b35d3862
3
+ size 48728373
battery_try/memory/normalizer.pt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:4abe1b4094c50e858a84ed4ca784d8d8d10e565ad14e64a383c90cd5c69b6831
3
+ size 8309
battery_try/policy.ckpt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:a40d19b37bfbace58b4d784a23b6fe3ccbf9871cc89d2e8eb29e120daa37a807
3
+ size 1097423783
blocks_ranking_try/memory/best_model.pt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:7303dc7bbaab700f9a4b9c05978ab32b03592a6f7137650cc0a26eea9e655f6a
3
+ size 48780853
blocks_ranking_try/memory/normalizer.pt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:ea4fc34749c85de310882f585e68483afe43078e62accb5b2d216408dae80684
3
+ size 8309
blocks_ranking_try/policy.ckpt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:15c08f7776126459a5a70cdb157a38246af6e56ad6c113fb8496d0a9d34e0b06
3
+ size 1097476327
put_back_block/memory/best_model.pt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:2192a1ae0ac0bbc9391f3cacfe2599bc703cba4718227da32d8c6639d19d1a0f
3
+ size 48712373
put_back_block/memory/normalizer.pt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:22ded27b555e2d002b59f546beb1c585257e711156d052af1c8aa96c1b4d3105
3
+ size 8309
put_back_block/policy.ckpt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:36f2040334fbd872abd6ca064186ab36f145445946b97a6f28346fa11bb98748
3
+ size 1097407847
rearrange_blocks/memory/best_model.pt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:d6514124ba176f86a6e222d585963531b472675d3b1e92ce323dd72b26dfc0c1
3
+ size 48714293
rearrange_blocks/memory/normalizer.pt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:151e5a05b9c625cdbbffa146f2fb9a5a60f5cb3cd522cbc29f53a39635650a16
3
+ size 8309
rearrange_blocks/policy.ckpt ADDED
@@ -0,0 +1,3 @@
 
 
 
 
1
+ version https://git-lfs.github.com/spec/v1
2
+ oid sha256:555f0d878b0a0af5d0afad3ea9fc0dd180a702ede106821f6ee767daba5584fd
3
+ size 1097409767