Simulated manipulation โ€” trained policy adapters and probes

Model weights trained in simulation for robot manipulation and 2-D navigation experiments.

All weights here were trained by the uploader. Pretrained backbones are not redistributed; the training code downloads them from their own sources under their own licences.

Contents

Path What it is
star/pi0_anchor_lora.pt LoRA + projector adapter, two-cube tabletop task
star/pi0_noplan_lora.pt the same adapter trained with the conditioning dropped
star/pi0_richplan_lora_rp.pt adapter trained on a real-prompt demonstration mixture
star/pi0_richplan_lora_dag.pt the above, refined with a privileged-expert DAgger stage
star/dino_probe_spoon448.pt spatial-softmax probe over frozen visual features
star/p2_unified_ckpt.pt a single generator trained jointly over four task families
star/pi0_general_reasoner.pt joint generator over four demonstration sets
star/robocasa_steer_lora.pt adapter for a kitchen-simulator pick-and-place task
star/adapters_{cf,v2,mixture}/ adapters for a bowl-selection benchmark, several data mixtures
star/heads_cf_fixed/ small conditional generators, multiple seeds
star/dual_channel_head_s0.pt a two-channel conditioning head
maze/saved_models/ 18 multi-path maze navigation runs, several architectures

Verifying a download

A manifest pinning every file by content hash accompanies the training code.

Caveat

The exact training invocation for pi0_richplan_lora_rp.pt was not preserved. Use these weights rather than trying to reconstruct it.

Downloads last month

-

Downloads are not tracked for this model. How to track
Video Preview
loading