Commit History

Invalidate WebShop SPrPO training seeds 1 and 2 affected by request-seed bug
37541c4
verified

PengxinWang commited on

Use explicit download and catalog URLs in model card
d550f14
verified

PengxinWang commited on

Standardize retained training artifacts at step200 and organize ablations
136f147
verified

PengxinWang commited on

Add reproducible training overhead table for main experiments
fe9e3c9
verified

PengxinWang commited on

Reduce edge margins while preserving the original motivation canvas size
aa65eff
verified

PengxinWang commited on

Center clean-rate glyphs and tighten motivation whitespace
104e478
verified

PengxinWang commited on

Align motivation condition names exactly with bar centers
d61caac
verified

PengxinWang commited on

Equalize motivation panel widths and reduce condition typography
03dcf6c
verified

PengxinWang commited on

Fit relative-drop labels between bars and clean reference lines
e404e3e
verified

PengxinWang commited on

Place ALFWorld perturbation labels to the right of bars
898f770
verified

PengxinWang commited on

Move motivation condition labels closer to the plots
a2cc857
verified

PengxinWang commited on

Tighten motivation layout and remove explanatory footer
58ab98b
verified

PengxinWang commited on

Restyle introduction figure with relative-drop labels and method-inspired pastel colors
8b708a8
verified

PengxinWang commited on

Redesign half-width motivation figure with matched Gaussian 0.20 and dropout 0.05
7cfe5a4
verified

PengxinWang commited on

Update ALFWorld SAM three-seed main results
42ab6dc
verified

PengxinWang commited on

Add ALFWorld 1.5B single-layer FFN Channel-SAM step100/step200 (rho=0.005, cap=0.05) (#6)
f667567

PengxinWang Davidk1 commited on

Add qwen2_5_1_5_stable_gaussian_noise_v11_step_200_seed0 step200 LoRA adapter and run config (#5)
6a3865b

PengxinWang Davidk1 commited on

Add qwen2_5_1_5_stable_gaussian_noise_v9_step_200_seed0 step200 LoRA adapter and run config (#4)
3844509

PengxinWang Davidk1 commited on

Tighten main-result panel spacing in LaTeX and preview
de9a061
verified

PengxinWang commited on

Publish step-200 main-result figures, plotting data, and LaTeX layout
f631c23
verified

PengxinWang commited on

Rename shared analysis group to perturbation_effect_across_stage
a606550
verified

PengxinWang commited on

Publish WebShop perturbation-geometry analysis figure and data
f389664
verified

PengxinWang commited on

Normalize step100 LoRA adapter configs and add step100 to manifests
90e03a6
verified

PengxinWang commited on

Add results/alfworld/robust_training/qwen2_5_7_gaussian_noise_step_200_seed0/checkpoints/global_step_100/actor/lora_adapter
b5d10c9
verified

PengxinWang commited on

Add results/alfworld/training/qwen2_5_7_step_200_seed0/checkpoints/global_step_100/actor/lora_adapter
b571fa8
verified

PengxinWang commited on

Add results/webshop/training/qwen2_5_7_step_200_seed0/checkpoints/global_step_100/actor
b5a08cc
verified

PengxinWang commited on

Add WebShop 7B single-layer FFN Channel-SAM step200 (rho=0.005, cap=0.05) (#2)
56fdbbc

PengxinWang Davidk1 commited on

Publish ALFWorld Qwen2.5-7B step200 clean-vs-gaussian evaluation results
15aff1e
verified

PengxinWang commited on

Publish qwen2_5_7_stable_gaussian_noise_step_200_seed0 step200 LoRA adapter for collaborator ALFWorld 7B handoff
230907a
verified

PengxinWang commited on

Publish qwen2_5_7_stable_gaussian_noise_step_200_seed0 step200 LoRA adapter for collaborator ALFWorld 7B handoff
b30a182
verified

PengxinWang commited on

Publish qwen2_5_7_stable_gaussian_noise_step_200_seed0 step200 LoRA adapter for collaborator ALFWorld 7B handoff
e87bc1e
verified

PengxinWang commited on

Publish qwen2_5_7_stable_gaussian_noise_step_200_seed0 step200 LoRA adapter for collaborator ALFWorld 7B handoff
da23783
verified

PengxinWang commited on

Remove redundant motivation README
707f7f1
verified

PengxinWang commited on

Remove redundant reward surface README
56454c7
verified

PengxinWang commited on

Organize reward surface and motivation figures with reproducible CSV inputs
bbaca61
verified

PengxinWang commited on

Publish complete WebShop step200 three-seed evaluation results
230ae6e
verified

PengxinWang commited on

Remove superseded SAM-named evaluation figure
f322c84
verified

PengxinWang commited on

Publish complete WebShop step200 three-seed evaluation results
5ad165d
verified

PengxinWang commited on

Use fresh cap 0.05 SAM as official step200 baseline
b756470
verified

PengxinWang commited on

Add WebShop single-layer FFN Channel-SAM step200 (rho=0.005, cap=0.05) (#1)
aa7d451

PengxinWang Davidk1 commited on

Remove redundant evaluation package readme and manifest
c8c5c49
verified

PengxinWang commited on

Update curated evaluation figures and reproducible data
adbb800
verified

PengxinWang commited on

Use Vanilla, Plain Gaussian and Stable Gaussian algorithm labels
057000e
verified

PengxinWang commited on

Publish WebShop 1.5B Stable Gaussian step200 main result
9d5f63d
verified

PengxinWang commited on

Simplify model card around step200 baseline and model download table
3cdd8a6
verified

PengxinWang commited on

Document unified 200-step training default for 1.5B and 7B
b562238
verified

PengxinWang commited on

Publish four completed Qwen2.5-7B step-200 baseline adapters
2f9d4a3
verified

PengxinWang commited on

Document four finalized step-300 evaluation baselines
7903637
verified

PengxinWang commited on

Publish evaluation-only step-300 LoRA baselines; remove training archive
903ca99
verified

PengxinWang commited on

Archive baseline: alfworld/training/qwen2_5_1_5_step_300_seed0
5dfb70c
verified

PengxinWang commited on