training / docs

Commit History

log: retract v7-v12 KL conclusion; metric artifact + logp_floor fix
9d5bc84
verified

cazyundee commited on

log: repetition fix necessary but not sufficient
7e47006
verified

cazyundee commited on

log: v7-v12 stability campaign, 3 bugs fixed, no strength gain
f9afc8e
verified

cazyundee commited on

log: repetition-feature bug, greedy baseline, learn bench
f14f327
verified

cazyundee commited on

fix eval-on-resume (silently skipped); add learn bench; v4 analysis
24c12cb
verified

cazyundee commited on

add learn-phase thread benchmark
3f4a221
verified

cazyundee commited on

revert logp_old change (ablation-falsified); keep only_player + kl estimator fixes
5d627db
verified

cazyundee commited on

draw-penalty: remove guaranteed-0 repetition optimum
35dde2e
verified

cazyundee commited on

parallel self-play (--workers), conversion experiments, whitelist
b97f43c
verified

cazyundee commited on

whitelist analysis scripts; fix pid-recycling in launch/job_table
2f4aa7f
verified

cazyundee commited on

research log: env correction + thread bench
83d8bce
verified

cazyundee commited on

tinychess: full research artifact - self-play RL, ablations, compute scaling, faithfulness, precision, plasticity
227f251
verified

cazyundee commited on