fix 2nd train/inference mismatch: last-move feature lost in FEN round-trip c8eb6ce verified cazyundee commited on 15 days ago
fix 2nd train/inference mismatch: last-move feature lost in FEN round-trip c682388 verified cazyundee commited on 15 days ago
fix: repetition feature lost in FEN round-trip (train/inference mismatch); greedy-health exp 79997ea verified cazyundee commited on 15 days ago
fix: repetition feature lost in FEN round-trip (train/inference mismatch); greedy-health exp c94f46f verified cazyundee commited on 15 days ago
fix: repetition feature lost in FEN round-trip (train/inference mismatch); greedy-health exp 24c6c78 verified cazyundee commited on 15 days ago
revert logp_old change (ablation-falsified); keep only_player + kl estimator fixes 5d627db verified cazyundee commited on 15 days ago
adjudicate claimed repetition/50-move draws on material 0f3f386 verified cazyundee commited on 15 days ago
draw-penalty: remove guaranteed-0 repetition optimum 35dde2e verified cazyundee commited on 15 days ago
PPO correctness: logp_old from behaviour dist (temp+dirichlet), not raw policy 09929ef verified cazyundee commited on 15 days ago
PPO fix: exclude opponent positions from buffer + staleness bound 3f1e7d3 verified cazyundee commited on 15 days ago
spawn-based persistent self-play pool (fork/libgomp deadlock fix) fb6e523 verified cazyundee commited on 16 days ago
parallel self-play (--workers), conversion experiments, whitelist b97f43c verified cazyundee commited on 16 days ago
tinychess: full research artifact - self-play RL, ablations, compute scaling, faithfulness, precision, plasticity 227f251 verified cazyundee commited on 17 days ago
tinychess: self-play research substrate (phase 1-4) 3495881 verified cazyundee commited on 17 days ago