Persona Dosing: Calibrated Activation Steering for Graded Trait Control Paper • 2609.36388 • Published 8 days ago • 51
Occamy-1.0: Open Pareto-frontier 35B Intelligence for Co-work Paper • 2609.11977 • Published Sep 4 • 117
Difficulty-Adaptive Tree-Structured Policy Optimization for Expanding Reasoning Coverage in RLVR Paper • 2609.08650 • Published 28 days ago • 11
StatePlay: State-Aware Game World Models for Mechanics-Consistent Generation Paper • 2607.26754 • Published Jul 29 • 19