The Mirage of Optimizing Training Policies: Monotonic Inference Policies as the Real Objective for LLM Reinforcement Learning Paper • 2606.29526 • Published Jun 28 • 170
OmniDirector: General Multi-Shot Camera Cloning without Cross-Paired Data Paper • 2606.13432 • Published Jun 11 • 113
lafrancef/ssd-math-v1p1-qwen3-5-4b-scaled-4k-20260604T073132Z-train-v1p1-checkpoint-13-1221b7f8 5B • Updated Jun 4 • 2 • 1
StreamChar: Long-Horizon Streaming Character Audio-Video Generation with Decoupled Orchestration Paper • 2605.25659 • Published May 25 • 17