PreviewDiff: Multimodal Critic-Guided Search over Diffusion Latents Paper • 2609.36199 • Published 3 days ago • 4
TabFM-Auto: Self-Evolving Pipelines for Tabular Foundation Models Paper • 2609.37989 • Published 3 days ago • 10
Disaggregated Quantization: Specializing LLM Prefill and Decode Paper • 2609.26333 • Published 10 days ago • 90
Duplex-MPE: Benchmarking Multi-Party Interaction in Full-Duplex Dialogue Paper • 2609.31948 • Published 7 days ago • 86
Self-Evolving Coding Agents: From Digital Programs to Physical-World Intelligence Paper • 2609.35432 • Published 4 days ago • 102
Beyond Teacher Assignment: Domain-Normalized Multi-Teacher On-Policy Distillation Paper • 2609.35347 • Published 4 days ago • 152
Post-Training Leaves Behavioral Shadows on Unrelated Decisions Paper • 2609.29233 • Published 8 days ago • 267
Selecting Diverse SFT Traces Improves Post-RL Generalization Paper • 2609.33780 • Published 5 days ago • 32
VisionHOPE: Visual Backbones as Self-Modifying Learning Systems Paper • 2609.33325 • Published 5 days ago • 304
The Cartesian Shortcut: Re-evaluate Vision Reasoning in Polar Coordinate Space Paper • 2605.09883 • Published May 29 • 1
Groupwise Agentic Grading and Advantage Redistribution for Code Agent RL Paper • 2609.32577 • Published 6 days ago • 115
QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents Paper • 2609.33848 • Published 5 days ago • 38
Draft-KV: Learning Useful Latent Communication Between Language Models Paper • 2609.34754 • Published 4 days ago • 8
MassAlloc Attention: Let Attention Allocate Its Own Compute Paper • 2609.32712 • Published 6 days ago • 68