MetroLLM-Bench: Evaluating Language Models as Transit Kiosk Runtimes Paper • 2609.10016 • Published 22 days ago • 29
Mask Forcing: Improving Autoregressive Video Diffusion Distillation via Dual-Noise Masking Rollout Paper • 2609.09123 • Published 23 days ago • 55
Reason Through the Latent! Making Latent Visual Reasoning Necessary Paper • 2609.06746 • Published 25 days ago • 36
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published Aug 27 • 155
Training Chemical Plausibility-Aware Large Language Models for Single-Step Retrosynthesis Paper • 2608.18940 • Published Aug 19 • 35
LycheeMemory V2: Efficient Long-Term Memory for LLM Agents via Semantic Segment-Level Consolidation Paper • 2608.12990 • Published Aug 13 • 14
SPIEval: Evaluating Large Language Models as Mobile Assistants over Scattered Personal Information Paper • 2608.10692 • Published Aug 11 • 13
Activity Frames: Deterministic Screen-Activity Compilation for Agent Memory and Replay Paper • 2608.05784 • Published Aug 6 • 25
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement Learning Paper • 2608.05987 • Published Aug 6 • 104
PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning Paper • 2608.01837 • Published Aug 3 • 40