Sample Count Is Not Enough: Candidate-Generation Strategy Shapes the Energy and Performance of LLM Test-Time Scaling Paper • 2609.19499 • Published 9 days ago • 37
Studying Image Tokenizers as Visual Languages in Unified Multimodal Models Paper • 2609.09143 • Published 17 days ago • 29
CosmoH2G: A Hand-to-Gripper Transfer Dataset and Baseline Method for Object Manipulation with Complex Spatial Movements Paper • 2609.07498 • Published 18 days ago • 34
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 22 days ago • 186
ForgeWM: Progressive Causal Training for Few-Step Action-Conditioned Video World Models Paper • 2608.14022 • Published Aug 14 • 24
Training Leaves Traces: Centered Residual Signatures for Language Model Lineage Verification Paper • 2608.14929 • Published Aug 14 • 18
Massive Activations in Hybrid Linear Attention Large Language Models: Pre-Attention Spikes and Inter-Spike Plateaus Paper • 2608.12149 • Published Aug 12 • 30
CalibForge: Adversarial Solver Calibration for Scaling Learnable Terminal Tasks Paper • 2608.06352 • Published Aug 6 • 24
On-Policy Delta Distillation for Multilingual Math Reasoning Paper • 2608.05802 • Published Aug 6 • 33
ContinualSkillBench: Can LLM Agents Truly Evolve Their Capabilities? Paper • 2608.03874 • Published Aug 4 • 15
UniWorld-Design: From Pixel Generation to Layer-Native Design Paper • 2608.03971 • Published Aug 4 • 25