CrossBFM: Distilling a Shared Latent Behavior Space Across Humanoid Embodiments Paper • 2609.38087 • Published 6 days ago • 24
Structured Residual Connectivity Matters for Diffusion Transformers Paper • 2609.33203 • Published 8 days ago • 17
H3-World: Turning Language Understanding into World Control Paper • 2609.01560 • Published Sep 1 • 53
BadWAM: When World-Action Models Dream Right but Act Wrong Paper • 2607.15207 • Published Jul 16 • 45
You Don't Need Strong Assumptions: Visual Representation Learning via Temporal Differences Paper • 2606.15956 • Published Jun 14 • 13
SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning Paper • 2606.13673 • Published Jun 11 • 115
Mix-Quant: Quantized Prefilling, Precise Decoding for Agentic LLMs Paper • 2605.20315 • Published May 19 • 26
On-Policy Self-Evolution via Failure Trajectories for Agentic Safety Alignment Paper • 2605.11882 • Published May 12 • 15
Visual Generation in the New Era: An Evolution from Atomic Mapping to Agentic World Modeling Paper • 2604.28185 • Published Apr 30 • 90
Vision-Language-Action Safety: Threats, Challenges, Evaluations, and Mechanisms Paper • 2604.23775 • Published Apr 26 • 45
Gated Condition Injection without Multimodal Attention: Towards Controllable Linear-Attention Transformers Paper • 2603.27666 • Published Mar 29 • 16