DataMagic: Authoring Data Videos through Declarative Multi-Agent Orchestration Paper • 2609.33403 • Published 7 days ago • 10
Do Audio LLMs Listen Before They Act? Diagnosing Acoustic-Context Gating in Voice Agents Paper • 2609.32536 • Published 8 days ago • 12
PixelDense: Dense Prediction as Representation Alignment for Pixel Diffusion Paper • 2610.00483 • Published 4 days ago • 16
DiffNR: Diffusion-Enhanced Neural Representation Optimization for Sparse-View 3D Tomographic Reconstruction Paper • 2604.21518 • Published Apr 23 • 27
DVD: Deterministic Video Depth Estimation with Generative Priors Paper • 2603.12250 • Published Mar 12 • 28
MASQuant: Modality-Aware Smoothing Quantization for Multimodal Large Language Models Paper • 2603.04800 • Published Mar 5 • 25
FinToolBench: Evaluating LLM Agents for Real-World Financial Tool Use Paper • 2603.08262 • Published Mar 9 • 42
Thinking in Uncertainty: Mitigating Hallucinations in MLRMs with Latent Entropy-Aware Decoding Paper • 2603.13366 • Published Mar 9 • 95