PDE-JEPA: Predictive Representation Learning of Latent Dynamics Modeling for Parametric PDEs Paper • 2609.34715 • Published 9 days ago • 52
PhyEdit: Towards Real-World Object Manipulation via Physically-Grounded Image Editing Paper • 2604.07230 • Published Apr 9 • 18
From 2D Grids to 1D Tokens: Reforming Shared Representations for Multimodal Image Fusion Paper • 2606.12303 • Published Jun 10 • 33
VIA-SD: Verification via Intra-Model Routing for Speculative Decoding Paper • 2606.12243 • Published Jun 10 • 37
Comprehensive Benchmarking of Long-Form Speech Generation in Diverse Scenarios Paper • 2605.28618 • Published May 27 • 30
Towards Streaming Synchronized Spatial Audio Generation via Autoregressive Diffusion Transformer Paper • 2605.30940 • Published May 29 • 34
SwanVoice: Expressive Long-Form Zero-Shot Speech Synthesis for Both Monologue and Dialogue Paper • 2605.30993 • Published May 29 • 55
RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details Paper • 2604.06870 • Published Apr 8 • 44
Dynamic Experts Search: Enhancing Reasoning in Mixture-of-Experts LLMs at Test Time Paper • 2509.22572 • Published Sep 26, 2025 • 13
Dynamic Experts Search: Enhancing Reasoning in Mixture-of-Experts LLMs at Test Time Paper • 2509.22572 • Published Sep 26, 2025 • 13
view article Article From PyTorch DDP to Accelerate to Trainer, mastery of distributed training with ease muellerzr • Oct 21, 2022 • 44