Grouped Query Experts: Mixture-of-Experts on GQA Self-Attention Paper • 2606.20945 • Published Jun 18 • 80
ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing? Paper • 2606.19531 • Published Jun 17 • 24
JanusMesh: Fast and Zero-Shot 3D Visual Illusion Generation via Cross-Space Denoising Paper • 2606.20563 • Published Jun 18 • 20
Web Retrieval-Aware Chunking (W-RAC) for Efficient and Cost-Effective Retrieval-Augmented Generation Systems Paper • 2604.04936 • Published Jan 8 • 26
UniFork: Exploring Modality Alignment for Unified Multimodal Understanding and Generation Paper • 2506.17202 • Published Jun 20, 2025 • 10
Hunyuan3D 2.1: From Images to High-Fidelity 3D Assets with Production-Ready PBR Material Paper • 2506.15442 • Published Jun 18, 2025 • 19
Reranking-based Generation for Unbiased Perspective Summarization Paper • 2506.15925 • Published Jun 19, 2025 • 5
ZClip: Adaptive Spike Mitigation for LLM Pre-Training Paper • 2504.02507 • Published Apr 3, 2025 • 90
Vision-Guided Chunking Is All You Need: Enhancing RAG with Multimodal Document Understanding Paper • 2506.16035 • Published Jun 19, 2025 • 89