RecMem: Recurrence-based Memory Consolidation for Efficient and Effective Long-Running LLM Agents Paper • 2605.16045 • Published May 15 • 1
view article Article MLA: Redefining KV-Cache Through Low-Rank Projections and On-Demand Decompression NormalUhr • Feb 4, 2025 • 24
XSkill: Continual Learning from Experience and Skills in Multimodal Agents Paper • 2603.12056 • Published Jul 1 • 34
view article Article The Engineering Handbook for GRPO + LoRA with Verl: Training Qwen2.5 on Multi-GPU Weyaxi • Jan 2 • 24
Video Reality Test: Can AI-Generated ASMR Videos fool VLMs and Humans? Paper • 2512.13281 • Published Dec 15, 2025 • 65
RetrievalAttention: Accelerating Long-Context LLM Inference via Vector Retrieval Paper • 2409.10516 • Published Sep 16, 2024 • 43