view article Article Your Inference Server is Secretly a Learner: Reef Infrastructure for Continual Self-Improving Agents quao627 • 18 days ago • 15
Kalman Delta Networks: Uncertainty-aware Associative Memory Paper • 2609.07816 • Published 27 days ago • 29
Make Each Token Count: Towards Improving Long-Context Performance with KV Cache Eviction Paper • 2605.09649 • Published May 10 • 13
Meta Llama 3 Collection This collection hosts the transformers and original repos of the Meta Llama 3 and Llama Guard 2 releases • 5 items • Updated Dec 6, 2024 • 1.01k
MM-Eureka: Exploring Visual Aha Moment with Rule-based Large-scale Reinforcement Learning Paper • 2503.07365 • Published Mar 10, 2025 • 61
Taipan: Efficient and Expressive State Space Language Models with Selective Attention Paper • 2410.18572 • Published Oct 24, 2024 • 18
Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation Paper • 2410.05363 • Published Oct 7, 2024 • 45