Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning Paper • 2607.18722 • Published 12 days ago • 35
view article Article Welcome Inkling by Thinking Machines +3 burtenshaw, merve, pcuenq, ariG23498, andito • 18 days ago • 145
Harness Handbook: Making Evolving Agent Harnesses Readable,Navigable, and Editable Paper • 2607.13285 • Published 19 days ago • 231
Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling Paper • 2607.02980 • Published about 1 month ago • 83
Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling Paper • 2607.02980 • Published about 1 month ago • 83
Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling Paper • 2607.02980 • Published about 1 month ago • 83
view article Article A failed experiment: Infini-Attention, and why we should keep trying? +1 neuralink, lvwerra, thomwolf • Aug 14, 2024 • 76
BiFormer: Vision Transformer with Bi-Level Routing Attention Paper • 2303.08810 • Published Mar 15, 2023 • 1
RelayAttention for Efficient Large Language Model Serving with Long System Prompts Paper • 2402.14808 • Published Feb 22, 2024