Terminal-Universe: Turning Agent Trajectories into Scalable Terminal Environments Paper • 2609.04148 • Published 21 days ago • 245
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses Paper • 2609.24972 • Published 3 days ago • 181
Luce: Relightable Gaussians for 3D Asset Generation Paper • 2608.23943 • Published about 1 month ago • 16
It Takes Two to Match: Co-Evolving Generative Retriever with Reinforcement Learning Paper • 2609.00638 • Published 23 days ago • 66
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 7 days ago • 175
On the Design of Qwen3.8-Next Architecture: Evaluation, Efficiency, and Training Stability Paper • 2608.30320 • Published 24 days ago • 63
view article Article NeoMME: an efficient Multimodal-native and Multilingual Encoder Hcompany • 21 days ago • 110
view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego • 21 days ago • 134