view article Article Kimi K3 Model Overview: 2.8T Parameters, MXFP4 Quantization, and What the Open Weights Mean for the Community ResterChed • 11 days ago • 170
RLCSD: Reinforcement Learning with Contrastive On-Policy Self-Distillation Paper • 2606.11709 • Published Jun 10 • 1 • 1
RLCSD: Reinforcement Learning with Contrastive On-Policy Self-Distillation Paper • 2606.11709 • Published Jun 10 • 1
Learning from Language Feedback via Variational Policy Distillation Paper • 2605.15113 • Published May 18 • 13
Dr-DCI: Scaling Direct Corpus Interaction via Dynamic Workspace Expansion Paper • 2606.14885 • Published Jun 12 • 11
Rethinking Continual Experience Internalization for Self-Evolving LLM Agents Paper • 2606.04703 • Published Jun 3 • 26
ResearchClawBench: A Benchmark for End-to-End Autonomous Scientific Research Paper • 2606.07591 • Published May 28 • 102
K-BrowseComp: A Web Browsing Agent Benchmark Grounded in Korean Contexts Paper • 2606.02404 • Published Jun 1 • 59
Agentic Context Engineering: Evolving Contexts for Self-Improving Language Models Paper • 2510.04618 • Published Oct 6, 2025 • 134