Learning from Teacher Continuations at Student States Paper • 2609.36246 • Published 12 days ago • 40
Selecting Diverse SFT Traces Improves Post-RL Generalization Paper • 2609.33780 • Published 13 days ago • 38
Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning Paper • 2609.03430 • Published Sep 3 • 188
Modular Cognitive Architecture Emerges in Large Language Models Paper • 2608.13567 • Published Jun 27 • 16
LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers Paper • 2608.06867 • Published Aug 7 • 114
Lie Detection Collection Did you lie? Evaluating Lie Detectors across Model Scale and Belief-Verified Model Organisms • 9 items • Updated Jun 10 • 6
TELEClass: Taxonomy Enrichment and LLM-Enhanced Hierarchical Text Classification with Minimal Supervision Paper • 2403.00165 • Published Feb 29, 2024 • 1
Rethinking the Reranker: Boundary-Aware Evidence Selection for Robust Retrieval-Augmented Generation Paper • 2602.03689 • Published Feb 3
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Paper • 2606.02373 • Published Jun 1 • 59
Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses Paper • 2606.02373 • Published Jun 1 • 59
TELEClass: Taxonomy Enrichment and LLM-Enhanced Hierarchical Text Classification with Minimal Supervision Paper • 2403.00165 • Published Feb 29, 2024 • 1
Learning to Predict Future-Aligned Research Proposals with Language Models Paper • 2603.27146 • Published Apr 6 • 6
Retrieval is Cheap, Show Me the Code: Executable Multi-Hop Reasoning for Retrieval-Augmented Generation Paper • 2605.12975 • Published May 13 • 9
Useful Memories Become Faulty When Continuously Updated by LLMs Paper • 2605.12978 • Published May 13 • 19
Steer2Adapt: Dynamically Composing Steering Vectors Elicits Efficient Adaptation of LLMs Paper • 2602.07276 • Published Feb 7 • 11
Steer2Adapt: Dynamically Composing Steering Vectors Elicits Efficient Adaptation of LLMs Paper • 2602.07276 • Published Feb 7 • 11
Weak-Driven Learning: How Weak Agents make Strong Agents Stronger Paper • 2602.08222 • Published Feb 9 • 183