Org-Agent: Beyond Personal Assistants Towards Organizational Agents Paper • 2609.34392 • Published 11 days ago • 7
When Does Dense Retrieval Need Asymmetric Geometry? A Bias-Variance Theory of Shared and Dual Projections Paper • 2609.32488 • Published 13 days ago • 27
Routing Drift Alone Does Not Diagnose Failure in Merged MoE LLMs Paper • 2609.32821 • Published 13 days ago • 5
RelightFormer: Feed-forward Generative Transformer for Multiview Object Relighting Paper • 2609.07414 • Published Sep 7 • 11
$\boldsymbol{f}$-OPD: Stabilizing Long-Horizon On-Policy Distillation with Freshness-Aware Control Paper • 2605.17862 • Published May 18 • 1
ReLaX: Reasoning with Latent Exploration for Large Reasoning Models Paper • 2512.07558 • Published Dec 8, 2025 • 1
Persistent Recursive Worlds Enable Autonomous Software Evolution Paper • 2608.10450 • Published Aug 12 • 8
MMOOC: A Comprehensive Benchmark for Out-of-Context Evaluation in Multimodal Large Language Models Paper • 2607.27637 • Published Aug 1 • 7
SAE Interventions are Unreliable: Post-Intervention Recovery of Suppressed Behavior Paper • 2606.18322 • Published Jun 16 • 16
Dialogue Language Model with Large-Scale Persona Data Engineering Paper • 2412.09034 • Published Dec 12, 2024
InfantCryNet: A Data-driven Framework for Intelligent Analysis of Infant Cries Paper • 2409.19689 • Published Sep 29, 2024 • 1
Not All Disagreement Is Learnable: Token Teachability in On-Policy Distillation Paper • 2605.26844 • Published May 26 • 24
E-PMQ: Expert-Guided Post-Merge Quantization with Merged-Weight Anchoring Paper • 2605.16882 • Published May 16 • 2
ShadowPEFT: Shadow Network for Parameter-Efficient Fine-Tuning Paper • 2604.19254 • Published Apr 21 • 32
ShadowPEFT: Shadow Network for Parameter-Efficient Fine-Tuning Paper • 2604.19254 • Published Apr 21 • 32