ComplexityMT: Benchmarking the Interaction Between Text Complexity and Machine Translation Paper • 2606.05421 • Published Jun 3 • 1
EditHero: A Benchmark for Long-Horizon Part-Level 3D Editing and Vibe Modeling Paper • 2610.02298 • Published 9 days ago • 57
CompoWorld: Compositional Environment Scaling for General Agents Paper • 2609.33665 • Published 13 days ago • 41
Raven: The Harness of Harnesses for Composable Agentic Intelligence Paper • 2609.33439 • Published 13 days ago • 672
TabFM-Auto: Self-Evolving Pipelines for Tabular Foundation Models Paper • 2609.37989 • Published 11 days ago • 12
What Makes Recurrence Effective in Looped Language Models? Paper • 2609.36636 • Published 11 days ago • 11
QwenGyre: An Elastic Reinforcement Learning Framework for Training xLong-Horizon Agents Paper • 2609.33848 • Published 13 days ago • 47
Knowing When Thinking Is Not Enough: Teaching Small Reasoning Models to Reason Beyond Their Parametric Knowledge Paper • 2609.34327 • Published 12 days ago • 41
FlowTool: Controlling Tool Parameter in Image Retouching via Flow Matching Paper • 2609.35673 • Published 12 days ago • 34
NanoForecast v0.5: Competitive Time Series Forecasting Through Training Pipeline Optimization Paper • 2609.31669 • Published 25 days ago • 3
Disaggregated Quantization: Specializing LLM Prefill and Decode Paper • 2609.26333 • Published 18 days ago • 93
TimeEvo: Failure-Driven Self-Evolution of a Time Series Agent Paper • 2609.27277 • Published 17 days ago • 33
Jev thinks "I don't know'', but doesn't say it: Introducing Sys1Cal-v1 Dataset for Probability Calibration Paper • 2609.35342 • Published 12 days ago • 8
Paragraph Boundaries Are Not White Space:Compression Depth as the Signature of Hierarchical Structure Paper • 2609.23551 • Published 20 days ago • 7
Do Implicit Personalization and Explicit Styles Conflict? PsPLUG: A Lightweight Plug-in for Balancing Personalization and Style in Customized LLMs Paper • 2601.06362 • Published 20 days ago • 10