AREX-2: Advancing Self-Improving Agents through Long-Horizon Reflective Tasks Paper • 2609.38288 • Published 7 days ago • 140
MassAlloc Attention: Let Attention Allocate Its Own Compute Paper • 2609.32712 • Published 10 days ago • 75
CoWindow Attention: Full Causal Coverage Is a Collective Property Paper • 2609.32704 • Published 10 days ago • 67
CoWindow Attention: Full Causal Coverage Is a Collective Property Paper • 2609.32704 • Published 10 days ago • 67
MassAlloc Attention: Let Attention Allocate Its Own Compute Paper • 2609.32712 • Published 10 days ago • 75
MotionVLA: Vision-Language-Action Model for Humanoid Motion Paper • 2606.15142 • Published Jun 13 • 6
DragMesh-2: Physically Plausible Dexterous Hand-Object Interaction with Articulated Objects Paper • 2606.15133 • Published Jun 13 • 34
ImpossibleRubrics: Stress-Testing Generated Rubrics as Reward Signals Paper • 2609.16816 • Published 21 days ago • 11
ImpossibleRubrics: Stress-Testing Generated Rubrics as Reward Signals Paper • 2609.16816 • Published 21 days ago • 11
FLM-101B: An Open LLM and How to Train It with $100K Budget Paper • 2309.03852 • Published Sep 7, 2023 • 45
Can LLM Already Serve as A Database Interface? A BIg Bench for Large-Scale Database Grounded Text-to-SQLs Paper • 2305.03111 • Published May 4, 2023 • 12
Graphix-T5: Mixing Pre-Trained Transformers with Graph-Aware Layers for Text-to-SQL Parsing Paper • 2301.07507 • Published Jan 18, 2023