BoundInk: Boundary-Aware Online Handwriting Generation Paper • 2604.02103 • Published 11 days ago • 6
Just Ask Jev: Reinforcement Learning for Calibrated Decisions as a Zero-Shot Detector of AI Alignment Failures Paper • 2609.29429 • Published 7 days ago • 23
RewardVerse: Rubric-Guided Policy Optimization for Video Reward Modeling Paper • 2609.22947 • Published 12 days ago • 41
Just-in-Time Memory: Learning to Curate Task-Adaptive Memory for LLM Agents Paper • 2609.27334 • Published 8 days ago • 51
StableVQ: Practical Guidelines for Stable Vector-Quantized Tokenizer Training Paper • 2609.26774 • Published 9 days ago • 56
Geometric and Semantic Coupling for Interaction Understanding in 3D Scenes Paper • 2609.25247 • Published 10 days ago • 10
RULER: Instance-aware Rubric Rewards for SVG Generation Paper • 2609.25270 • Published 10 days ago • 101
A Lie Detector Test for Language Models: Reading Knowledge a Model Won't Reveal Paper • 2609.21996 • Published 13 days ago • 9
CADWorld: Computer-Use Benchmark for Long-Horizon Computer-Aided Design Paper • 2609.16251 • Published 17 days ago • 14
SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Agent Harness Paper • 2609.20519 • Published 14 days ago • 138
DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 14 days ago • 191
PACT: Can Enterprise AI Assistants Be Trusted Under Pressure? Paper • 2609.18605 • Published 15 days ago • 41
VABench: Measuring Embodied Spatial Intelligence through Visual Demonstrations, Active Perception, and Metric Control Paper • 2609.19554 • Published 14 days ago • 43
LimiX-2: A Contextual Mechanism Network Towards General Structured-Data Intelligence Paper • 2609.17488 • Published 16 days ago • 812
Generalized Agent Iteration: One Formal Framework for Iterative Policy Improvement and Recursive Self-Improvement Paper • 2609.13406 • Published 20 days ago • 84