Region-Level Policy Optimization for Fine-grained MLLM Perception Paper • 2609.19745 • Published 7 days ago • 41
SpectralShift: Effective Context Window Extension of Gated DeltaNet via Spectral Reparameterization Paper • 2609.14320 • Published 11 days ago • 33
TempCloze: Can Video-LLMs Identify the Missing Middle? Paper • 2609.01515 • Published 23 days ago • 31
ActReview: Rebuttal-Guided Training Data and Rubric Rewards for Actionable Peer Review Generation Paper • 2609.09076 • Published 16 days ago • 24
microsoft/VibeVoice-ASR-Streaming-7B Automatic Speech Recognition • 9B • Updated 21 days ago • 6.46k • 247
Knowing When Not to Reuse: Conditional Experience Transfer in Autonomous LLM Post-Training Paper • 2608.26730 • Published 28 days ago • 155
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published 21 days ago • 186
GOAG: Generative and Object-Agnostic Grasp Planner for Dexterous Robotic Manipulation Paper • 2608.19759 • Published Aug 20 • 8
Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination Paper • 2608.14391 • Published Aug 14 • 285
CW-BASS v2: Saturation-Aware Pseudo-Label Selection for Semi-Supervised Segmentation under Foundation-Model Teachers Paper • 2608.12773 • Published Aug 13 • 9
Specification-first convergence with an AI coding agent: a case study of dismantling a core architectural invariant across 189 files in a 717k-line codebase with no test oracle and no human code review Paper • 2608.12440 • Published Aug 12 • 11
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 265