The Optimizer Is the Agent: Reasoning-Driven Search across Prompts, Programs, and ML Workflows Paper • 2608.06714 • Published 4 days ago • 4
Characterizing the Quality Profile of AI-Generated C++ in Production Paper • 2608.06640 • Published 5 days ago • 4
Modular TTT: Rethinking Test-Time Training as Composable Modules Paper • 2608.07110 • Published 4 days ago • 5
GST-Bench: Can VLMs Develop Global Spatial Awareness from Video? Paper • 2608.05747 • Published 5 days ago • 42
HarnessOpt-Bench: Evaluating LLMs at Harness Optimization Paper • 2608.06301 • Published 5 days ago • 33
DyPES-VLA: Learning Shared Dynamics Priors and Embodiment-Specific Control for Cross-Embodiment Manipulation Paper • 2608.06374 • Published 5 days ago • 22
ToolArtist: Tool-Using Unified Multimodal Models for Agentic Image Generation Paper • 2608.04436 • Published 6 days ago • 55
HelloWorld: Enabling Socially Interactive Characters in Video World Models Paper • 2608.05070 • Published 6 days ago • 35
OPD-V: Visual On-Policy Self-Distillation with Modality Balance Paper • 2608.05131 • Published 6 days ago • 9
PAST-Bench: Benchmarking the Foundations of Recursive Self-Improvement in Personal Agents Paper • 2608.04003 • Published 7 days ago • 33
JoyAI-Video-Edit: Real-Time Open-Ended Video Editing with Autoregressive Diffusion Paper • 2608.03974 • Published 7 days ago • 90
Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing Paper • 2608.02711 • Published 8 days ago • 86
Evaluation-Verification Reward for Consistent Multi-Reference Image Editing Paper • 2607.29025 • Published 11 days ago • 16
ExtractBench: A Benchmark for Schema-Guided Enterprise Document Extraction Paper • 2607.29677 • Published 11 days ago • 23
Meshy T2: Fast Native Mesh Generation with Flow Matching Paper • 2607.28675 • Published 14 days ago • 55