Rethinking Critic Learning in PPO: Understanding and Mitigating Value Flattening Paper • 2609.18708 • Published 16 days ago • 84
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 25 days ago • 376
Studying Image Tokenizers as Visual Languages in Unified Multimodal Models Paper • 2609.09143 • Published 24 days ago • 29
DriveZero: End-to-End Driving Beyond Human Demonstrations Paper • 2609.06055 • Published 27 days ago • 57
RoboTok: An Internet-Scale Data Engine for Human Demonstration Retrieval and Dexterous Manipulation Learning Paper • 2609.03199 • Published about 1 month ago • 31
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 266
GaussianSelector: Lightweight Human-Guided Object Selection in 3D Gaussian Splatting with Graph Optimization Paper • 2608.01492 • Published Aug 2 • 15
DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces Paper • 2608.03451 • Published Aug 4 • 34
Poly-OPD: Heterogeneous Multi-Teacher On-Policy Distillation for Capability-Selectable Flow Models Paper • 2608.04349 • Published Aug 5 • 13
StyleForge: Indoor Furniture Styling by Counterfactual Reasoning in a Hypergraph Field Paper • 2608.01954 • Published Aug 3 • 13