Systematically Exploring the Capabilities of GPT-6 Astra as Embodied Policies Paper • 2609.38537 • Published 5 days ago • 30
Systematically Exploring the Capabilities of GPT-6 Astra as Embodied Policies Paper • 2609.38537 • Published 5 days ago • 30
DeformGen Collection Official DeformGen artifacts. Paper: arxiv.org/abs/2606.25939. Code: github.com/Zili2002/DeformGen • 3 items • Updated Jun 25 • 1
ImageWAM Collection Models of ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing? • 6 items • Updated Jul 29 • 2
HumanTracker: Towards Comprehensive and Human-Aligned Motion Tracking Benchmark Paper • 2608.13555 • Published Aug 13 • 18
HumanTracker: Towards Comprehensive and Human-Aligned Motion Tracking Benchmark Paper • 2608.13555 • Published Aug 13 • 18
PerceptionRubrics: Calibrating Multimodal Evaluation to Human Perception Paper • 2606.28322 • Published Jun 26 • 42
SenseNova-U1 Collection SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-Unify Architecture • 12 items • Updated Aug 14 • 77
ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing? Paper • 2606.19531 • Published Jun 17 • 28
ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing? Paper • 2606.19531 • Published Jun 17 • 28
Hybrid-grained Feature Aggregation with Coarse-to-fine Language Guidance for Self-supervised Monocular Depth Estimation Paper • 2510.09320 • Published Oct 10, 2025 • 3
VLA-JEPA: Enhancing Vision-Language-Action Model with Latent World Model Paper • 2602.10098 • Published Feb 10 • 24