Video Generation Models are General-Purpose Vision Learners Paper • 2607.09024 • Published 27 days ago • 87
view article Article Welcome Inkling by Thinking Machines +3 burtenshaw, merve, pcuenq, ariG23498, andito • 22 days ago • 150
SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer Paper • 2605.30409 • Published May 28 • 42
DINOv3 Collection DINOv3: foundation models producing excellent dense features, outperforming SotA w/o fine-tuning - https://arxiv.org/abs/2508.10104 • 15 items • Updated Mar 10 • 728
GraphLocator: Graph-guided Causal Reasoning for Issue Localization Paper • 2512.22469 • Published Dec 27, 2025 • 4
SpatialTree: How Spatial Abilities Branch Out in MLLMs Paper • 2512.20617 • Published Dec 23, 2025 • 44
view article Article Continuous batching from first principles +1 ror, ArthurZ, mcpotato • Nov 25, 2025 • 430
view article Article Building the Open Agent Ecosystem Together: Introducing OpenEnv +8 spisakjo, darktex, zkwentz, mortimerp9, Sanyam, Hamid-Nazeri, Pankit01, emre0, lewtun, reach-vb • Oct 23, 2025 • 166
The Well Collection A 15TB collection of physics simulation datasets. • 18 items • Updated Mar 24, 2025 • 53
MM Grounding DINO Collection See: https://github.com/huggingface/transformers/pull/37925 • 8 items • Updated Jun 26, 2025 • 5
LLMDet Collection See: https://github.com/huggingface/transformers/pull/37925 • 3 items • Updated Jun 26, 2025 • 3
SmolDocling datasets Collection Datasets used to train SmolDocling • 6 items • Updated Jul 31, 2025 • 31