Collections
Discover the best community collections!
Collections including paper arxiv:2607.20709
-
TIPSv2: Advancing Vision-Language Pretraining with Enhanced Patch-Text Alignment
Paper • 2604.12012 • Published • 15 -
InferenceSupport
💥609Discussions about the Inference Providers feature on the Hub
-
NVIDIA-labs OO Agents: Native Python Object-Oriented Agents
Paper • 2607.20709 • Published • 25 -
SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation
Paper • 2607.21553 • Published • 25
-
Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs
Paper • 2501.18585 • Published • 61 -
RWKV-7 "Goose" with Expressive Dynamic State Evolution
Paper • 2503.14456 • Published • 154 -
DeepMesh: Auto-Regressive Artist-mesh Creation with Reinforcement Learning
Paper • 2503.15265 • Published • 46 -
Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning
Paper • 2503.15558 • Published • 52
-
PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion
Paper • 2605.23902 • Published • 47 -
minWM: A Full-Stack Open-Source Framework for Real-Time Interactive Video World Models
Paper • 2605.30263 • Published • 59 -
From Pixels to Words -- Towards Native One-Vision Models at Scale
Paper • 2605.28820 • Published • 76 -
open-thoughts/AgentTrove
Viewer • Updated • 1.7M • 2.17k • 189
-
Agentic Reasoning and Tool Integration for LLMs via Reinforcement Learning
Paper • 2505.01441 • Published • 39 -
LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities
Paper • 2504.16078 • Published • 21 -
Emergent Agentic Transformer from Chain of Hindsight Experience
Paper • 2305.16554 • Published -
DiaTool-DPO: Multi-Turn Direct Preference Optimization for Tool-Augmented Large Language Models
Paper • 2504.02882 • Published • 7
-
PiD: Fast and High-Resolution Latent Decoding with Pixel Diffusion
Paper • 2605.23902 • Published • 47 -
minWM: A Full-Stack Open-Source Framework for Real-Time Interactive Video World Models
Paper • 2605.30263 • Published • 59 -
From Pixels to Words -- Towards Native One-Vision Models at Scale
Paper • 2605.28820 • Published • 76 -
open-thoughts/AgentTrove
Viewer • Updated • 1.7M • 2.17k • 189
-
TIPSv2: Advancing Vision-Language Pretraining with Enhanced Patch-Text Alignment
Paper • 2604.12012 • Published • 15 -
InferenceSupport
💥609Discussions about the Inference Providers feature on the Hub
-
NVIDIA-labs OO Agents: Native Python Object-Oriented Agents
Paper • 2607.20709 • Published • 25 -
SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals for Efficient Video Generation
Paper • 2607.21553 • Published • 25
-
Agentic Reasoning and Tool Integration for LLMs via Reinforcement Learning
Paper • 2505.01441 • Published • 39 -
LLMs are Greedy Agents: Effects of RL Fine-tuning on Decision-Making Abilities
Paper • 2504.16078 • Published • 21 -
Emergent Agentic Transformer from Chain of Hindsight Experience
Paper • 2305.16554 • Published -
DiaTool-DPO: Multi-Turn Direct Preference Optimization for Tool-Augmented Large Language Models
Paper • 2504.02882 • Published • 7
-
Thoughts Are All Over the Place: On the Underthinking of o1-Like LLMs
Paper • 2501.18585 • Published • 61 -
RWKV-7 "Goose" with Expressive Dynamic State Evolution
Paper • 2503.14456 • Published • 154 -
DeepMesh: Auto-Regressive Artist-mesh Creation with Reinforcement Learning
Paper • 2503.15265 • Published • 46 -
Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning
Paper • 2503.15558 • Published • 52