How to Use NVIDIA Warp and MjWarp to Accelerate Robotics Simulation and Learning Workflows 1 day ago • 22
**Know Who Spoke When: Build Real-Time, Multi-Speaker AI with NVIDIA Nemotron 3 Diarization** 2 days ago • 50
Build Low-Latency Multilingual Voice Agents: Open Weights & Full Deployment Control with NVIDIA Magpie TTS Aug 10 • 38
Introducing NVIDIA Nemotron 3 Nano Omni: Long-Context Multimodal Intelligence for Documents, Audio and Video Agents Apr 28 • 66
The First Healthcare Robotics Dataset and Foundational Physical AI Models for Healthcare Robotics Mar 16 • 34
Beyond Semantic Similarity: Introducing NVIDIA NeMo Retriever’s Generalizable Agentic Retrieval Pipeline Mar 13 • 43
Build an Agent That Thinks Like a Data Scientist: How We Hit #1 on DABStep with Reusable Tool Generation Mar 13 • 19
NVIDIA Nemotron 2 Nano 9B Japanese: State-of-the-Art Small Language Model Customized for Japanese Sovereign AI Feb 17 • 3
Small Yet Mighty: Improve Accuracy In Multimodal Search and Visual Document Retrieval with Llama Nemotron RAG Models Jan 6 • 30
The Open Evaluation Standard: Benchmarking NVIDIA Nemotron 3 Nano with NeMo Evaluator Dec 17, 2025 • 51
Nemotron 3 Nano \- A new Standard for Efficient, Open, and Intelligent Agentic Models Dec 15, 2025 • 114
How to Build a Healthcare Robot from Simulation to Deployment with NVIDIA Isaac for Healthcare Oct 28, 2025 • 20
NVIDIA Releases 8 Million Sample Open Dataset and Tooling for OCR, Image Reasoning, Image and Video QA Tasks Oct 28, 2025 • 18
Cosmos Predict 2.5 & Transfer 2.5: Evolving the World Foundation Models for Physical AI Oct 28, 2025 • 22
Nemotron’s Open Secret: Accelerating AI Development with Open Models, Data, and Recipes Oct 22, 2025 • 12
Llama‑Embed‑Nemotron‑8B Text Embedding Model Ranks First on Multilingual MTEB Leaderboard Oct 21, 2025 • 15
Scaling Test-Time Compute to Achieve Gold Medal at IOI 2025 with Open-Weight Models Oct 20, 2025 • 19
📢 NVIDIA Releases Nemotron-CC-Math Pre-Training Dataset: A High-Quality, Web-Scale Math Corpus for Pretraining Large Language Models Aug 18, 2025 • 6
NVIDIA Releases Improved Pretraining Dataset: Preserves High Value Math & Code, and Augments with Multi-Lingual Aug 18, 2025 • 5
NVIDIA Releases 3 Million Sample Dataset for OCR, Visual Question Answering, and Captioning Tasks Aug 11, 2025 • 77
Llama-NeMoRetriever-ColEmbed: Developer-Focused Guide to NVIDIA's State-of-the-Art Text-Image Retrieval Jul 9, 2025 • 5
Nemotron-Personas: Improve AI Training With the First Synthetic Personas Dataset Aligned to Real-World Distributions Jun 10, 2025 • 25
Submitted by Min-Hung Chen 113 SpatialClaw: Rethinking Action Interface for Agentic Spatial Reasoning NVIDIA 387 4
Submitted by Siyi Chen 6 VoLo: A Physical Orchestrator for Open-Vocabulary Long-Horizon Manipulation NVIDIA 2
Submitted by taesiri 9 GRAIL: Generating Humanoid Loco-Manipulation from 3D Assets and Video Priors NVIDIA 531 1
Submitted by Yaosheng Fu 5 SparDA: Sparse Decoupled Attention for Efficient Long-Context LLM Inference NVIDIA 73 2
Submitted by Yoad Tewel 19 Bootstrap Your Generator: Unpaired Visual Editing with Flow Matching NVIDIA 2
Submitted by taesiri 21 NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation NVIDIA 342 1
Submitted by Wei Huang 19 LongLive-RAG: A General Retrieval-Augmented Framework for Long Video Generation NVIDIA 113 1
Submitted by taesiri 3 Light Interaction: Training-Free Inference Acceleration for Interactive Video World Models NVIDIA 2
Submitted by Chan Hee Song 54 Why Far Looks Up: Probing Spatial Representation in Vision-Language Models NVIDIA 17 3
Submitted by Yuyang 38 SANA-Streaming: Real-time Streaming Video Editing with Hybrid Diffusion Transformer NVIDIA 2
Submitted by Rishit Dagli 2 FreeForm: Reduced-Order Deformable Simulation from Particle-Based Skinning Eigenmodes NVIDIA 5.18k 2