nvidia/nemotron-3.5-asr-streaming-0.6b Automatic Speech Recognition • 0.6B • Updated 23 days ago • 961k • • 963
EnterpriseOps-Gym: Environments and Evaluations for Stateful Agentic Planning and Tool Use in Enterprise Settings Paper • 2603.13594 • Published Mar 13 • 150
Running on CPU Upgrade Agents 14 LLM Beer Distribution Game Public 👁 14 Play an interactive beer distribution game with AI
view article Article Nemotron 3 Nano \- A new Standard for Efficient, Open, and Intelligent Agentic Models nvidia • Dec 15, 2025 • 114
view article Article We Got Claude to Fine-Tune an Open Source LLM burtenshaw, evalstate • Dec 4, 2025 • 631
Running on Zero MCP Featured 2.25k Qwen Image Edit Camera Control 🎬 2.25k Fast 4 step inference with Qwen Image Edit 2509
Running Agents 2 Cache-to-Cache Communication Demo 🔗 2 Compare Single, Text-to-Text, and Cache-to-Cache inference
Running on CPU Upgrade Featured 3.25k The Smol Training Playbook 📚 3.25k The secrets to building world-class LLMs
view article Article Supercharge your OCR Pipelines with Open Models +5 merve, ariG23498, davanstrien, hynky, andito, reach-vb, pcuenq • Oct 21, 2025 • 317
view article Article StackLLaMA: A hands-on guide to train LLaMA with RLHF +5 edbeeching, kashif, ybelkada, lewtun, lvwerra, nazneen, natolambert • Apr 5, 2023 • 48
rogue-security/prompt-injection-jailbreak-sentinel-v2 Text Classification • 0.6B • Updated Mar 11 • 26.2k • 36
MCPMark: A Benchmark for Stress-Testing Realistic and Comprehensive MCP Use Paper • 2509.24002 • Published Sep 28, 2025 • 180
The Dragon Hatchling: The Missing Link between the Transformer and Models of the Brain Paper • 2509.26507 • Published Sep 30, 2025 • 551
Less is More: Recursive Reasoning with Tiny Networks Paper • 2510.04871 • Published Oct 6, 2025 • 518