Sergio Paniego PRO
AI & ML interests
Recent Activity
Organizations
- Running4.05k
The Ultra-Scale Playbook
🌌4.05kThe ultimate guide to training LLM on large GPU Clusters
- Running on CPU UpgradeFeatured3.31k
The Smol Training Playbook
📚3.31kThe secrets to building world-class LLMs
- Running361
Evaluation Guidebook
📝361Explore LLM benchmark scores over time
- Running236
FineVision: Open Data is All You Need
📝236A new open-source dataset for training VLMs
- RunningAgents43
comparevlms
🏃43Compare Vision Language Models
- Runtime errorAgents4
Gemma3 License Plate Detection
📈4Gemma 3 for license plate detection
- Running on ZeroAgentsFeatured144
Gemma 3n E4B It
⚡144Chat with an AI that understands text, images, audio, and video
- Running on ZeroAgentsFeatured46
Moondream3
🏢46Image and video tasks with moondream3.
-
sergiopaniego/rl-envs-youtube-livestream-4-scripts
Updated - Running on CPU UpgradeRL
Coding Environment Server
💻Run Python code in an interactive coding environment
- SleepingAgents
Trackio Training Agents 4
🎯Show live tracking data in a visual interface
-
sergiopaniego/qwen3-1.7b-mbpp-grpo
Text Generation • 2B • Updated • 491
- Runtime errorRL
CARLA Environment Server
🚗Control a Carla driving simulation with custom actions
- Runtime errorRL
CARLA Environment Server
🚗Control a CARLA driving simulator with custom actions
- SleepingAgents
Carla Grpo Trolley
🚀Visualize your program’s I/O activity in real time
-
sergiopaniego/Qwen3-0.6B-carla-trolley-escape
0.8B • Updated • 15
- RunningAgents43
comparevlms
🏃43Compare Vision Language Models
- Running on ZeroAgents68
OCR Time Machine
📚68Extract text from images and XML files using OCR models
- SleepingAgents26
Compare Docvqa Models
🦀26Compare different visual question answering
- Running on CPU UpgradeAgents23
Compare Clip Siglip
🏃23Compare strong zero-shot image classification models
-
Qwen/Qwen2.5-Omni-7B
Any-to-Any • 11B • Updated • 332k • 1.95k - RunningAgentsFeatured375
Qwen2.5 Omni 7B Demo
🏆375Chat with text, audio, images, and video, get spoken replies
-
Qwen2.5-Omni Technical Report
Paper • 2503.20215 • Published • 174 -
openbmb/MiniCPM-o-2_6
Any-to-Any • 9B • Updated • 321k • 1.3k
-
sergiopaniego/rl-envs-youtube-livestream-4-scripts
Updated - Running on CPU UpgradeRL
Coding Environment Server
💻Run Python code in an interactive coding environment
- SleepingAgents
Trackio Training Agents 4
🎯Show live tracking data in a visual interface
-
sergiopaniego/qwen3-1.7b-mbpp-grpo
Text Generation • 2B • Updated • 491
- Runtime errorRL
CARLA Environment Server
🚗Control a Carla driving simulation with custom actions
- Runtime errorRL
CARLA Environment Server
🚗Control a CARLA driving simulator with custom actions
- SleepingAgents
Carla Grpo Trolley
🚀Visualize your program’s I/O activity in real time
-
sergiopaniego/Qwen3-0.6B-carla-trolley-escape
0.8B • Updated • 15
- Running4.05k
The Ultra-Scale Playbook
🌌4.05kThe ultimate guide to training LLM on large GPU Clusters
- Running on CPU UpgradeFeatured3.31k
The Smol Training Playbook
📚3.31kThe secrets to building world-class LLMs
- Running361
Evaluation Guidebook
📝361Explore LLM benchmark scores over time
- Running236
FineVision: Open Data is All You Need
📝236A new open-source dataset for training VLMs
- RunningAgents43
comparevlms
🏃43Compare Vision Language Models
- Running on ZeroAgents68
OCR Time Machine
📚68Extract text from images and XML files using OCR models
- SleepingAgents26
Compare Docvqa Models
🦀26Compare different visual question answering
- Running on CPU UpgradeAgents23
Compare Clip Siglip
🏃23Compare strong zero-shot image classification models
- RunningAgents43
comparevlms
🏃43Compare Vision Language Models
- Runtime errorAgents4
Gemma3 License Plate Detection
📈4Gemma 3 for license plate detection
- Running on ZeroAgentsFeatured144
Gemma 3n E4B It
⚡144Chat with an AI that understands text, images, audio, and video
- Running on ZeroAgentsFeatured46
Moondream3
🏢46Image and video tasks with moondream3.
-
Qwen/Qwen2.5-Omni-7B
Any-to-Any • 11B • Updated • 332k • 1.95k - RunningAgentsFeatured375
Qwen2.5 Omni 7B Demo
🏆375Chat with text, audio, images, and video, get spoken replies
-
Qwen2.5-Omni Technical Report
Paper • 2503.20215 • Published • 174 -
openbmb/MiniCPM-o-2_6
Any-to-Any • 9B • Updated • 321k • 1.3k