alvanlii/whisper-small-cantonese Automatic Speech Recognition • 0.2B • Updated Nov 20, 2025 • 219k • 118
Running on CPU Upgrade 602 Visualize Dataset (v2.0+ latest dataset format) 💻 602 Explore and visualize LeRobot datasets
HY-Embodied-0.5: Embodied Foundation Models for Real-World Agents Paper • 2604.07430 • Published Apr 8 • 181
Sommelier: Scalable Open Multi-turn Audio Pre-processing for Full-duplex Speech Language Models Paper • 2603.25750 • Published Mar 20 • 36
Running on CPU Upgrade Agents Featured 122 Cohere Multilingual ASR 🎙 122 Transcribe audio clips to text in multiple languages
OpenResearcher: A Fully Open Pipeline for Long-Horizon Deep Research Trajectory Synthesis Paper • 2603.20278 • Published Mar 17 • 102
Jackrong/Qwen3.5-9B-Claude-4.6-Opus-Reasoning-Distilled-v2-GGUF Image-Text-to-Text • 9B • Updated Apr 6 • 37.6k • 385