nvidia/nemotron-3.5-asr-streaming-0.6b Automatic Speech Recognition • 0.6B • Updated 23 days ago • 1.26M • • 1.16k
yuxinlu1/gemma-4-12B-agentic-fable5-composer2.5-v2-3.5x-tau2-GGUF Text Generation • 12B • Updated Jun 19 • 698k • 1.63k
DreamActor-M1: Holistic, Expressive and Robust Human Image Animation with Hybrid Guidance Paper • 2504.01724 • Published Apr 2, 2025 • 68
F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching Paper • 2410.06885 • Published Oct 9, 2024 • 49