Running Agents 73 Qwen3.5 Omni Online Demo 📚 73 Chat with a multimodal AI using text, image, audio, or video
Running Agents Featured 412 Qwen3 TTS Demo 🚀 412 Generate spoken audio from your text in many voices
Running Agents 115 Qwen3 TTS Voice Design 📈 115 Generate custom speech from text and voice description
SoulX-Podcast: Towards Realistic Long-form Podcasts with Dialectal and Paralinguistic Diversity Paper • 2510.23541 • Published Oct 27, 2025 • 17
HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning Paper • 2509.08519 • Published Sep 10, 2025 • 130
ThinkSound: Chain-of-Thought Reasoning in Multimodal Large Language Models for Audio Generation and Editing Paper • 2506.21448 • Published Jun 26, 2025 • 9
Running on Zero Agents Featured 675 ACE Step 😻 675 A Step Towards Music Generation Foundation Model