Running on Zero Agents Featured 5.15k FLUX.1 [Schnell] π 5.15k Generate images from text prompts with FLUX.1-schnell
Running Agents Featured 1.14k OpenVoice π€ 1.14k Generate speech in a cloned voice from a short audio clip
Running on CPU Upgrade Featured 977 TTS Arena V2 π£ 977 Compare two text-to-speech voices and vote for the better
Runtime error Agents 408 HierSpeech++ (Zero-shot TTS) β‘ 408 Generate high-quality speech from text using a prompt audio
Running on Zero Agents Featured 399 Playground V2 π 399 Generate images from text prompts with customizable options
playgroundai/playground-v2-1024px-aesthetic Text-to-Image β’ 3B β’ Updated Feb 23, 2024 β’ 333 β’ 560
pyannote/speaker-diarization-3.1 Automatic Speech Recognition β’ Updated May 10, 2024 β’ 7.26M β’ 4k
Running on Zero Agents Featured 734 StyleTTS 2 π£ 734 Efficient, fast, and natural text to speech with StyleTTS 2!
stabilityai/stable-video-diffusion-img2vid-xt Image-to-Video β’ 2B β’ Updated Jul 10, 2024 β’ 210k β’ 3.43k
Running Agents 190 Gradio Lipsync Wav2lip π 190 Generate lipβsynced video from a face image and audio
pyannote/speaker-diarization Automatic Speech Recognition β’ Updated May 10, 2024 β’ 338k β’ 1.34k