-
facebook/seamless-streaming
Text-to-Speech ⢠Updated ⢠289 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition ⢠Updated ⢠7.38M ⢠3.91k -
pyannote/segmentation
Voice Activity Detection ⢠Updated ⢠809k ⢠699 -
pyannote/segmentation-3.0
Voice Activity Detection ⢠Updated ⢠5.47M ⢠1.99k
Myem
Sergioso
AI & ML interests
None yet
Organizations
None yet
GOOD
- RunningAgents523
tts Text To Speech
š523Text-to-speech (TTS) with Next-gen Kaldi
- PausedAgents11
Text To Speech
š11Transcribe audio to text
- Build errorAgents18
UTMOS Demo
š¢18Evaluate audio quality with MOS score
- Runtime errorAgentsFeatured220
SpeechT5 Speech Synthesis Demo
š©220
SD_Comfy_IMG
SDModels
AudioVideo
TTSSS
- Runtime errorAgents21
Youtube Video Translator
šØ21Translate YouTube videos to different languages
- RunningAgents523
tts Text To Speech
š523Text-to-speech (TTS) with Next-gen Kaldi
- Running on ZeroAgents521
AICoverGen
š521Launch a web UI for interacting with the model
- Runtime errorAgents316
Tortoise Tts
š¢316ExpressivText-to-Speech
Subtitle
Sdiff
AISTS
-
facebook/seamless-streaming
Text-to-Speech ⢠Updated ⢠289 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition ⢠Updated ⢠7.38M ⢠3.91k -
pyannote/segmentation
Voice Activity Detection ⢠Updated ⢠809k ⢠699 -
pyannote/segmentation-3.0
Voice Activity Detection ⢠Updated ⢠5.47M ⢠1.99k
AudioVideo
GOOD
- RunningAgents523
tts Text To Speech
š523Text-to-speech (TTS) with Next-gen Kaldi
- PausedAgents11
Text To Speech
š11Transcribe audio to text
- Build errorAgents18
UTMOS Demo
š¢18Evaluate audio quality with MOS score
- Runtime errorAgentsFeatured220
SpeechT5 Speech Synthesis Demo
š©220
TTSSS
- Runtime errorAgents21
Youtube Video Translator
šØ21Translate YouTube videos to different languages
- RunningAgents523
tts Text To Speech
š523Text-to-speech (TTS) with Next-gen Kaldi
- Running on ZeroAgents521
AICoverGen
š521Launch a web UI for interacting with the model
- Runtime errorAgents316
Tortoise Tts
š¢316ExpressivText-to-Speech
SD_Comfy_IMG
Subtitle
SDModels
Sdiff