Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
ViuAI
/
ViuAI_TTS_200M
Like
0
Text-to-Speech
PyTorch
Hindi
English
tts
flow-matching
diffusion-transformer
voice-cloning
zero-shot
elevenlabs-style
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
ViuAI_TTS_200M
/
models
43.8 kB
Ctrl+K
Ctrl+K
1 contributor
History:
9 commits
ViuAI
Fix: Peak audio normalization (-1 dB), robust speaker search, and tuned CFG for crystal-clear benchmark generation
d62a3fe
verified
7 days ago
__init__.py
Safe
422 Bytes
Initial commit: ViuAI_TTS_200M architecture, prosody engine, and weights
10 days ago
dit_backbone.py
10.7 kB
Fix: Peak audio normalization (-1 dB), robust speaker search, and tuned CFG for crystal-clear benchmark generation
7 days ago
duration_predictor.py
Safe
3.11 kB
Initial commit: ViuAI_TTS_200M architecture, prosody engine, and weights
10 days ago
flow_matching.py
Safe
4.08 kB
Initial commit: ViuAI_TTS_200M architecture, prosody engine, and weights
10 days ago
text_encoder.py
Safe
5.65 kB
Update: training fixes, LICENSE + requirements + scripts
9 days ago
viuai_tts.py
11.6 kB
Fix: Peak audio normalization (-1 dB), robust speaker search, and tuned CFG for crystal-clear benchmark generation
7 days ago
vocoder.py
8.22 kB
Fix: Peak audio normalization (-1 dB), robust speaker search, and tuned CFG for crystal-clear benchmark generation
7 days ago