Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

ViuAI
/
ViuAI_TTS_200M

Text-to-Speech
PyTorch
Hindi
English
tts
flow-matching
diffusion-transformer
voice-cloning
zero-shot
elevenlabs-style
Model card Files Files and versions
xet
Community
ViuAI_TTS_200M / models
43.8 kB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 9 commits
ViuAI's picture
ViuAI
Fix: Peak audio normalization (-1 dB), robust speaker search, and tuned CFG for crystal-clear benchmark generation
d62a3fe verified 7 days ago
  • __init__.py
    422 Bytes
    Initial commit: ViuAI_TTS_200M architecture, prosody engine, and weights 10 days ago
  • dit_backbone.py
    10.7 kB
    Fix: Peak audio normalization (-1 dB), robust speaker search, and tuned CFG for crystal-clear benchmark generation 7 days ago
  • duration_predictor.py
    3.11 kB
    Initial commit: ViuAI_TTS_200M architecture, prosody engine, and weights 10 days ago
  • flow_matching.py
    4.08 kB
    Initial commit: ViuAI_TTS_200M architecture, prosody engine, and weights 10 days ago
  • text_encoder.py
    5.65 kB
    Update: training fixes, LICENSE + requirements + scripts 9 days ago
  • viuai_tts.py
    11.6 kB
    Fix: Peak audio normalization (-1 dB), robust speaker search, and tuned CFG for crystal-clear benchmark generation 7 days ago
  • vocoder.py
    8.22 kB
    Fix: Peak audio normalization (-1 dB), robust speaker search, and tuned CFG for crystal-clear benchmark generation 7 days ago