MeloTTS-ZH: Optimized for Qualcomm Devices

MeloTTS is a high-quality multi-lingual text-to-speech library for English, Chinese and Spanish language.

This is based on the implementation of MeloTTS-ZH found here. This repository contains pre-exported model files optimized for Qualcomm® devices. You can use the Qualcomm® AI Hub Models library to export with custom configurations. More details on model performance across various devices, can be found here.

Qualcomm AI Hub Models uses Qualcomm AI Hub Workbench to compile, profile, and evaluate this model. Sign up to run these models on a hosted Qualcomm® device.

Deploying MeloTTS-ZH on-device

This model is compatible with the Qualcomm Voice AI SDK. Download the SDK from the Qualcomm Package Manager to deploy this model on-device.

Getting Started

There are two ways to deploy this model on your device:

Option 1: Download Pre-Exported Models

Below are pre-exported model assets ready for deployment.

Runtime Precision Chipset SDK Versions Download
VOICE_AI mixed_with_float Snapdragon® 8 Elite Gen 5 For Galaxy Mobile QAIRT 2.50 Download
VOICE_AI mixed_with_float Snapdragon® 8 Elite For Galaxy Mobile QAIRT 2.50 Download
VOICE_AI mixed_with_float Snapdragon® X2 Elite QAIRT 2.50 Download
VOICE_AI mixed_with_float Snapdragon® X Elite QAIRT 2.50 Download
VOICE_AI mixed_with_float Snapdragon® 8 Gen 3 Mobile QAIRT 2.50 Download
VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-8275 QAIRT 2.50 Download
VOICE_AI mixed_with_float Qualcomm® Dragonwing™ QCS8550 (Proxy) QAIRT 2.50 Download
VOICE_AI mixed_with_float Qualcomm® SA8775P QAIRT 2.50 Download
VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-9075 QAIRT 2.50 Download
VOICE_AI mixed_with_float Qualcomm® SA7255P QAIRT 2.50 Download

For more device-specific assets and performance metrics, visit MeloTTS-ZH on Qualcomm® AI Hub.

Option 2: Export with Custom Configurations

Use the Qualcomm® AI Hub Models Python library to compile and export the model with your own:

  • Custom weights (e.g., fine-tuned checkpoints)
  • Custom input shapes
  • Target device and runtime configurations

This option is ideal if you need to customize the model beyond the default configuration provided here.

See our repository for MeloTTS-ZH on GitHub for usage instructions.

Model Details

Model Type: Model_use_case.audio_generation

Model Stats:

  • Max decoded sequence length: 512 tokens
  • Model checkpoint: myshell-ai/MeloTTS-Chinese
  • Model size (bert_wrapper) (float): 581 MB
  • Model size (decoder) (float): 55.5 MB
  • Model size (encoder) (float): 31.9 MB
  • Model size (flow) (float): 76.9 MB
  • Number of parameters (bert_wrapper): 152M
  • Number of parameters (decoder): 14.5M
  • Number of parameters (encoder): 8.34M
  • Number of parameters (flow): 20.1M

Performance Summary

Model Runtime Precision Chipset Inference Time (ms) Peak Memory Range (MB) Primary Compute Unit
bert_wrapper VOICE_AI mixed_with_float Snapdragon® 8 Elite Gen 5 For Galaxy Mobile 2.623 ms 0 - 7 MB NPU
bert_wrapper VOICE_AI mixed_with_float Snapdragon® 8 Elite For Galaxy Mobile 3.298 ms 0 - 8 MB NPU
bert_wrapper VOICE_AI mixed_with_float Snapdragon® X2 Elite 3.303 ms 0 - 0 MB NPU
bert_wrapper VOICE_AI mixed_with_float Snapdragon® X Elite 7.681 ms 0 - 0 MB NPU
bert_wrapper VOICE_AI mixed_with_float Snapdragon® 8 Gen 3 Mobile 4.996 ms 0 - 7 MB NPU
bert_wrapper VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-8275 9.91 ms 0 - 4 MB NPU
bert_wrapper VOICE_AI mixed_with_float Qualcomm® Dragonwing™ QCS8550 (Proxy) 7.021 ms 0 - 2 MB NPU
bert_wrapper VOICE_AI mixed_with_float Qualcomm® SA8650P 9.295 ms 0 - 9 MB NPU
bert_wrapper VOICE_AI mixed_with_float Qualcomm® SA8255P 9.295 ms 0 - 9 MB NPU
bert_wrapper VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-9075 8.935 ms 2 - 5 MB NPU
bert_wrapper VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-X7181 7.681 ms 0 - 0 MB NPU
bert_wrapper VOICE_AI mixed_with_float Qualcomm® Dragonwing™ Q-8750 3.298 ms 0 - 8 MB NPU
decoder VOICE_AI mixed_with_float Snapdragon® 8 Elite Gen 5 For Galaxy Mobile 36.755 ms 0 - 7 MB NPU
decoder VOICE_AI mixed_with_float Snapdragon® 8 Elite For Galaxy Mobile 42.641 ms 0 - 8 MB NPU
decoder VOICE_AI mixed_with_float Snapdragon® X2 Elite 36.335 ms 0 - 0 MB NPU
decoder VOICE_AI mixed_with_float Snapdragon® X Elite 72.477 ms 1 - 1 MB NPU
decoder VOICE_AI mixed_with_float Snapdragon® 8 Gen 3 Mobile 51.843 ms 0 - 8 MB NPU
decoder VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-8275 64.406 ms 0 - 3 MB NPU
decoder VOICE_AI mixed_with_float Qualcomm® Dragonwing™ QCS8550 (Proxy) 72.318 ms 1 - 3 MB NPU
decoder VOICE_AI mixed_with_float Qualcomm® SA8650P 83.902 ms 0 - 10 MB NPU
decoder VOICE_AI mixed_with_float Qualcomm® SA8255P 83.902 ms 0 - 10 MB NPU
decoder VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-9075 71.27 ms 0 - 2 MB NPU
decoder VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-X7181 72.477 ms 1 - 1 MB NPU
decoder VOICE_AI mixed_with_float Qualcomm® Dragonwing™ Q-8750 42.641 ms 0 - 8 MB NPU
encoder VOICE_AI mixed_with_float Snapdragon® 8 Elite Gen 5 For Galaxy Mobile 15.396 ms 2 - 10 MB NPU
encoder VOICE_AI mixed_with_float Snapdragon® 8 Elite For Galaxy Mobile 16.375 ms 2 - 10 MB NPU
encoder VOICE_AI mixed_with_float Snapdragon® X2 Elite 15.247 ms 4 - 4 MB NPU
encoder VOICE_AI mixed_with_float Snapdragon® X Elite 26.235 ms 4 - 4 MB NPU
encoder VOICE_AI mixed_with_float Snapdragon® 8 Gen 3 Mobile 18.458 ms 4 - 11 MB NPU
encoder VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-8275 27.264 ms 4 - 11 MB NPU
encoder VOICE_AI mixed_with_float Qualcomm® Dragonwing™ QCS8550 (Proxy) 26.028 ms 4 - 5 MB NPU
encoder VOICE_AI mixed_with_float Qualcomm® SA8650P 43.115 ms 2 - 11 MB NPU
encoder VOICE_AI mixed_with_float Qualcomm® SA8255P 43.115 ms 2 - 11 MB NPU
encoder VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-9075 27.43 ms 4 - 10 MB NPU
encoder VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-X7181 26.235 ms 4 - 4 MB NPU
encoder VOICE_AI mixed_with_float Qualcomm® Dragonwing™ Q-8750 16.375 ms 2 - 10 MB NPU
flow VOICE_AI mixed_with_float Snapdragon® 8 Elite Gen 5 For Galaxy Mobile 31.883 ms 2 - 10 MB NPU
flow VOICE_AI mixed_with_float Snapdragon® 8 Elite For Galaxy Mobile 42.259 ms 2 - 10 MB NPU
flow VOICE_AI mixed_with_float Snapdragon® X2 Elite 31.799 ms 2 - 2 MB NPU
flow VOICE_AI mixed_with_float Snapdragon® X Elite 74.697 ms 2 - 2 MB NPU
flow VOICE_AI mixed_with_float Snapdragon® 8 Gen 3 Mobile 51.804 ms 2 - 9 MB NPU
flow VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-8275 69.106 ms 2 - 7 MB NPU
flow VOICE_AI mixed_with_float Qualcomm® Dragonwing™ QCS8550 (Proxy) 73.464 ms 3 - 4 MB NPU
flow VOICE_AI mixed_with_float Qualcomm® SA8650P 128.354 ms 2 - 12 MB NPU
flow VOICE_AI mixed_with_float Qualcomm® SA8255P 128.354 ms 2 - 12 MB NPU
flow VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-9075 72.859 ms 1 - 5 MB NPU
flow VOICE_AI mixed_with_float Qualcomm® Dragonwing™ IQ-X7181 74.697 ms 2 - 2 MB NPU
flow VOICE_AI mixed_with_float Qualcomm® Dragonwing™ Q-8750 42.259 ms 2 - 10 MB NPU

License

  • The license for the original implementation of MeloTTS-ZH can be found here.

References

Community

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support