FastSpeech2
FastSpeech2 是一种高效的端到端语音合成模型。相比 FastSpeech,FastSpeech2 引入了多尺度时长预测器和能量 / 基频预测分支,优化了时长预测模块并新增韵律特征建模,在合成速度和语音自然度上均有大幅提升。
Mirror Metadata
- Hugging Face repo: shadow-cann/hispark-modelzoo-fastspeech2
- Portal model id: jrc61eo19400
- Created at: 2026-05-29 09:26:03
- Updated at: 2026-09-01 16:13:07
- Category: 音频
Framework
- PyTorch
Supported OS
- Linux
- OpenHarmony
Computing Power
- Hi3403V100 SVP_NNN
- Hi3403V100 NNN
Tags
- 文本转语音
Detail Parameters
- 输入: 1x40
- 参数量: 35.266M
- 计算量: 29.162GFLOPs
Files In This Repo
- fastspeech_hifigan_en_nnn.onnx (源模型 / 源模型元数据)
- fastspeech_hifigan_en_svp_nnn.onnx (源模型 / 源模型元数据)
- fastspeech_hifigan_en.om (编译模型 / OM 元数据 / A16W8)
- fastspeech_hifigan_en_nnn.om (编译模型 / OM 元数据 / FP16)
- fastspeech_hifigan_en.onnx (源模型 / 镜像补充)
Upstream Links
- Portal card: https://gitbubble.github.io/hisilicon-developer-portal-mirror/model-detail.html?id=jrc61eo19400
- Upstream repository: https://gitcode.com/HiSpark/modelzoo/blob/master/samples/built-in/audio/FastSpeech2/README.md
- License reference: https://github.com/ming024/FastSpeech2/blob/master/LICENSE
Notes
- This repository was mirrored from the HiSilicon Developer Portal model card and local downloads captured on 2026-03-27.
- File ownership follows the portal card mapping, not just filename similarity.
- Cover image: 1722265270026243_fastspeech2.jpg
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support