|
Download README.md from ICTNLP/HybridEmo: direct link, hf CLI and curl.
- Browser
- Download file 1.34 kB
-
https://huggingface.co/ICTNLP/HybridEmo/resolve/main/README.md
- Command line
-
hf download hf://ICTNLP/HybridEmo/README.md
-
curl -L -o README.md https://huggingface.co/ICTNLP/HybridEmo/resolve/main/README.md
1.34 kB
metadata
language:
- en
tags:
- text-to-speech
- speech
- emotional-speech
- instruction-following
- emotion-trajectory
- emotion-blending
base_model: FunAudioLLM/Fun-CosyVoice3-0.5B-2512
HybridEmo
HybridEmo is an instruction-following multi-emotion text-to-speech model developed by post-training CosyVoice 3. It supports sequential emotion trajectories and simultaneous emotion blending within a single utterance.
Resources
Usage
Please refer to the GitHub repository for environment setup and inference instructions.
Citation
If you use HybridEmo in your research, please cite our paper:
@misc{zhou2026sequentialtrajectoriessimultaneousblending,
title={Sequential Trajectories and Simultaneous Blending: Multi-Emotion Modeling for Instruction-Following TTS},
author={Yan Zhou and Yun Hong and Yang Feng},
year={2026},
eprint={2608.30325},
archivePrefix={arXiv},
primaryClass={cs.CL},
url={https://arxiv.org/abs/2608.30325},
}