AffectF5 / README.md
ScreamingTony's picture
Keep evaluation details on GitHub and simplify model card
e4c5138 verified
|
Raw History Blame Contribute Delete
1.89 kB
---
license: cc-by-nc-4.0
language:
- en
library_name: pytorch
pipeline_tag: text-to-speech
base_model: SWivid/F5-TTS
tags:
- text-to-speech
- f5-tts
- emotion
- arousal
- voice-cloning
---
# AffectF5 C7
[![GitHub](https://img.shields.io/badge/GitHub-ScreamingTony%2FAffectF5-181717?logo=github)](https://github.com/ScreamingTony/AffectF5)
**想让它开口说话?[前往 GitHub:获取代码、安装指南与 Demo](https://github.com/ScreamingTony/AffectF5)。**
AffectF5 为 F5-TTS v1 Base 增加五类 Emotion 与连续 Arousal 控制。该仓库只保存轻量 adapter;模型运行时仍需下载官方 F5-TTS v1 Base。
## 控制接口
- Emotion:Neutral、Happy、Sad、Angry、Surprise。
- Arousal:`[-1, 1]`。
- Emotion 与 Arousal 分别建模,以轻量适配控制待生成区域,同时兼顾文本准确度与音色保持。
- 运行配置见 `config.yaml`,参考静音裁剪可在推理时开启。
## 使用
```bash
git clone https://github.com/ScreamingTony/AffectF5.git
cd AffectF5
pip install -e ".[demo]"
affect-f5 --help
```
## 权重内容
`affectf5_c7_adapter.safetensors` 包含:
- 轻量 LoRA 适配参数;
- Emotion/Arousal 条件编码与生成区域控制模块。
不包含官方底模、优化器状态或训练数据。
## 许可与限制
Adapter 采用 CC BY-NC 4.0,仅供非商业用途。当前主要验证英文;自动情感指标不能替代人工听测。
本项目基于 [F5-TTS](https://github.com/SWivid/F5-TTS),保留官方底模归属与许可限制。代码遵循 MIT,详见 [NOTICE](https://github.com/ScreamingTony/AffectF5/blob/main/NOTICE.md) 与 [模型许可说明](https://github.com/ScreamingTony/AffectF5/blob/main/MODEL_LICENSE.md)。必要结构和采样配置保留在 `config.yaml` 中,以维持权重加载和推理兼容性。
## 联系
[ye.chen2@siat.ac.cn](mailto:ye.chen2@siat.ac.cn)