File size: 1,397 Bytes
c3f9ad5
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
35
36
37
---
license: apache-2.0
tags:
- tts
- arabic
- kokoro
- styletts2
- checkpoint
---

# Sofelia TTS — Training Checkpoints (private)

Raw StyleTTS2 Stage-2 checkpoints for the Palestinian Arabic Kokoro fine-tune.
These are training-state checkpoints (`{'net': {...module.prefixed...}, 'optimizer', 'epoch', ...}`),
**not** the inference model. For inference use the public model
[hamdallah/Sofelia-TTS-82M](https://huggingface.co/hamdallah/Sofelia-TTS-82M).

## Checkpoints

| File | What | Val |
|---|---|---|
| `checkpoints/sirin_eliaa_final_epoch2nd.pth` | **Eliaa** single-voice final (powers the public model) | 0.463 (Sirin val set) |
| `checkpoints/multispeaker_final_epoch2nd.pth` | multispeaker base (Kore + Sirin) | 0.446 (mixed val set) |
| `checkpoints/stage1_first_stage.pth` | Stage-1 acoustic checkpoint | — |

## Provenance

Kokoro-82M → Stage 1 (multispeaker, Palestinian Arabic) → Stage 2 (multispeaker)
→ Stage 2 continuation on the single human voice (Eliaa/Sirin), stopped at the
validation plateau / overfit point. `training/` holds the configs, the dialect
frontend (`prepare_arabic.py`, `ar_lexicon.json`), and the voicepack/inference
scripts needed to reproduce or extend.

To convert a checkpoint to the public Kokoro inference format, strip the
`module.` prefixes from each `net` component and load via `kokoro.KModel`
(see `training/test_inference_arabic.py`).