VOXCPM-2: Finetuning for a regional language

#18
by bardicsaucer - opened

Let’s say you are targeting a specific regional language that isn't widely known, and you have about 50 hours of audio data. How and where could I find baseline benchmarks or metrics for training purposes? Is the tuning completely a process of trial and error? For context, the language I am targeting is Kannada.

Sign up or log in to comment