|
Download README.md from espnet/s3prl_adapter_model: direct link, hf CLI and curl.
- Browser
- Download file 515 Bytes
-
https://huggingface.co/espnet/s3prl_adapter_model/resolve/main/README.md
- Command line
-
hf download hf://espnet/s3prl_adapter_model/README.md
-
curl -L -o README.md https://huggingface.co/espnet/s3prl_adapter_model/resolve/main/README.md
515 Bytes
| pipeline_tag: automatic-speech-recognition | |
| ## Usage | |
| ```python | |
| import librosa | |
| from espnet2.bin.asr_inference import Speech2Text | |
| speech2text = Speech2Text.from_pretrained(model_tag="espnet/s3prl_adapter_model") | |
| # librosa resamples and mixes to one channel, so any file works; 16000 is | |
| # what nearly every espnet recogniser is trained on - check this model's | |
| # config if its audio is not 16 kHz | |
| speech, rate = librosa.load("audio.wav", sr=16000, mono=True) | |
| text, *_ = speech2text(speech)[0] | |
| print(text) | |
| ``` | |