Automatic Speech Recognition
Transformers
TensorBoard
Safetensors
msp
Generated from Trainer
custom_code
Instructions to use MahmoodAnaam/MSP with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use MahmoodAnaam/MSP with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("automatic-speech-recognition", model="MahmoodAnaam/MSP", trust_remote_code=True)# Load model directly from transformers import AutoModelForCTC model = AutoModelForCTC.from_pretrained("MahmoodAnaam/MSP", trust_remote_code=True, device_map="auto") - Notebooks
- Google Colab
- Kaggle
File size: 1,178 Bytes
dc0059b | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 29 30 31 32 33 34 35 36 37 38 39 40 41 42 43 44 45 | {
"auto_map": {
"AutoProcessor": "processing_msp.MSPProcessor"
},
"feature_extractor": {
"auto_map": {
"AutoFeatureExtractor": "feature_extraction_msp_audio.MSPAudioFeatureExtractor",
"AutoProcessor": "processing_msp.MSPProcessor"
},
"do_normalize": true,
"feature_extractor_type": "MSPAudioFeatureExtractor",
"feature_size": 1,
"padding_side": "right",
"padding_value": 0,
"return_attention_mask": true,
"sampling_rate": 16000
},
"processor_class": "MSPProcessor",
"video_processor": {
"auto_map": {
"AutoProcessor": "processing_msp.MSPProcessor",
"AutoVideoProcessor": "video_processing_msp_visual.MSPVisualVideoProcessor"
},
"crop_size": {
"height": 88,
"width": 88
},
"do_center_crop": true,
"do_convert_rgb_to_grayscale": true,
"do_normalize": true,
"do_rescale": true,
"do_resize": true,
"image_mean": 0.421,
"image_std": 0.165,
"resample": 2,
"rescale_factor": 0.00392156862745098,
"return_metadata": false,
"size": {
"height": 96,
"width": 96
},
"video_processor_type": "MSPVisualVideoProcessor"
}
}
|