Audio-Text-to-Text
Transformers
Safetensors
Chinese
English
edgeinstant
feature-extraction
audio
speech-recognition
speech-translation
audio-question-answering
custom_code
Instructions to use chenjz24/EdgeIn-v1 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use chenjz24/EdgeIn-v1 with Transformers:
# pip install -U transformers accelerate # Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("chenjz24/EdgeIn-v1", trust_remote_code=True, device_map="auto") - Notebooks
- Google Colab
- Kaggle
File size: 818 Bytes
f74eb65 | 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 | {
"source": "/default-filesys/workspace/chenjunzhe/EdgeInstant-1.5b/runs/bilingual_s5_full/final",
"output": "/default-filesys/workspace/chenjunzhe/EdgeInstant-1.5b/runs/bilingual_s5_full/final_compact",
"precision": "torch.bfloat16",
"preserved_modalities": [
"text",
"audio_input",
"image",
"video",
"speech_output"
],
"source_parameters": 2031972995,
"exported_parameters": 2031710851,
"removed_weights": [
"codec.quantizer.rvq_first.input_proj.weight",
"codec.quantizer.rvq_rest.input_proj.weight"
],
"source_weight_bytes": 5677248858,
"exported_weight_bytes": 4115251826,
"equivalence_reference": "Source loaded with dtype=auto, as in run_mmau.py; align and audio_special remain FP32.",
"cache_implementation": "dynamic",
"recurrent_decode": "cuda_graph"
}
|