mlx-community/Qwen2-VL-2B-Instruct-4bit

This model was converted to MLX format from Qwen/Qwen2-VL-2B-Instruct using mlx-vlm version 0.0.13. Refer to the original model card for more details on the model.

Use with mlx

pip install -U mlx-vlm
python -m mlx_vlm.generate --model mlx-community/Qwen2-VL-2B-Instruct-4bit --max-tokens 100 --temp 0.0
Downloads last month
17,168
Safetensors
Model size
2B params
Tensor type
U32
Β·
F16
Β·
MLX
Hardware compatibility
Log In to add your hardware

Quantized

Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Spaces using mlx-community/Qwen2-VL-2B-Instruct-4bit 4