MiMo-V2.5-Pro

No longer available on HF due to storage restrictions - archived here

See MiMo-V2.5-Pro in action: demonstration videos

Tested with an M3 Ultra 512 GiB and M4 Max 128 GiB RAM using Inferencer app distributed compute

  • Distributed inference: ~13 tokens/s @ 1000 tokens ~450 GiB / ~67 GiB (debug build)
Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for inferencerlabs/MiMo-V2.5-Pro-MLX-Q4.3-INF

Quantized
(9)
this model