Quant Experts: Token-aware Adaptive Error Reconstruction with Mixture of Experts for Large Vision-Language Models Quantization Paper • 2602.24059 • Published Feb 27 • 1
Running on Zero Agents 420 Stable Diffusion 3.5 Large Turbo 🏃 420 Generate images fast with SD3.5 turbo