view article Article Speculative Decoding in Practice: How EAGLE3 Makes LLMs Faster Without Changing Their Outputs lujangusface β’ Apr 3 β’ 9
view article Article 2x Faster on a 229B MoE: EAGLE3 Speculative Decoding for MiniMax-M2.5 lujangusface β’ Apr 9 β’ 3
cyankiwi/Qwen3-30B-A3B-Instruct-2507-AWQ-4bit Text Generation β’ 5B β’ Updated 8 days ago β’ 881k β’ 32