RoboICL: Embodied In-Context Learning with GPT-6 Astra Paper • 2609.34261 • Published 2 days ago • 13
view post Post 1272 Pangu Pro MoE 🔥 Huawei's first open model!Paper: Pangu Pro MoE: Mixture of Grouped Experts for Efficient Sparsity (2505.21411)Model: https://gitcode.com/ascend-tribe/pangu-pro-moe-model✨ MoGE: Mixture of Grouped Experts✨ 16B activated params - 48 layers✨ Trained on 15T tokens✨ Natively optimized for Ascend hardware See translation 1 reply · 🔥 4 4 + Reply
Pangu Pro MoE: Mixture of Grouped Experts for Efficient Sparsity Paper • 2505.21411 • Published May 27, 2025 • 17
Pangu Ultra MoE: How to Train Your Big MoE on Ascend NPUs Paper • 2505.04519 • Published May 7, 2025 • 5