lmstudio-community/Qwen3-Coder-30B-A3B-Instruct-MLX-4bit Text Generation • 31B • Updated Jul 31, 2025 • 143k • 40
view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego • 21 days ago • 134
ApexAgents-SkyRL-Recipe Collection Checkpoints and eval traces from "Training frontier knowledge work agents: A 397B open recipe with SkyRL" • 4 items • Updated 23 days ago • 4
view article Article Training and Finetuning Multi-Vector Embedding Models with Sentence Transformers tomaarsen • 29 days ago • 144
nvidia/NVIDIA-Nemotron-Labs-Teacher-Competition-Coding Text Generation • 561B • Updated Aug 14 • 666 • 6
view article Article Meta is back with Muse Glimmer: local, agentic, multimodal, and open source +2 pcuenq, merve, burtenshaw, ariG23498 • Aug 10 • 113