Instructions to use squaredcuber/forge-optimizer-qwen3.6-35b-a3b with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Local Apps Settings
- Unsloth Desktop
forge-optimizer
A LoRA fine-tune of Qwen3.6-35B-A3B (MoE) that turns unoptimized backend code into efficient, correct code — the skill measured by forger-bench, an efficiency-aware benchmark for AI-generated InsForge SDK code.
Trained with Unsloth (bf16 LoRA) + an agentic GRPO loop adopting CUDA-Agent (arXiv 2602.24286): the model writes a solution, the forger-bench grader runs+verifies+ measures real server metrics, and a discrete milestone reward (-1 incorrect/scaleBug, 1 wasteful, 2 beats-naive, 3 near-optimal) drives RL.
Contamination control
Never trained on a sealed test task; held-out concepts measure optimization skill vs template memorization.
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support