Tested the model on coding with OpenCode and agentic work

#32
by curiousily - opened

No. Summarize here.

This is the summary of his video. No need to click on the link.

On complex coding tasks, GM 5.3 Flash delivers surprisingly strong results building browser games and a full newsletter app with drag-and-drop email builder and drip sequences. It is arguably the best model available in the 300-billion-parameter range if you can self-host it. Its biggest weakness is inference speed: over OpenRouter throughput fluctuates wildly between 4 and 60 tokens per second, making it often too slow and unstable for daily coding work. That said, results were less impressive on one-shot HTML/CSS/JavaScript generation without reasoning, where the model sometimes shipped underwhelming output and occasionally needed "babysitting" with continuations to finish.

After a couple of weeks I find it excellent

  • Very good 'reasoning' - short on simple prompts
  • Gave correct analysis on a couple difficult cases where ds4f failed
  • Very nice to work with, pushes back at user mistakes more than most
  • Good long term recall, goal rememering
  • Good user constraint and situational awareness
  • A bit slower than ds4f

Hard to pick my favorite.

Sign up or log in to comment