🐻 Kodiak-v0.2-1B is out: an open 1B encoder that makes decisions instead of writing text.
Give it a state and typed questions. You get back calibrated answers in one forward pass, and it says "can't tell" when it doesn't know.
On tasks it was never trained for (our frozen eval set v0.2): • 0.689 accuracy (3-run mean) vs 0.688 for Qwen3-8B, at ~40× the speed • calibration error 0.085 vs 0.293 • accuracy mode (3 models averaged): 0.706
New in v0.2: grounding checks, picking an assistant's next tool step from API specs, claim verification, and product relevance.