GPUburnout-3B-75K-Chat (RETIRED)

This model is retired. It was trained in April 2026 with a buggy SFT data formatter (plain string concatenation instead of apply_chat_template), which caused loop collapse at inference and made it underperform the smaller 2B chat model on most benchmarks.

Use GPUburnout/GPUburnout-3B-75K-Chat-v2 instead.

Full postmortem: It Took Me Two Weeks to Read My Own Code

This repo is kept as a historical artifact, not for production use.

Downloads last month
63
Safetensors
Model size
3B params
Tensor type
F16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Space using GPUburnout/GPUburnout-3B-75K-Chat 1