Nanbeige4.1-3B: A Small General Model that Reasons, Aligns, and Acts Paper • 2602.13367 • Published Feb 13 • 37
Nanbeige4.2-3B: Unlocking Agentic Capabilities in a Compact Model Paper • 2607.22083 • Published 13 days ago • 6
Inkling Collection Inkling by Thinking Machines, self-quantized to GGUF by Atomic Chat. 975B MoE (41B active), 1M context. • 1 item • Updated 23 days ago • 1
Hy3 Collection Atomic Chat GGUF builds of Tencent Hy3 (295B MoE): coding-optimized imatrix quants • 1 item • Updated 25 days ago • 2
DFlash speculative drafts Collection DFlash block-diffusion draft heads (GGUF). Attach to any target GGUF to speed up llama.cpp. • 8 items • Updated 26 days ago • 1
Laguna XS 2.1 Collection poolside Laguna XS 2.1 (33B/3B MoE) quantized by Atomic Chat - GGUF (llama.cpp) and MLX (Apple Silicon) ladders. • 4 items • Updated Jul 2 • 2
Ornith 1.0 Collection DeepReinforce's Ornith 1.0 agentic-coding models, self-quantized to GGUF by Atomic Chat with a per-tensor importance matrix. Runs fully offline. • 10 items • Updated Jul 1 • 6
Qwen 3.6 UDT MTP Collection Dynamic-imatrix GGUF quants of Qwen 3.6 27B & 35B-A3B. TurboQuant3 KV + shared-model NextN ready. • 2 items • Updated May 14 • 5
Gemma 4 Assistant GGUF Collection Gemma 4 MTP assistant drafters as GGUF (F16/Q8_0/Q5_K_M/Q4_K_M/Q4_K_S). Speculative-decoding heads for the atomic-llama-cpp-turboquant fork. • 4 items • Updated May 7 • 13