Inference Provider

VERIFIED
70,949 monthly requests

AI & ML interests

None defined yet.

Recent Activity

Articles

danielhanchen 
posted an update 3 days ago
view post
Post
799
We compared 1-bit Kimi K3 to Claude Opus 5 and GPT 5.6. 🤯

We gave 4 models the same prompt: Create a glass aquarium whose side panel develops a visible crack and then bursts...

1-bit Kimi K3 GGUF ran locally on 4x B200s at 36 tok/s.

GGUF: unsloth/Kimi-K3-GGUF
GitHub repo: https://github.com/unslothai/unsloth
  • 1 reply
·
danielhanchen 
posted an update 5 days ago
view post
Post
3636
Kimi K3 can now be run locally! ✨

The 1-bit model retains ~78.9% accuracy after we shrunk it from 1.56TB to 594GB (-62% size).

Run on a Mac Studio connected with 128GB RAM device. Kimi K3 is the strongest open model to date.

GGUF: unsloth/Kimi-K3-GGUF
Guide: https://unsloth.ai/docs/models/kimi-k3
  • 5 replies
·
danielhanchen 
posted an update 14 days ago
view post
Post
4723
Introducing Unsloth for AMD 🚀
You can now train & run LLMs on your AMD hardware

• We collaborated with AMD to enable you to train & run 500+ models on AMD GPUs
• Works on Windows, WSL, Linux
• Train Qwen, Gemma on just 3GB VRAM

GitHub: https://github.com/unslothai/unsloth
Blog + Guide: https://unsloth.ai/docs/basics/amd
  • 3 replies
·
danielhanchen 
posted an update 16 days ago
danielhanchen 
posted an update 20 days ago
danielhanchen 
posted an update 24 days ago
danielhanchen 
posted an update 27 days ago
stvincent-cohere 
published an article 27 days ago
view article
Article

Meet Cohere Transcribe Arabic

CohereLabs
10
danielhanchen 
posted an update about 1 month ago
view post
Post
3367
1-bit GLM-5.2 GGUF vs. Claude 4.8 Opus vs. GPT-5.5

We gave 3 models the same prompt and compared one-shot outputs.

The 1-bit GLM-5.2 GGUF ran locally on a Mac Studio M3 Ultra with 256GB RAM at ~21.6 tok/s.

Which output do you like best?
GGUF: unsloth/GLM-5.2-GGUF
  • 3 replies
·
danielhanchen 
posted an update about 2 months ago