Community Blog & Articles
NEW Articles from Team or Enterprise organizations will get promoted to the main section. Kimi K3 Model Overview: 2.8T Parameters, MXFP4 Quantization, and What the Open Weights Mean for the Community
ResterChed
• • 101
Introducing Cosmos 3 Edge
Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 Diffusers
Aether-7B-5Attn: A 100% Open-Source Sovereign Foundation Model — and a Controlled Experiment in Heterogeneous Attention
FINAL-Bench
• • 21
NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval
nvidia
• • 56
Be Ready Before the Attack: A Practical Guide to Self-Hosting an Open Model for Cyber Defense
jeffboudier
• • 14
Hugging Face on AMD Instinct MI455X: First Transformers Results
badaoui
• • 13
POCKET: a 35-billion-parameter model that runs on your iPhone — and on your PC with no GPU
FINAL-Bench
• • 11
KV Caching Explained: Optimizing Transformer Inference Efficiency
not-lain
• • 380
Uncensor any LLM with abliteration
mlabonne
• • 882
The influx of specialist models on the Open SLM Leaderboard
Banaxi-Tech
• • 7
One Adapter, Both Modalities: Field Notes from Building and Serving a Multimodal Reranker
lightonai
• • 17
Introduction to State Space Models (SSM)
lbourdois
• • 238
From GRPO to DAPO and GSPO: What, Why, and How
NormalUhr
• • 133
Code a simple RAG from scratch
ngxson
• • 366
Tokenization is Killing our Multilingual LLM Dream
Introducing North Mini Code: Cohere’s First Model For Developers
J-Space: Yet Another LLM Mind Reader?
dlouapre
• • 34
We Just Surgically Changed What Your Model Believes
ApolloRaines
• • 3
ArmBench-LLM 1.0: Benchmarking LLMs on Armenian Language Tasks