Kimi K3 Model Overview: 2.8T Parameters, MXFP4 Quantization, and What the Open Weights Mean for the Community ResterChed • 11 days ago • 177
POCKET: a 35-billion-parameter model that runs on your iPhone — and on your PC with no GPU FINAL-Bench • 6 days ago • 12
Be Ready Before the Attack: A Practical Guide to Self-Hosting an Open Model for Cyber Defense jeffboudier • 8 days ago • 16
ECMWF's AI forecasting model is open source: now let's make it easy to run. hugging-science • about 7 hours ago • 6
FLUX 3 Model Overview: Multimodal Flow Models for Image, Video, Audio, and Action Prediction ResterChed • 5 days ago • 5
NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval nvidia • 12 days ago • 57
Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 Diffusers nvidia • 11 days ago • 80
AI, Physical AI, World Models, VLA, VLM, and Other Terms We Should Stop Mixing Together barakor • May 17 • 5
Kimi K3 Model Overview: 2.8T Parameters, MXFP4 Quantization, and What the Open Weights Mean for the Community ResterChed • 11 days ago • 177
POCKET: a 35-billion-parameter model that runs on your iPhone — and on your PC with no GPU FINAL-Bench • 6 days ago • 12
Be Ready Before the Attack: A Practical Guide to Self-Hosting an Open Model for Cyber Defense jeffboudier • 8 days ago • 16
ECMWF's AI forecasting model is open source: now let's make it easy to run. hugging-science • about 7 hours ago • 6
FLUX 3 Model Overview: Multimodal Flow Models for Image, Video, Audio, and Action Prediction ResterChed • 5 days ago • 5
NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval nvidia • 12 days ago • 57
Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 Diffusers nvidia • 11 days ago • 80
AI, Physical AI, World Models, VLA, VLM, and Other Terms We Should Stop Mixing Together barakor • May 17 • 5