DeepSeek-V4.1-Flash: Pushing the Limits of KV Cache Compression Paper • 2609.19969 • Published 17 days ago • 206
view article Article Jun Kim, oMLX creator and maintainer, joins Hugging Face to support the MLX community +3 pcuenq, lysandre, victor, julien-c, Jundot • 12 days ago • 84
view article Article Navigating the RLHF Landscape: From Policy Gradients to PPO, GAE, and DPO for LLM Alignment NormalUhr • Feb 11, 2025 • 134
view reply awesome high-quality blogpost! 🔥 great use of different fun elements, very nice to have embedded gradio space as well 🤗
view article Article ColPali: Efficient Document Retrieval with Vision Language Models 👀 manu • Jul 5, 2024 • 333
view article Article DeepSeek-R1 Dissection: Understanding PPO & GRPO Without Any Prior Reinforcement Learning Knowledge NormalUhr • Feb 7, 2025 • 302
view article Article Building Moon Bot: A Slack-Native Coding Agent Backed by HuggingFace Buckets huggingface • Jun 24 • 56
view article Article How to Use NVIDIA Warp and MjWarp to Accelerate Robotics Simulation and Learning Workflows nvidia • 10 days ago • 54
view article Article Predicting the Effects of Mutations on Protein Function with ESM-2 AmelieSchreiber • Dec 13, 2023 • 29
view article Article How UK AISI and EvalEval Are Making Benchmark Results Reproducible +7 evijit, j-chim, deeplumiere, srishtiy, wmmkennedy, irenesolaiman, mcfadyen-aisi, lynn-aisi, coz-aisi • 12 days ago • 27
view article Article NVIDIA Kumo Tabular Sets a New Accuracy-Efficiency Frontier for Tabular Prediction nvidia • 5 days ago • 69