view article Article 150k Tokens Deep: Why Context Recall Beats Compaction by up to 8x dacorvo ⢠3 days ago ⢠2
view article Article Introducing Storage Buckets on the Hugging Face Hub +10 Wauplin, coyotte508, XciD, victor, julien-c, lhoestq, pierric, Sylvestre, hlarcher, rajatarya, seanses, assafvayner ⢠Mar 10 ⢠199
view article Article Welcome GPT OSS, the new open-source model family from OpenAI! +10 reach-vb, pcuenq, lewtun, clem, Rocketknight1, clefourrier, celinah, Wauplin, marcsun13, pagezyhf, ahadnagy, joaogante ⢠Aug 5, 2025 ⢠514
view article Article Introducing the AMD 5th Gen EPYC⢠CPU mohitsha, mfuntowicz ⢠Oct 10, 2024 ⢠7
view article Article Sensitivity Aware Mixed Precision Quantization V1 badaoui ⢠Jun 13, 2025 ⢠29
view article Article Hugging Face and AMD partner on accelerating state-of-the-art models for CPU and GPU platforms juliensimon ⢠Jun 13, 2023 ⢠5
view article Article Building Tensors from Scratch in Rust (Part 1.1): Core Structure and Indexing KeighBee ⢠Jun 12, 2025 ⢠8
view article Article Introducing multi-backends (TRT-LLM, vLLM) support for Text Generation Inference mfuntowicz, hlarcher ⢠Jan 16, 2025 ⢠77