V100-Validated Quantizations Collection Models built for and runtime-validated on NVIDIA Tesla V100/SM70, with frozen recipes, checksums, benchmarks, and explicit quality limits. • 8 items • Updated 26 days ago • 4
How Lossless Is Lossless Speculative Decoding? The Role of Numerical Precision in Orthrus Paper • 2609.15504 • Published 22 days ago • 37
Audio Models Quantized Collection INT8, FP8 and INT4 builds of the Aniemore speech-emotion models. Load one with from_pretrained(repo, subfolder="int4"). • 9 items • Updated Aug 1 • 1
T-One Finetuned Collection T-One fine-tuned models on corporative internal expert-labeled data • 5 items • Updated Feb 26 • 1
Whisper Finetuned Collection OpenAI's Whisper models fine-tuned on corporative internal expert-labeled data • 3 items • Updated Feb 26 • 1
VulnLLM-R: Specialized Reasoning LLM with Agent Scaffold for Vulnerability Detection Paper • 2512.07533 • Published Dec 8, 2025 • 4
view article Article Transformers v5: Simple model definitions powering the AI ecosystem +2 lysandre, ArthurZ, cyrilvallez, reach-vb • Dec 1, 2025 • 315
WeDLM: Reconciling Diffusion Language Models with Standard Causal Attention for Fast Inference Paper • 2512.22737 • Published Dec 28, 2025 • 2
Kandinsky 5.0: A Family of Foundation Models for Image and Video Generation Paper • 2511.14993 • Published Nov 19, 2025 • 236
view article Article How Financial News Can Be Used to Train Good Financial Models SelmaNajih001 • Oct 8, 2025 • 10
view article Article Introducing RTEB: A New Standard for Retrieval Evaluation +4 fzliu, KennethEnevoldsen, Samoed, isaacchung, tomaarsen, fzoll • Oct 1, 2025 • 149
InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency Paper • 2508.18265 • Published Aug 25, 2025 • 222
Cosmos-Preidct1 Collection ⚠️ This collection is archived. 👉 https://huggingface.co/collections/nvidia/cosmos3 • 14 items • Updated Aug 11 • 302