RAGU: A Multi-Step GraphRAG Engine with a Compact Domain-Adapted LLM Paper • 2607.11683 • Published Jul 13 • 149
facebook/dinov3-vitb16-pretrain-lvd1689m Image Feature Extraction • 85.7M • Updated Aug 19, 2025 • 606k • 346
google/siglip2-base-patch16-224 Zero-Shot Image Classification • 0.4B • Updated Feb 21, 2025 • 1.64M • 140
DINOv3 Collection DINOv3: foundation models producing excellent dense features, outperforming SotA w/o fine-tuning - https://arxiv.org/abs/2508.10104 • 15 items • Updated Mar 10 • 786
google/siglip2-base-patch16-naflex Zero-Shot Image Classification • 0.4B • Updated Feb 21, 2025 • 1.04M • 40
FastVLM Collection Efficient Vision Encoding for Vision Language Models • 8 items • Updated Mar 2 • 115
MobileCLIP2 Collection MobileCLIP2: Mobile-friendly image-text models with SOTA zero-shot capabilities trained on DFNDR-2B • 30 items • Updated Apr 23 • 64
view article Article SigLIP 2: A better multilingual vision language encoder +1 ariG23498, merve, qubvel-hf • Feb 21, 2025 • 228
SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features Paper • 2502.14786 • Published Feb 20, 2025 • 169
google/siglip2-so400m-patch16-naflex Zero-Shot Image Classification • 1B • Updated Feb 21, 2025 • 379k • 89