Collections
Discover the best community collections!
Collections trending this week
-
unsloth/Qwen3-VL-30B-A3B-Instruct-GGUF
Image-Text-to-Text • 31B • Updated • 26.1k • 118 -
unsloth/Qwen3-VL-30B-A3B-Thinking-GGUF
Image-Text-to-Text • 31B • Updated • 5.6k • 47 -
unsloth/Qwen3-VL-4B-Instruct-GGUF
Image-Text-to-Text • 4B • Updated • 330k • 75 -
unsloth/Qwen3-VL-4B-Thinking-GGUF
Image-Text-to-Text • 4B • Updated • 8.25k • 25
-
WAON: Large-Scale and High-Quality Japanese Image-Text Pair Dataset for Vision-Language Models
Paper • 2510.22276 • Published • 3 -
llm-jp/WAON-Bench
Viewer • Updated • 1.87k • 171 • 2 -
llm-jp/waon-siglip2-base-patch16-256
Zero-Shot Image Classification • 0.4B • Updated • 765 • 1 -
llm-jp/WAON
Updated • 133 • 8
-
Teaching Pretrained Language Models to Think Deeper with Retrofitted Recurrence
Paper • 2511.07384 • Published • 21 -
smcleish/Recurrent-Llama-3.2-train-recurrence-32
Text Generation • 1B • Updated • 1.5k • 1 -
smcleish/Recurrent-Llama-3.2-train-recurrence-16
Text Generation • 1B • Updated • 295 -
smcleish/Recurrent-Llama-3.2-train-recurrence-8
Text Generation • 1B • Updated • 165
-
ByteDance/Ouro-1.4B
Text Generation • 1B • Updated • 25.5k • 163 -
ByteDance/Ouro-1.4B-Thinking
Text Generation • 1B • Updated • 17.6k • 54 -
ByteDance/Ouro-2.6B
Text Generation • 3B • Updated • 12.2k • 95 -
ByteDance/Ouro-2.6B-Thinking
Text Generation • 3B • Updated • 15.1k • 158
-
cerebras/Qwen3-Coder-REAP-363B-A35B-FP8
Text Generation • 363B • Updated • 66 • 17 -
cerebras/Qwen3-Coder-REAP-246B-A35B-FP8
Text Generation • 246B • Updated • 49 • 22 -
cerebras/Qwen3-Coder-REAP-363B-A35B
Text Generation • 363B • Updated • 143 • 6 -
cerebras/Qwen3-Coder-REAP-246B-A35B
Text Generation • 246B • Updated • 90 • 8
-
Teaching Pretrained Language Models to Think Deeper with Retrofitted Recurrence
Paper • 2511.07384 • Published • 21 -
smcleish/Recurrent-Llama-3.2-train-recurrence-32
Text Generation • 1B • Updated • 1.5k • 1 -
smcleish/Recurrent-Llama-3.2-train-recurrence-16
Text Generation • 1B • Updated • 295 -
smcleish/Recurrent-Llama-3.2-train-recurrence-8
Text Generation • 1B • Updated • 165
-
unsloth/Qwen3-VL-30B-A3B-Instruct-GGUF
Image-Text-to-Text • 31B • Updated • 26.1k • 118 -
unsloth/Qwen3-VL-30B-A3B-Thinking-GGUF
Image-Text-to-Text • 31B • Updated • 5.6k • 47 -
unsloth/Qwen3-VL-4B-Instruct-GGUF
Image-Text-to-Text • 4B • Updated • 330k • 75 -
unsloth/Qwen3-VL-4B-Thinking-GGUF
Image-Text-to-Text • 4B • Updated • 8.25k • 25
-
ByteDance/Ouro-1.4B
Text Generation • 1B • Updated • 25.5k • 163 -
ByteDance/Ouro-1.4B-Thinking
Text Generation • 1B • Updated • 17.6k • 54 -
ByteDance/Ouro-2.6B
Text Generation • 3B • Updated • 12.2k • 95 -
ByteDance/Ouro-2.6B-Thinking
Text Generation • 3B • Updated • 15.1k • 158
-
WAON: Large-Scale and High-Quality Japanese Image-Text Pair Dataset for Vision-Language Models
Paper • 2510.22276 • Published • 3 -
llm-jp/WAON-Bench
Viewer • Updated • 1.87k • 171 • 2 -
llm-jp/waon-siglip2-base-patch16-256
Zero-Shot Image Classification • 0.4B • Updated • 765 • 1 -
llm-jp/WAON
Updated • 133 • 8
-
cerebras/Qwen3-Coder-REAP-363B-A35B-FP8
Text Generation • 363B • Updated • 66 • 17 -
cerebras/Qwen3-Coder-REAP-246B-A35B-FP8
Text Generation • 246B • Updated • 49 • 22 -
cerebras/Qwen3-Coder-REAP-363B-A35B
Text Generation • 363B • Updated • 143 • 6 -
cerebras/Qwen3-Coder-REAP-246B-A35B
Text Generation • 246B • Updated • 90 • 8