Models to try These are some model I want to try / don't have the hardware for unsloth/Qwen3.6-27B-MTP-GGUF Image-Text-to-Text • 27B • Updated May 26 • 815k • 1.32k deepseek-ai/DeepSeek-R1 Text Generation • 684B • Updated Mar 27, 2025 • 1.02M • • 14.3k
Collection For My DeepSeek R1 Qwen Distill 7B Retraining This is the collection of datasets I used / am going to use to train my DeepSeek R1 Distill Qwen 7B-based model. PrimeIntellect/SYNTHETIC-1 Viewer • Updated Feb 21, 2025 • 1.99M • 2.4k • 64 open-thoughts/OpenThoughts3-1.2M Viewer • Updated Jun 9, 2025 • 1.2M • 23.4k • 264 tokyotech-llm/Swallow-Nemotron-Post-Training-Dataset-v1 Viewer • Updated Feb 21 • 8.84M • 579 • 6 BAAI/Infinity-Instruct Viewer • Updated Dec 4, 2025 • 21.9M • 2.06k • 765
Models to try These are some model I want to try / don't have the hardware for unsloth/Qwen3.6-27B-MTP-GGUF Image-Text-to-Text • 27B • Updated May 26 • 815k • 1.32k deepseek-ai/DeepSeek-R1 Text Generation • 684B • Updated Mar 27, 2025 • 1.02M • • 14.3k
Collection For My DeepSeek R1 Qwen Distill 7B Retraining This is the collection of datasets I used / am going to use to train my DeepSeek R1 Distill Qwen 7B-based model. PrimeIntellect/SYNTHETIC-1 Viewer • Updated Feb 21, 2025 • 1.99M • 2.4k • 64 open-thoughts/OpenThoughts3-1.2M Viewer • Updated Jun 9, 2025 • 1.2M • 23.4k • 264 tokyotech-llm/Swallow-Nemotron-Post-Training-Dataset-v1 Viewer • Updated Feb 21 • 8.84M • 579 • 6 BAAI/Infinity-Instruct Viewer • Updated Dec 4, 2025 • 21.9M • 2.06k • 765