ukisai/Swift-Qwen3.8-27B-GGUF Image-Text-to-Text ⢠27B ⢠Updated about 14 hours ago ⢠176k ⢠390
view post Post 4236 š Introducing Halo 1.0Today, we are open-sourcing Halo, the training framework we use to train every model at White Circle.It comes with: š§ Full post-training stack: SFT, DPO/KTO/SMPO, reward modeling, GRPO, distillationš¤ Async multi-turn RL with vLLM/SGLang rollouts and sandboxed tool useā” ~2.8Ć TRL throughput on 8Ć B300 (EP+FSDPv2, FA4, fp8/fp4)š¤ Dense HF models + 15 MoE families (Qwen, GLM, Mistral, DeepSeek-V4ā¦)š ļø One halo command, prebuilt Docker images, and docs for humans and agentsš» https://github.com/whitecircle/haloTry it and tell us what you're training See translation 1 reply Ā· š 9 9 ā¤ļø 4 4 + Reply
peculiar-ragdoll/Sharp-Spark-X2.5-4B-GGUF Text Generation ⢠4B ⢠Updated 4 days ago ⢠3.74k ⢠65
dealignai/Bonsai-2-27B-1bit-CRACK-GGUF Text Generation ⢠27B ⢠Updated 6 days ago ⢠27.3k ⢠78
dealignai/Bonsai-2-27B-Ternary-CRACK-GGUF Text Generation ⢠27B ⢠Updated 6 days ago ⢠52.7k ⢠137
prism-ml/Ternary-Bonsai-2-27B-gguf Text Generation ⢠27B ⢠Updated about 7 hours ago ⢠2.99M ⢠1.99k
peculiar-ragdoll/Tiel-Coder-35B-A3B-GGUF Image-Text-to-Text ⢠35B ⢠Updated 14 days ago ⢠507k ⢠336
ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF Image-Text-to-Text ⢠27B ⢠Updated 22 days ago ⢠1.47M ⢠1.63k
peculiar-ragdoll/Cyber-Tiel-Coder-35B-A3B-GGUF Image-Text-to-Text ⢠35B ⢠Updated 14 days ago ⢠24.4k ⢠60
AnonimousA/Qwen3.8-Flash-Next-REAP-320-GGUF Text Generation ⢠132B ⢠Updated 15 days ago ⢠33.2k ⢠20
deepseek-ai/DeepSeek-V4.1-Flash Image-Text-to-Text ⢠763B ⢠Updated 14 days ago ⢠606k ⢠⢠3.69k
peculiar-ragdoll/Cyber-Tiel-Coder-35B-A3B-GGUF-MTP Image-Text-to-Text ⢠36B ⢠Updated 6 days ago ⢠56.9k ⢠130
peculiar-ragdoll/Tiel-Coder-35B-A3B-GGUF-MTP Image-Text-to-Text ⢠0.4B ⢠Updated 14 days ago ⢠1.2M ⢠212