AI & ML interests

Deep experimentation on ablation and quanting. Goal: trim the fat on models. Question: Is all data, good data, or, is some just noise?

Recent Activity

deucebucket  updated a collection about 2 months ago
24 GB cards and CPU offload
deucebucket  updated a collection about 2 months ago
Qwen 3.6 - Cerebellum
deucebucket  updated a collection about 2 months ago
North-Mini-Code - Cerebellum
View all activity

DB-Cerebellum 's collections 8

Fits 16 GB cards
Weights under ~14 GB so a 16 GB GPU runs them fully loaded with room for context. Measured footprints on each card.
Gemma 4 - Cerebellum
24 GB cards and CPU offload
The big ones. Includes the 122B with its measured expert-offload recipe (about 18 GB VRAM plus 34 GB RAM).
Fits 16 GB cards
Weights under ~14 GB so a 16 GB GPU runs them fully loaded with room for context. Measured footprints on each card.
24 GB cards and CPU offload
The big ones. Includes the 122B with its measured expert-offload recipe (about 18 GB VRAM plus 34 GB RAM).
Gemma 4 - Cerebellum