ThakiCloud/eg2-text-q4
Feature Extraction • 0.1B • Updated
Quantized EmbeddingGemma 2 text towers that query an existing BF16 index without re-embedding the corpus.
Note 4-bit text tower, 152,669,232 B of weights. Keeps 0.936 of the BF16 top-10 neighborhood (BF16 = 1.000); margin 0.1712 vs BF16 0.1817. Internally measured on a private frozen set; serving not validated.
Note Mixed-precision text tower (GPTQ 4-bit transformer, 2-bit vocabulary with 30% of rows at 4-bit), 129,532,796 B — 85% of eg2-text-q4. Beats uniform 3-bit on every held-out slice; within 25% of the 3→4-bit gap of the Q4 recipe. Internally measured; serving not validated.