Rhett1 commited on
Commit
db538fc
·
verified ·
1 Parent(s): 7a0e7c4

Upload svd_llm2/unheal/llama2-13b-base/README.md with huggingface_hub

Browse files
svd_llm2/unheal/llama2-13b-base/README.md ADDED
@@ -0,0 +1,20 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ license: llama2
3
+ base_model: NousResearch/Llama-2-13b-hf
4
+ tags: [svd, low-rank, model-compression, llama2]
5
+ ---
6
+
7
+ # SVD-LLM V2 (NAACL 2025, official code unreleased; our implementation of Alg. 1+2) @ 7.74B, UNHEALED (LLaMA-2-13B base)
8
+
9
+ - **UNHEALED checkpoint**: raw post-compression weights, NO distillation. The KD-healed
10
+ counterpart is at `../heal/llama2-13b-base/` (same `ranks.json`; weights differ).
11
+ - Factored params: **7,740,277,760** (budget 7.7403B, matched to the TR-3 llama2-13b row;
12
+ SVD-LLM V2 overall keep-ratio 0.5740 across the factored projections).
13
+ - probe MMLU (limit 40/subtask, n=2280, seed 1234, fp16 — measured as the step-0 eval of
14
+ the sibling heal run): **26.14%**, i.e. essentially chance level; FP16 dense 13B base
15
+ teacher on the same probe: 54.25%. Compression alone, without healing, destroys base-model
16
+ capability at this budget.
17
+
18
+ Load with `lowrank_linear.py` in this dir (identical recipe to the sibling dirs).
19
+ Compression-provenance details (source revision, sha256 of every artifact, allocation
20
+ files) are in `provenance.json`.