Commit History

Upload Llama-3.2-3B-Instruct-cuda-q8.base with huggingface_hub
c111e66
verified

prabodbc commited on

Upload Llama-3.2-3B-Instruct-cuda-q4mix.base with huggingface_hub
c652bae
verified

prabodbc commited on

Re-convert with fixed rotary q/k permutation (baseRT-internal#114); NIAH-validated at 2k/8k. Q4 now uses the default-q4 quality profile (embeddings/head at higher precision).
0f746b7
verified

prabodbc commited on

Re-convert with fixed rotary q/k permutation (baseRT-internal#114); NIAH-validated at 2k/8k. Q4 now uses the default-q4 quality profile (embeddings/head at higher precision).
e3fc87e
verified

prabodbc commited on

Add config.json (enables HF download tracking)
614046d
verified

prabodbc commited on

Upload README.md with huggingface_hub
511408f
verified

prabodbc commited on

Upload Llama-3.2-3B-Instruct-Q8.base with huggingface_hub
b11ec4f
verified

prabodbc commited on

Upload Llama-3.2-3B-Instruct-Q4.base with huggingface_hub
b61b5fe
verified

prabodbc commited on

initial commit
b40a4cf
verified

prabodbc commited on