Download quantization_config.json from LeaderboardModel1/experiment024b-AutoRound-W4A16-RTN: direct link, hf CLI and curl.
- Browser
- Download file 302 Bytes
-
https://huggingface.co/LeaderboardModel1/experiment024b-AutoRound-W4A16-RTN/resolve/main/quantization_config.json
- Command line
-
hf download hf://LeaderboardModel1/experiment024b-AutoRound-W4A16-RTN/quantization_config.json
-
curl -L -o quantization_config.json https://huggingface.co/LeaderboardModel1/experiment024b-AutoRound-W4A16-RTN/resolve/main/quantization_config.json
302 Bytes
| { | |
| "bits": 4, | |
| "data_type": "int", | |
| "group_size": 128, | |
| "sym": true, | |
| "enable_quanted_input": false, | |
| "iters": 0, | |
| "low_gpu_mem_usage": true, | |
| "autoround_version": "0.13.1", | |
| "block_name_to_quantize": "model.layers", | |
| "quant_method": "auto-round", | |
| "packing_format": "auto_round:auto_gptq" | |
| } |