Views
No views yet
| Branch | Bits | lm_head | Dataset | Size | Description |
|---|---|---|---|---|---|
| 8_0 | 8.0 | 8.0 | Default | 9.8 GB | Maximum quality that ExLlamaV2 can produce, near unquantized performance. |
| 6_5 | 6.5 | 8.0 | Default | 8.6 GB | Very similar to 8.0, good tradeoff of size vs performance, recommended. |
| 5_0 | 5.0 | 6.0 | Default | 7.4 GB | Slightly lower perplexity vs 6.5. |
| 4_25 | 4.25 | 6.0 | Default | 6.7 GB | GPTQ equivalent bits per weight. |
| 3_5 | 3.5 | 6.0 | Default | 6.1 GB | Lower quality, only use if you have to. |
git clone --single-branch --branch 6_5 https://huggingface.co/bartowski/internlm2-7b-llama-exl2 internlm2-7b-llama-exl2-6_5pip3 install huggingface-hubmain (only useful if you only care about measurement.json) branch to a folder called internlm2-7b-llama-exl2:1mkdir internlm2-7b-llama-exl2
2huggingface-cli download bartowski/internlm2-7b-llama-exl2 --local-dir internlm2-7b-llama-exl2 --local-dir-use-symlinks False--revision parameter:1mkdir internlm2-7b-llama-exl2-6_5
2huggingface-cli download bartowski/internlm2-7b-llama-exl2 --revision 6_5 --local-dir internlm2-7b-llama-exl2-6_5 --local-dir-use-symlinks False