NB-Llama-3.1-8B-Instruct โ EXL2
ExLlamaV2 / EXL2 quants of NbAiLab/nb-llama-3.1-8B-Instruct.
Official Hub files are BF16 and GGUF. There was no EXL2 pack.
Converted with ExLlamaV2 0.3.2, lm_head at 6-bit, built-in default calibration. One measurement pass, then each bitrate from measurement.json.
Branches
hf download oxfrug/nb-llama-3.1-8B-Instruct-exl2 --revision 5.0bpw --local-dir ./nb-llama-3.1-8B-Instruct-exl2-5.0bpw
Notes
- License: Meta Llama 3.1 Community License. Keep
NOTICE. - Loader: ExLlamaV2. Not GGUF.
- On PyTorch 2.13 without Flash Attention 2.5.7+, set
config.no_sdpa = Truebefore load.
Source
NbAiLab/nb-llama-3.1-8B-Instruct
โ meta-llama/Llama-3.1-8B-Instruct
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐ Ask for provider support
Model tree for oxfrug/nb-llama-3.1-8B-Instruct-exl2
Base model
meta-llama/Llama-3.1-8B Finetuned
meta-llama/Llama-3.1-8B-Instruct Finetuned
NbAiLab/nb-notram-llama-3.1-8b-instruct