HPC-Coder-v2-1.3b-Q8_0-GGUF

This is the HPC-Coder-v2-6.7b model with 8 bit quantized weights in the GGUF format that can be used with llama.cpp. Refer to the original model card for more details on the model.

Use with llama.cpp

See the llama.cpp repo for installation instructions. You can then use the model as:

llama-cli --hf-repo hpcgroup/hpc-coder-v2-1.3b-Q8_0-GGUF --hf-file hpc-coder-v2-1.3b-q8_0.gguf -r "Below is an instruction that describes a task. Write a response that appropriately completes the request.\n\n### Instruction:" --in-prefix "\n" --in-suffix "\n### Response:\n" -c 8096 -p "your prompt here"

Downloads last month: 5

GGUF

Model size

1B params

Architecture

llama

Hardware compatibility

8-bit

Inference Providers NEW

This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for hpcgroup/hpc-coder-v2-1.3b-Q8_0-GGUF

Base model

hpcgroup/hpc-coder-v2-1.3b

Quantized

(4)

this model

Collection including hpcgroup/hpc-coder-v2-1.3b-Q8_0-GGUF

HPC-Coder-v2 Quantizations

Collection

4 items • Updated Aug 9, 2024