Qwen2.5-Coder-0.5B-Instruct โ€” Q4_K_M GGUF

A Q4_K_M quantisation of Qwen2.5-Coder-0.5B-Instruct, in GGUF format, for running with llama.cpp.

The base model writes and completes code from a natural-language instruction. Quantised to Q4_K_M so it fits and runs on ordinary consumer hardware, including phones, without a GPU.

Nothing about the model's behaviour has been changed. This is a format and precision conversion only โ€” no fine-tuning, no merging, no adapters.

File

file size
qwen2.5-coder-0.5b-instruct-q4_k_m.gguf 398 MB

Use

llama-cli -m qwen2.5-coder-0.5b-instruct-q4_k_m.gguf -p "write a function that ..."

Any llama.cpp-based runtime will load it.

Licence

Apache 2.0, inherited from the base model. See the base model card for the model's own documentation, capabilities and limitations.

Downloads last month
90
GGUF
Model size
0.5B params
Architecture
qwen2
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Kj0rdan/Qwen2.5-Coder-0.5B-Instruct-Q4_K_M-GGUF

Quantized
(85)
this model