PR Poet Q4_K_M GGUF

A merged and Q4_K_M-quantized version of PR Poet LoRA, based on Qwen/Qwen3-0.6B.

This model powers PR Poet, a GitHub Action that turns a pull request's commit message into a four-line rhyming poem.

The GGUF is intended for CPU inference with llama.cpp. For the original PEFT adapter, use mkly/pr-poet-lora.

Downloads last month
91
GGUF
Model size
0.8B params
Architecture
qwen3
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for mkly/pr-poet-Q4_K_M.gguf

Finetuned
Qwen/Qwen3-0.6B
Quantized
(412)
this model