Incorrecter GGUF (Q8_0)

Quantized build of Avicennasis/incorrecter for ollama. This is the quant that passed the evaluation gate; Q4_K_M was rejected (1–3-word edits 0.60 < 0.70 on the 40-text check).

ollama run hf.co/Avicennasis/incorrecter-GGUF:Q8_0

The repo's params file sets the recommended sampling, so the command above needs no extra flags. To build it yourself with a Modelfile:

FROM ./incorrecter-Q8_0.gguf
PARAMETER temperature 1.0
PARAMETER top_p 0.9
PARAMETER top_k 40
PARAMETER repeat_penalty 1.0

Keep repeat_penalty at 1.0. The model's job is to copy your text with a few mistakes, and a repeat penalty punishes the copying. With ollama's default (1.1) at temperature 0.9, only 0.62 of changed texts stayed within 1–3 word edits. The top_p cap keeps the copy exact while leaving room for the intended typos.

Measured through ollama (40 held-out texts, 3 draws, settings above): 0.93 of texts changed, 0.85 of those with 1–3 word edits, 0.98 line count kept, 0.97 sign-off kept, and 0.94 judged to keep their meaning (Llama-3.3-70B judge). Feed it clean text as the user message. See the main model card for training data, evaluation and limitations.

Downloads last month
92
GGUF
Model size
0.6B params
Architecture
qwen2
Hardware compatibility
Log In to add your hardware

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Avicennasis/incorrecter-GGUF

Quantized
(1)
this model