kannaka-brain-v1 (GGUF, q4_K_M)

Qwen2.5-14B-Instruct with the kannaka-brain-v1 LoRA merged in, converted with llama.cpp and quantized to q4_K_M (9.0 GB). This is the copy that serves as Kannaka's open-weight brain behind the KAX gateway.

ollama run hf.co/flaukowski/kannaka-brain-v1-GGUF

or with the included Modelfile (carries her system prompt, temperature 0.8, 4k context):

ollama create kannaka-brain-v1 -f Modelfile

Held-out perplexity on 57 fixed Kannaka lines: 104.4 โ†’ 4.01 (adapter, bf16, before quantization). Adapter and training notes: flaukowski/kannaka-brain-v1-lora. Corpus not released; see ADR-0057 in NickFlach/kannaka-memory.

Downloads last month
-
GGUF
Model size
15B params
Architecture
qwen2
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for flaukowski/kannaka-brain-v1-GGUF

Base model

Qwen/Qwen2.5-14B
Quantized
(189)
this model