Iki cleanup model — Gemma 4 E2B "iki6" (Q4_0)

Fine-tuned cleanup model for Iki, the on-device macOS dictation app. QLoRA r32/α32, 3 epochs, merged into google/gemma-4-E2B-it and quantized to Q4_0 with llama.cpp.

Training data is synthetic/public only — no user dictation, no PHI: in-repo synthetic clinical/social-work lines, templated numeral-convention lines, the Iki benchmark references, and a wikitext-103 sample (CC BY-SA), corrupted with Iki's generator (iki-bench generate-cleanup-dataset, dataset v6, 9,151 rows). See docs/MODEL_PROVENANCE.md in the Iki repo.

Gates vs stock E2B Q4_0 (Iki's harness): cleanup-corpus exact match 54.7% → 73.6%, numeral golden 18/31 → 28/31, profanity preservation 20/20, hand check 50/50 with zero invented content, WER within noise, p95 cleanup latency 1531 → 509 ms.

The Iki app downloads gemma-4-E2B-it-iki6-Q4_0.gguf from this repo on first launch.

Downloads last month
27
GGUF
Model size
5B params
Architecture
gemma4
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for the-acai/iki-gemma-cleanup

Quantized
(330)
this model