Kuza Gemma 4 E2B

Full training-run archive for Kuza (East Africa agricultural assistant), fine-tuned from unsloth/gemma-4-E2B-it-qat-q4_0-unquantized. Weights, logs, checkpoints, GGUFs, and provenance are stored with the same layout as $KUZA_WORK_DIR/kuza-gemma-4-e2b/.

This derivative is subject to the Gemma license.

Training mix

  • 100% English train from kuzaai/kuza_sft_english
  • 35% Swahili from kuzaai/kuza_sft_swahili
  • 8% HuggingFaceH4/no_robots
  • 5% adversarial from kuzaai/kuza_sft_adversarial
  • all multiturn from kuzaai/kuza_sft_multiturn

LoRA: RsLoRA r=32, alpha=64, QAT int4. Sequence length 1024, 2 epochs, LR 2e-05. Thinking is off. GGUFs are text-only (PLE kept; vision/audio dropped).

Files

  • adapter/ โ€” PEFT adapter, tokenizer, SFT metrics and manifests
  • training/ โ€” Trainer checkpoints including checkpoint-best
  • merged_bf16/ โ€” text-only merged Hugging Face BF16 weights
  • reference/ โ€” text-only BF16 GGUF (kuza-bf16.gguf) and smoke log
  • imatrix/ โ€” calibration corpus, eval corpus, imatrix, logs
  • quants/ โ€” quantized GGUF candidates:
  • q4_k_m_ud_style/kuza-q4_k_m-ud-style.gguf (q4_k_m)
  • ud_q4_k_xl/kuza-ud-q4_k_xl.gguf (q4_k_m)
  • q4_0_qat_aligned/kuza-q4_0-qat-aligned.gguf (q4_0)
  • screen/ โ€” GPU KLD diagnostic, hidden-set scores, and results.json winner
  • provenance/ โ€” copied adapter metrics, recipes, screen JSON
  • upload_manifest.json โ€” path, size, and sha256 for every uploaded file

Download

huggingface-cli download kuzaai/kuza-gemma-4-e2b --local-dir ./kuza-gemma-4-e2b

screen/results.json ranks by hidden-set accuracy, then GGUF size, then GPU TPS. Do not treat GPU TPS as an ADTC laptop measurement.

Downloads last month
24
GGUF
Model size
5B params
Architecture
gemma4
Hardware compatibility
Log In to add your hardware

4-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for kuzaai/kuza-gemma-4-e2b-old

Adapter
(8)
this model