CareDraft E4B (fine-tuned)

A LoRA fine-tune of Gemma 4 E4B QAT (google/gemma-4-E4B-it), trained by the CareDraft team to turn rough, speech-to-text home-health dictation into polished, Medicare-compliant clinical documentation across physical therapy, occupational therapy, speech-language pathology, and skilled nursing.

Training data is 100% synthetic — no real patient information, transcripts, or notes were used or are reproducible from this model. All training pairs were generated and mechanically validated (numeric-fact preservation, denial/ negation integrity, no unsupported claims) before training.

  • Format: GGUF, Q4_K_M quantization
  • Size: ~5.30 GB
  • SHA-256: 9e03f6af0b55fcb88358624055fa622b64a46b92803599b37e1016aa5eb011ca
  • Base: Gemma 4 E4B QAT (4-bit)
  • Intended use: On-device note generation inside the CareDraft app. Not intended for standalone medical use — all output requires clinician review before it becomes part of a medical record.

License

This model is a derivative of Google's Gemma models and is distributed under the Gemma Terms of Use. By downloading or using this model you agree to those terms, including the Gemma Prohibited Use Policy.

Usage

This model is built for CareDraft's local llama.cpp-based inference pipeline (Gemma chat template, <end_of_turn>/<eos> stop tokens) and is not packaged as a general-purpose chat assistant.

Downloads last month
61
GGUF
Model size
7B params
Architecture
gemma4
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for clesterdpt/caredraft-e4b

Quantized
(359)
this model