Gemma 4 26B A4B LINE LoRA — FP8_DYNAMIC

This checkpoint merges the LINE product-relevance LoRA into its exact BF16 training base and exports it as FP8_DYNAMIC.

  • LoRA SHA256: ebfe712be3a94a3c949f62c8d383c05d5b8fd88377a5ee1737e4de437c830c94
  • LoRA targets: 205
  • LoRA rank / alpha: 32 / 64
  • Referenced weight files: 3
  • Referenced weight bytes: 27165275244
  • Quantization method: compressed-tensors

The quantized Gemma 4 MoE experts use the linearized compressed-tensors layout intended for vLLM/SGLang. Validate quality through the serving engine; a plain Transformers load may not reconstruct this serving layout.

See export_manifest.json for exact provenance and package versions.

Downloads last month
115
Safetensors
Model size
26B params
Tensor type
BF16
·
F8_E4M3
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for MariusAlonso/gemma-4-26B-A4B-it-FP8-Dynamic-lora-v2

Adapter
(10)
this model