gemma-3-12b-wmt26-int4

A compressed, modified derivative of google/gemma-3-12b-it for the WMT26 Model Compression shared task (constrained track). This file has been modified from the original Gemma 3 release: INT4 W4A16 (GPTQ โ†’ compressed-tensors), near-lossless.

  • Base model: google/gemma-3-12b-it (Google)
  • Compression: INT4 W4A16 (GPTQ โ†’ compressed-tensors), near-lossless
  • Translation directions: ces-deu eng-zho_Hans eng-ara_EG
  • Serving: vLLM, validated stack vllm==0.23.0 / transformers==5.12.1 / compressed-tensors==0.17.0. Inference entrypoint, environment, and run instructions are in the submission package (run.sh / setup.sh / requirements.txt).

Built with Gemma โ€” license & use

Gemma is provided under and subject to the Gemma Terms of Use, found at https://ai.google.dev/gemma/terms . Use is additionally governed by the Gemma Prohibited Use Policy (https://ai.google.dev/gemma/prohibited_use_policy). This is a derivative of google/gemma-3-12b-it; Gemma and its trademarks are the property of Google. By using these weights you agree to the Gemma Terms of Use.

Downloads last month
37
Safetensors
Model size
2B params
Tensor type
I64
ยท
I32
ยท
BF16
ยท
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for soksof/gemma-3-12b-wmt26-int4

Quantized
(162)
this model