gemma-3-12b-wmt26-vocaball-int4

A compressed, modified derivative of google/gemma-3-12b-it for the WMT26 Model Compression shared task (constrained track). This file has been modified from the original Gemma 3 release: shared Latin+Han+Arabic vocab prune + INT4 W4A16 (GPTQ actorder=static).

  • Base model: google/gemma-3-12b-it (Google)
  • Compression: shared Latin+Han+Arabic vocab prune + INT4 W4A16 (GPTQ actorder=static)
  • Translation directions: ces-deu eng-zho_Hans eng-ara_EG
  • Serving: vLLM, validated stack vllm==0.23.0 / transformers==5.12.1 / compressed-tensors==0.17.0. Inference entrypoint, environment, and run instructions are in the submission package (run.sh / setup.sh / requirements.txt).

Built with Gemma — license & use

Gemma is provided under and subject to the Gemma Terms of Use, found at https://ai.google.dev/gemma/terms . Use is additionally governed by the Gemma Prohibited Use Policy (https://ai.google.dev/gemma/prohibited_use_policy). This is a derivative of google/gemma-3-12b-it; Gemma and its trademarks are the property of Google. By using these weights you agree to the Gemma Terms of Use.

Downloads last month
17
Safetensors
Model size
12B params
Tensor type
I64
·
I32
·
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for soksof/gemma-3-12b-wmt26-vocaball-int4

Quantized
(162)
this model