Qwen3-Coder 30B-A3B โ€” MER Q4_0 Qualification Artifact

Private engineering artifact containing the exact Q4_0 files used to qualify Micro-Expert-Router-SSD-Streamed-MoE (MER).

This is not yet a release-grade Amalgafy quantization or model-quality benchmark. The GGUF was requantized from an existing quantized GGUF with llama.cpp using --allow-requantize and --pure. Requantization may compound quantization error.

Artifacts

  • artifacts/gguf/Qwen3-Coder-30B-A3B-Instruct-pure-Q4_0.gguf
  • artifacts/mer/qwen3-coder-30b-a3b-mer-q4_0-v1.tar.zst
  • evidence/pr6-q4-parity.json

The MER archive contains 6,144 routed experts, 435 dense tensors, tokenizer, configuration, metadata, and the canonical ggml-standard-v1 Q4_0 layout.

Qualification

Qualified on an NVIDIA L4 through WGPU/Vulkan using MER commit:

dac1d213cf641ba79a48e74c24f80bc2eca66548

Results:

  • Seven raw WGSL Q4_0 cases passed
  • Three complete checkpoint-expert vectors passed
  • Initial expert installation occurred exactly once
  • Subsequent vectors uploaded zero expert-weight bytes
  • Zero CPU fallback or degraded expert execution
  • Worst complete-expert absolute error: 7.6293945e-06

Checksums

  • Pure Q4_0 GGUF: 8ddf61cadd354a5095905cc5ce535c44b777d0313ac241abcd2ceafa3362551b
  • MER archive: 659b8d31d0a83292c632aa109c8edb5301f4041b1a60ef43c6f23ec0404061fe
  • Parity report: 1d579a9e7ebc93191544ff162027e840dfbbd55ae7cc85e81021bb6e85784c60

Provenance

  • Upstream: Qwen/Qwen3-Coder-30B-A3B-Instruct
  • License: Apache-2.0
  • llama.cpp: 030ebb558a5820b444a8f836ed5cdd46c9b4bd7a
  • MER: dac1d213cf641ba79a48e74c24f80bc2eca66548

Qwen3-Coder is provided by the Qwen team. This repository preserves the upstream license and identifies the conversion and requantization changes.

Downloads last month
-
GGUF
Model size
31B params
Architecture
qwen3moe
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Amalgafy/Qwen3-Coder-30B-A3B-Instruct-MER-Q4-0

Quantized
(161)
this model