qwen3.5-9B-8bit-MTPLX

MTPLX-branded multi-token-prediction model for Apple Silicon (MLX). Forged with MTPLX Forge from Qwen/Qwen3.5-9B.

MTPLX Parameters

  • Body bits: 8-bit
  • Group Size: g128
  • MTP Policy: BF16
  • Precision: BF16

Verification

  • Best depth: D3
  • Multiplier vs autoregressive baseline: 2.49脳
  • Verified on: Apple M5
  • Sampler: temperature 0.6 路 top_p 0.95 路 top_k 20

See mtplx_runtime.json for the full verification record.

Usage

# MTPLX picks this model up automatically when downloaded:
mtplx pull FlatFootInternational/Qwen3.5-9B-MTPLX
mtplx start chat

License

See LICENSE.

Downloads last month
52
Safetensors
Model size
9B params
Tensor type
BF16
U32
F32
Inference Providers NEW
This model isn't deployed by any Inference Provider. 馃檵 Ask for provider support

Model tree for FlatFootInternational/Qwen3.5-9B-MTPLX

Quantized
(37)
this model