AX-Qwen3.8-Flash-Next-MLX-AXQ-4bit-MTP

An AXQuant (AXQ) mixed-precision MLX checkpoint for Apple Silicon, converted from Qwen/Qwen3.8-Flash-Next (qwen4_exp).

Development evidence — not a certified AXQuant release. Conversion and artifact-integrity records only. No quality, long-context, or MTP-speed claim.

Property Value
Base model Qwen/Qwen3.8-Flash-Next
Source revision de4b8e4d43b917e7706784d8bb445c9af86a3540
Product family qwen4-exp
Adapter qwen4-exp-v1
Recipe qwen38-flash-next-axq4-v0.1.yaml
Quant lane 4-bit affine
Measured BPW 6.04854356743359
Vision BF16-protected
PLE / n-gram embeddings embedding-floor (8-bit affine)
MTP packaged mtp.safetensors (BF16, not a speed claim)
Runtime MLX-VLM (qwen4_exp); AX Engine support is not claimed

Load with a mlx-vlm build that includes models.qwen4_exp.

Downloads last month
-
Safetensors
Model size
37B params
Tensor type
BF16
·
U32
·
I64
·
MLX
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for AutomatosX/AX-Qwen3.8-Flash-Next-MLX-AXQ-4bit-MTP

Quantized
(123)
this model