Qwen3.8-27B MLX MXFP4

Vision-enabled MLX conversion of Qwen/Qwen3.8-27B, pinned to revision 1d4bf0f2ff6012fd82039f2fa52739d0dd7c60c0.

  • Language model: MXFP4, 4-bit, group size 32
  • Vision tower: same-revision BF16 weights
  • Runtime: mlx-vlm 0.6.3 or newer

The Hub's approximately 5.5B safetensors count reflects packed MXFP storage elements; the underlying architecture remains the full 27B model.

Use

python -m mlx_vlm.generate \
  --model Shiftedx/Qwen3.8-27B-MLX-MXFP4 \
  --image image.jpg \
  --prompt "Describe this image."

Text generation and three image-understanding smoke tests passed locally. Quantization can still change behavior, so independently evaluate important use cases. Treat prompts, images, and outputs as untrusted: do not submit secrets, and sandbox tools or generated code with least-privilege access. This conversion adds no telemetry or remote execution.

The upstream Apache-2.0 license and model limitations continue to apply.

Downloads last month
100
Safetensors
Model size
6B params
Tensor type
U8
·
U32
·
BF16
·
MLX
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Shiftedx/Qwen3.8-27B-MLX-MXFP4

Base model

Qwen/Qwen3.8-27B
Quantized
(342)
this model

Collection including Shiftedx/Qwen3.8-27B-MLX-MXFP4