Mistral-Medium-3.5-128B-MLX-5bit

MLX 5bit (affine, group size 64) quantized variant of mistralai/Mistral-Medium-3.5-128B (text tower quantized; vision tower and projector retained in BF16; source checkpoint is FP8) for Apple silicon via mlx-lm.

Provenance

  • Source: mistralai/Mistral-Medium-3.5-128B @ revision 22b2b868a15677cfa6061277ed2f653d1349a9ab (Modified MIT — revenue carve-out below $20M/month).
  • Quantized with mlx_lm.convert (mlx-lm 0.31.3): affine, 5-bit, group size 64.

Smoke gate

Before upload this pack passed a deterministic coherence gate: greedy 48-token chat generation loaded through mlx_lm.load, judged for emptiness, repetition loops, multi-script gibberish, and special-token debris. Verdict: ok.

Usage

pip install mlx-lm
mlx_lm.generate --model majentik/Mistral-Medium-3.5-128B-MLX-5bit --prompt "Hello"

License

Modified MIT (upstream LICENSE): free use with a revenue carve-out — organisations above $20M/month revenue need a commercial agreement with Mistral AI. This is not a plain MIT license; read the upstream terms before commercial deployment.

Available tiers

Downloads last month
26
Safetensors
Model size
27B params
Tensor type
BF16
·
U32
·
MLX
Hardware compatibility
Log In to add your hardware

5-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for majentik/Mistral-Medium-3.5-128B-MLX-5bit

Quantized
(31)
this model