LFM2.5-8B-A1B-MLX-8bit

MLX 8bit (affine, group size 64) quantized variant of LiquidAI/LFM2.5-8B-A1B (8B/1B-active mixture-of-experts) for Apple silicon via mlx-lm.

Provenance

  • Source: LiquidAI/LFM2.5-8B-A1B @ revision b9aebfcbe28b6cb374042f495d733037550ab146 (LFM Open License v1.0 — see LICENSE in this repo).
  • Quantized with mlx_lm.convert (mlx-lm 0.31.3): affine, 8-bit, group size 64.

Smoke gate

Before upload this pack passed a deterministic coherence gate: greedy 32-token chat generation loaded through mlx_lm.load, judged for emptiness, repetition loops, multi-script gibberish, and special-token debris. Verdict: ok.

Usage

pip install mlx-lm
mlx_lm.generate --model majentik/LFM2.5-8B-A1B-MLX-8bit --prompt "Hello"

Evaluation

Benchmark Score
arc_easy_acc 0.4750
hellaswag_acc 0.4450

License

LFM Open License v1.0 (lfm1.0): commercial use permitted with attribution. The full text ships as LICENSE in this repository, copied verbatim from the upstream model.

Available tiers

Downloads last month
46
Safetensors
Model size
2B params
Tensor type
BF16
·
U32
·
F32
·
MLX
Hardware compatibility
Log In to add your hardware

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for majentik/LFM2.5-8B-A1B-MLX-8bit

Quantized
(86)
this model