Muse-Glimmer-30B-MLX-3bit

MLX 3bit (affine, group size 32) quantized variant of meta-models/Muse-Glimmer-30B (text tower quantized; vision tower and projector retained in BF16) for Apple silicon via mlx-lm.

Provenance

  • Source: meta-models/Muse-Glimmer-30B @ revision a4e59da52a7bc87ae7251dd5545c0dd437c44b68 (Apache-2.0 (upstream LICENSE)).
  • Quantized with mlx_lm.convert (mlx-lm 0.31.3): affine, 3-bit, group size 32.

Smoke gate

Before upload this pack passed a deterministic coherence gate: greedy 48-token chat generation loaded through mlx_lm.load, judged for emptiness, repetition loops, multi-script gibberish, and special-token debris. Verdict: ok.

Usage

pip install mlx-lm
mlx_lm.generate --model majentik/Muse-Glimmer-30B-MLX-3bit --prompt "Hello"

Evaluation

Text-benchmark scores are not published for this pack: the muse_glimmer architecture is not loadable by stock mlx_lm, which the eval harness uses. Quality gating is the deterministic coherence smoke above (greedy 48-token generation via mlx_vlm), which this tier passed before upload.

License

Apache-2.0 — see the upstream LICENSE file in meta-models/Muse-Glimmer-30B.

Available tiers

Downloads last month
108
Safetensors
Model size
6B params
Tensor type
BF16
·
U32
·
MLX
Hardware compatibility
Log In to add your hardware

3-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for majentik/Muse-Glimmer-30B-MLX-3bit

Quantized
(136)
this model