GPT4All-J-MXFP4-mlx

Introduce

Quant by mlx-lm 0.31.3 with new pipeline for GPT-J model_type

ref

Usage

mlx_lm.generate mlx_lm.generate --model /path/to/model --prompt "What is different between CPU and GPU?" --max-tokens 2048
Downloads last month
49
Safetensors
Model size
1B params
Tensor type
F32
U8
U32
MLX
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 馃檵 Ask for provider support

Model tree for AlphaOxO/GPT4All-J-MXFP4-mlx

Quantized
(1)
this model