Instructions to use carloslfu/Qwen3.8-Flash-Next-MLX-4bit with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use carloslfu/Qwen3.8-Flash-Next-MLX-4bit with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir Qwen3.8-Flash-Next-MLX-4bit carloslfu/Qwen3.8-Flash-Next-MLX-4bit
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
Qwen3.8-Flash-Next MLX 4-bit (mirror)
A byte-identical mirror of
pipenetwork/Qwen3.8-Flash-Next-MLX-4bit
at revision aa7c790e804bbf9d491ddb109c3d61bc4a555f7c. Every file has the same
sha256 as that revision. No requantization, no repacking, no changes.
It exists so that slotstream — which runs this model on Apple Silicon by streaming experts from SSD — has a download source under its own maintainer's control. slotstream tries this mirror first and the original second, and verifies every file against hashes compiled into the binary, so integrity never depends on which source served the bytes.
Credit for the conversion belongs to pipenetwork; the model is Qwen/Qwen3.8-Flash-Next and remains under the Qwen community license included here.
- Downloads last month
- 103
4-bit
Model tree for carloslfu/Qwen3.8-Flash-Next-MLX-4bit
Base model
Qwen/Qwen3.8-Flash-Next