Instructions to use antareslabs/hunch-0.6b-preview-MLX with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- MLX
How to use antareslabs/hunch-0.6b-preview-MLX with MLX:
# Download the model from the Hub pip install huggingface_hub[hf_xet] huggingface-cli download --local-dir hunch-0.6b-preview-MLX antareslabs/hunch-0.6b-preview-MLX
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- LM Studio
- Atomic Chat
hunch-0.6b-preview-MLX
The MLX conversion of antareslabs/hunch-0.6b-preview at the precisions that passed the
equivalence gate: every file was scored on all 6,000 held-out questions against the fp32 PyTorch run of the same
checkpoint.
hunch-0.6b-preview-mlx-f16: changes 0 of 6,000 answers against fp32 (bf16 alone changes 28), max TV 5.69e-03 against a floor of 7.68e-02; read on Apple silicon (Metal); gate: pass.
The pass rule and the builds that failed it are in FORMATS. Load it with
hunch.formats.hunch_mlx.load(path) from the Hunch repository; the release temperature
is embedded in the file and applied by default.
Files: hunch-0.6b-preview-mlx-f16.
License: Apache-2.0, as for the model it converts (model card).
Hardware compatibility
Log In to add your hardware
Quantized
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐ Ask for provider support