Measured peak <192 GB; estimated 256GB fit.
AI & ML interests
Exploring local AI and sharing MLX work with the community.
Recent Activity
View all activity
Organization Card
PERSONAL LOCAL-AI PROJECT
Vontra
Exploring local AI and sharing useful MLX work with the community.
I’m one person interested in what open models can do on hardware you can actually own. I work through the difficult parts of running them locally and share anything that might save somebody else time.
What I work on
- MLX and Apple silicon: conversions, quantized checkpoints and architecture compatibility.
- Local inference: practical serving notes, memory observations and fixes tested on real hardware.
- Open sharing: useful artifacts, honest caveats and clear credit to the original model creators.
The local lab
Most of the MLX work happens on a Mac Studio and MacBook Pro. I also use a ThinkStation PGX (DGX Spark) for local inference experiments outside the Apple ecosystem.
If something here saves you a conversion, a compatibility fix or a few hours of testing, then sharing it was worthwhile.
Personal project · Best effort · Upstream licences and attribution always apply
models 62
Vontra/DeepSeek-V4.1-Flash-MLX-2bit-MTP
Text Generation • 763B • Updated
Vontra/Nex-N2.5-mini-MLX-8bit
Image-Text-to-Text • 35B • Updated • 78 • 1
Vontra/Nex-N2.5-mini-MLX-6bit
Image-Text-to-Text • 35B • Updated • 28 • 1
Vontra/Nex-N2.5-mini-MLX-4bit
Image-Text-to-Text • 35B • Updated • 49
Vontra/Nex-N2.5-mini-MLX-oQ8
Image-Text-to-Text • 35B • Updated • 29
Vontra/Nex-N2.5-mini-MLX-oQ6
Image-Text-to-Text • 35B • Updated • 37 • 1
Vontra/Nex-N2.5-mini-MLX-oQ4
Image-Text-to-Text • 35B • Updated • 82
Vontra/Nex-N2.5-mini-MLX-oQ3
Image-Text-to-Text • 35B • Updated • 20
Vontra/Nex-N2.5-mini-MLX-oQ2
Image-Text-to-Text • 35B • Updated • 31 • 1
Vontra/GLM-5.3-Flash-MLX-2bit-MTP
Image-Text-to-Text • 352B • Updated • 1.9k • 1
datasets 0
None public yet