Llama-3.2-1B-Instruct

BaseRT .base builds of meta-llama/Llama-3.2-1B-Instruct for fast local inference on Apple Silicon (Metal).

Files

File Precision Size
Llama-3.2-1B-Instruct-Q4.base 4-bit 702 MB
Llama-3.2-1B-Instruct-Q8.base 8-bit 1.3 GB

Usage

curl -LsSf https://basecompute.co/install.sh | sh
basert pull basecompute/Llama-3.2-1B-Instruct
basert chat basecompute/Llama-3.2-1B-Instruct

Released under the llama3.2 license, inherited from the base model.

Downloads last month
30
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for basecompute/Llama-3.2-1B-Instruct

Finetuned
(1759)
this model