Mirror of meta-llama/Llama-3.2-3B-Instruct Q4_0 GGUF by pitcany. Mirrored 2026-09-15 as takedown protection.

Single-file Q4_0 GGUF (1.9 GB), built by Ollama from the official Meta checkpoint. Use with llama.cpp, Ollama, or any GGUF-compatible runtime.

Usage

# Ollama
ollama run llama3.2-3b

# llama.cpp
llama-cli -m llama3.2-3b-instruct-q4_0.gguf
Downloads last month
17
GGUF
Model size
3B params
Architecture
llama
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for pitcany/llama3.2-3b-instruct-q4_0

Quantized
(534)
this model