Text Generation
Transformers
GGUF
English
llama

This repo includes .gguf built for HuggingFace/Candle. They will not work with llama.cpp.

Refer to the original repo for more details.

Downloads last month
247
GGUF
Model size
1B params
Architecture
Hardware compatibility
Log In to add your hardware

4-bit

5-bit

8-bit

16-bit

Inference Providers NEW