Quantiazied tinyllama model (stories15m.pt). See original repo: https://github.com/karpathy/llama2.c See how to quantize here: https://github.com/ademyanchuk/llama2.c/blob/master/export.py#L282C14-L282C14 Use version 3 of export on stories15m.pt model.

Use the model in this rust port of llama2.c project: https://github.com/ademyanchuk/llama2-rs

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support