I made a couple more models just for fun. I can’t vouch for the quality, since I didn’t do any tests. Link to the original model and the script with which I created these models:

Downloads last month
50
GGUF
Model size
10.7B params
Architecture
llama
Hardware compatibility
Log In to view the estimation

4-bit

5-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support