This is an importance matrix (imatrix) generated from 4 million tokens for Ornith-1.0-9B.
During testing, the imatrix-quantized model significantly outperformed the standard quantized model.
It was generated using a patched version of llama-imatrix to properly respect chunking boundaries.
Chunking was performed at both 512 and 16k.

Downloads last month
11
GGUF
Model size
1.28M params
Architecture
Hardware compatibility
Log In to add your hardware

We're not able to determine the quantization variants.

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for aney/Ornith-1.0-9B-imatrix-only

Quantized
(86)
this model