This is an importance matrix (imatrix) generated from 1 million tokens for empero-ai/Qwythos-9B-v2.
During testing, the imatrix-quantized model significantly outperformed the standard quantized model.
It was generated using a patched version of llama-imatrix to properly respect chunking boundaries.
Chunking was performed at both 2048 and 16k.

Downloads last month
17
GGUF
Model size
1.28M params
Architecture
Hardware compatibility
Log In to add your hardware

We're not able to determine the quantization variants.

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for aney/Qwythos-9B-v2-imatrx-only

Finetuned
Qwen/Qwen3.5-9B
Quantized
(25)
this model