This is an attempted conversion of https://huggingface.co/UCloud-org/GLM-5.2-FP8-DFlash to gguf format. Both models function in ik_llama, but I cannot get either to get higher than 5% acceptance. I'm not sure which tokenizer to use, I think is the problem.

Downloads last month
473
GGUF
Model size
4B params
Architecture
dflash-draft
Hardware compatibility
Log In to add your hardware

We're not able to determine the quantization variants.

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for muzzy/GLM-5.2-FP8-DFlash-GGUF

Quantized
(1)
this model