NOTE!

These are random WIP files for GLM5.3-Flash from a runpod session. They will likely be removed in the comming days once I get quants figured out. In the meantime, the BF16's are both functional along with the LOGIT dump. I would recommend AGAINST using the imatrix files however as they seem to be broken in my testing.

EDIT:

Okay yeah, the imatrix files in here are 100% broken, DO NOT USE THEM! Please use the ones from AesSedai's repo if you are doing your own quant mixes!

https://huggingface.co/AesSedai/GLM-5.3-Flash-GGUF

Downloads last month
1,189
GGUF
Model size
321B params
Architecture
glm5-next
Hardware compatibility
Log In to add your hardware

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Sciguy429/GLM-5.3-Flash-BF16

Quantized
(83)
this model