GLM-4-9b Quantized with AutoFP8

使用 m-a-p/COIG-CQIA 的 COIG_pc 集作为校准量化的 glm-4-9b 基座

通常来讲你不会这样用基座模型。

Downloads last month
11
Safetensors
Model size
9.4B params
Tensor type
BF16
·
F8_E4M3
·
Inference Providers NEW
This model is not currently available via any of the supported Inference Providers.
The model cannot be deployed to the HF Inference API: The HF Inference API does not support model that require custom code execution.