chinse INT8 Smart V2

Source: jasort/chinse

Quantization: bitsandbytes LLM.int8()

Baseline NLL: 4.057724446058273

INT8 NLL: 4.010224133729935

Relative NLL increase: -0.011706145392519457

Downloads last month
-
Safetensors
Model size
5B params
Tensor type
F32
F16
I8
Inference Providers NEW
This model isn't deployed by any Inference Provider. 馃檵 Ask for provider support

Model tree for jasort/chinse-int8-smart-v2

Finetuned
jasort/chinse
Quantized
(1)
this model