Sharing W4A16 model for 3090 users :)

#2
by JC1DA - opened

This model was quantized using AutoRound.
More benchmarks are coming soon

Model Accuracy Acc Norm
Qwopus3.8-27B-Flash (FP16) 83.71% 77.06%
Qwopus3.8-27B-Flash (W4A16) 84.13% 78.54%

Sign up or log in to comment