B1-27B GGUF

Q8_0 quant of schneewolflabs/B1-27B plus the vision mmproj (f16, byte-identical to B0-27B's). MTP tensors intact — --spec-type draft-mtp works.

llama-server -m B1-27B-Q8_0.gguf --mmproj B1-27B-mmproj-f16.gguf \
    -ngl 99 -c 16384 --jinja -fa on -np 1 \
    --spec-type draft-mtp --spec-draft-n-max 4
Downloads last month
119
GGUF
Model size
27B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for schneewolflabs/B1-27B-GGUF

Quantized
(3)
this model