B0-27B GGUF

Q8_0 quant of schneewolflabs/B0-27B plus the vision mmproj (f16). MTP tensors intact — --spec-type draft-mtp works.

llama-server -m B0-27B-Q8_0.gguf --mmproj B0-27B-mmproj-f16.gguf \
    -ngl 99 -c 16384 --jinja -fa on -np 1 \
    --spec-type draft-mtp --spec-draft-n-max 4
Downloads last month
118
GGUF
Model size
27B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for schneewolflabs/B0-27B-GGUF

Quantized
(3)
this model