Qwable-27B GGUF (benchmark files)

The two files used for the PXA cross-engine speed comparisons on Tesla P100 and V100 cards, published so every number in the pxq_llama release notes and bench/fair-battle.md can be reproduced from a download. Both are conversions of the same 27B dense hybrid (Qwen3.8-27B lineage, abliterated), so the engines are compared on identical weights.

File Format Size sha256 (first 16) Runs on
Qwable-27B-PXQ4core.gguf PXQ4 (pxq_llama codec) 15.7 GB 6fb4a437684e8bbe pxq_llama only
Qwable-27B-MXFP4-lite.gguf MXFP4 15.0 GB 1d04b75817358cf7 mainline llama.cpp, ik_llama.cpp, pxq_llama

Protocol: bench/fair/protocol.md in the engine repo (temperature 0, median of 7, sha-checked files, unique prompts). Full sha256 sums are in checksums.sha256 here and in bench/fair/weights/MANIFEST.sha256 in the repo.

Community: Discord โ€” PXA Network

Downloads last month
548
GGUF
Model size
27B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support