Qwen3Guard-Gen-0.6B (GGUF, Q8_0)

A GGUF (llama.cpp) build of Qwen/Qwen3Guard-Gen-0.6B, quantized to Q8_0, used as the content-safety guard for Beyond Logic Labs' game Netshell. It runs in-process as an engine-managed llama-server child and classifies free text into the model's native safety categories.

Files

  • Qwen3Guard-Gen-0.6B.Q8_0.gguf

Usage

llama-server -m Qwen3Guard-Gen-0.6B.Q8_0.gguf --port 8091 -c 32768 --no-context-shift

Classification runs in response mode (a user turn plus the assistant turn under judgement); the model emits a short Safety: <verdict>\nCategories: ... block.

License and attribution

Apache-2.0, inherited from the base model. All credit for the model itself goes to the Qwen team; this repository only redistributes a quantized build. See the base model card for details.

Downloads last month
17
GGUF
Model size
0.8B params
Architecture
qwen3
Hardware compatibility
Log In to add your hardware

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for beyond-logic-labs/Qwen3Guard-Gen-0.6B-gguf

Finetuned
Qwen/Qwen3-0.6B
Quantized
(11)
this model