GGUF
conversational
How to use from
Docker Model Runner
docker model run hf.co/openresearchtools/Qwen3.5-4B-Instruct-GGUF:
Quick Links

This GGUF model was converted with llama.cpp from the original checkpoint: Qwen/Qwen3.5-4B. Model template has been changed to permanently turn the reasoning off.

Downloads last month
3,431
GGUF
Model size
4B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

4-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for openresearchtools/Qwen3.5-4B-Instruct-GGUF

Finetuned
Qwen/Qwen3.5-4B
Quantized
(368)
this model