GGUF
conversational
How to use from
Lemonade
Pull the model
# Download Lemonade from https://lemonade-server.ai/
lemonade pull openresearchtools/Qwen3.5-4B-Instruct-GGUF:
Run and chat with the model
lemonade run user.Qwen3.5-4B-Instruct-GGUF-
List all available models
lemonade list
Quick Links

This GGUF model was converted with llama.cpp from the original checkpoint: Qwen/Qwen3.5-4B. Model template has been changed to permanently turn the reasoning off.

Downloads last month
2,740
GGUF
Model size
4B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

4-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for openresearchtools/Qwen3.5-4B-Instruct-GGUF

Finetuned
Qwen/Qwen3.5-4B
Quantized
(338)
this model