Qwen Image 2.1 GGUF

Qwen Image 2.1 is a diffusion transformer for text-to-image generation and image editing. It uses the Qwen3-VL 8B text encoder and supports image references through ComfyUI. It needs to be loaded with the GGUF Loader nodes.

Available Quantizations

Quant File
Q4_0 qwen_image_2.1_Q4.gguf
Q8 qwen_image_2.1_Q8.gguf
Q8_CR qwen_image_2.1_Q8_CR.gguf
Q4_CR Work in progress

The standard Q4_0 and Q8 files use GGML quantization. Q8_CR uses the native INT8 ConvRot path.

Model Input and Output

Inputs

Input Description
Text prompt A text description for image generation or editing instructions.
Reference image Optional image input for image editing and visual conditioning.

Outputs

RGBA image output (Supports transparency)

ComfyUI Setup

Load the model with the GGUF loader node from comfyui-gguf-reboot. Use the Qwen Image 2.1 workflow supplied by the installed ComfyUI version.

Dependencies

Place the matching files from the upstream Qwen Image 2.1 release in the ComfyUI model folders.

Component Folder
Qwen Image 2.1 GGUF models/diffusion_models/ or models/unet/
Qwen3-VL 8B text encoder models/clip/ or text_encoders/
Qwen Image VAE models/vae/
Downloads last month
-
GGUF
Model size
7B params
Architecture
qwen_image21
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for molbal/Qwen-Image-2.1-GGUF

Quantized
(17)
this model