TaskFlow AI โ€” Qwen3 4B Instruct (Q4_K_M)

This repository contains the Q4_K_M GGUF quantization of Qwen3-4B-Instruct-2507, packaged by the official Ollama distribution as qwen3:4b-instruct-2507-q4_K_M.

This is an unmodified quantized base model. It has not been fine-tuned by TaskFlow AI. The upstream model is published under the Apache-2.0 license; see LICENSE and the upstream model card for details and attribution.

Files

  • Qwen3-4B-Instruct-2507-Q4_K_M.gguf โ€” quantized model weights (about 2.5 GB).
  • Modelfile โ€” settings to import the GGUF into Ollama.
  • LICENSE โ€” license copied from the upstream model distribution.

Run with Ollama

Download the GGUF file and Modelfile into the same directory, then run:

ollama create taskflow-qwen3 -f Modelfile
ollama run taskflow-qwen3

Quantization

  • Base model: Qwen/Qwen3-4B-Instruct-2507
  • Format: GGUF
  • Quantization: Q4_K_M
  • Source package: qwen3:4b-instruct-2507-q4_K_M from the Ollama model library
Downloads last month
7
GGUF
Model size
4B params
Architecture
qwen3
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for Lord230/TaskFlowAi

Quantized
(315)
this model