Raw

pipeline_tag: text-generation license: apache-2.0 license_link: https://huggingface.co/Qwen/Qwen3.6-27B/blob/main/LICENSE model_size: 27B quantization: Q4_K architecture: qwen35

Download with hf CLI

Copy download link History Blame Contribute Delete 62.6 kB metadata library_name: transformers license: apache-2.0 license_link: https://huggingface.co/Qwen/Qwen3.6-27B/blob/main/LICENSE pipeline_tag: image-text-to-text Qwen3.6-27B

This is an optimized, hybrid-quantized version of Qwen 3.6 27B, engineered to run smoothly on consumer hardware.

๐Ÿš€ Performance Breakthrough Hardware: Runs directly on CPU with only 16 GB RAM. (Slow but you can)

Minimall Swapping: Minimall lag or heavy disk swapping during inference.

Coding Capable: Tested and proven. Coded a working minigame on the very first try.

๐Ÿ› ๏ธ Quantization Setup To achieve this extreme memory reduction without destroying the model's intelligence, a mixed-precision strategy was used: FP32 Tensors Quantized to Q4_K FP16 Tensors Quantized to Q3_K

๐Ÿ”ฅ BIG THANKS & CREDITS

๐Ÿ’ฅ BIG THX to the Qwen Team! Thank you for giving the open-source community! ๐Ÿ™Œ

๐Ÿ’ฅ HUGE SHOUTOUT to the llama.cpp devs! Without your legendary inference engine.

Downloads last month
7
GGUF
Model size
27B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support