Raw

pipeline_tag: text-generation license: apache-2.0 license_link: https://huggingface.co/Qwen/Qwen3.6-27B/blob/main/LICENSE model_size: 27B quantization: Q4_K architecture: qwen35

Download with hf CLI

Copy download link History Blame Contribute Delete 62.6 kB metadata library_name: transformers license: apache-2.0 license_link: https://huggingface.co/Qwen/Qwen3.6-27B/blob/main/LICENSE pipeline_tag: image-text-to-text Qwen3.6-27B

This is an optimized, hybrid-quantized version of Qwen 3.6 27B, engineered to run smoothly on consumer hardware.

๐Ÿš€ Performance Breakthrough Hardware: Runs directly on CPU with only 16 GB RAM. (Slow but you can)

Minimall Swapping: Minimall lag or heavy disk swapping during inference.

Coding Capable: Tested and proven. Coded a working minigame on the very first try.

๐Ÿ› ๏ธ Quantization Setup To achieve this extreme memory reduction without destroying the model's intelligence, a mixed-precision strategy was used: FP32 Tensors Quantized to Q4_K FP16 Tensors Quantized to Q3_K

๐Ÿ”ฅ BIG THANKS & CREDITS

๐Ÿ’ฅ BIG THX to the Qwen Team! Thank you for giving the open-source community! ๐Ÿ™Œ

๐Ÿ’ฅ HUGE SHOUTOUT to the llama.cpp devs! Without your legendary inference engine.

Downloads last month
361
GGUF
Model size
27B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support