This repository stores Z Lab's Qwen3.6 27B DFlash drafter quantized to Q4_K_M and Q8_0 using upstream llama.cpp and following spiritbuun's guide.

Downloads last month
-
GGUF
Model size
2B params
Architecture
dflash
Hardware compatibility
Log In to add your hardware

4-bit

8-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for canhdu/Qwen3.6-27B-DFlash-GGUF

Quantized
(14)
this model