Built to run with antirez / ds4 inference engine. Not tested or built for llama cpp. Q4K available with dspark drafter.

Downloads last month
-
GGUF
Model size
20B params
Architecture
deepseek4-dspark
Hardware compatibility
Log In to add your hardware

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for sm54/deepseek-v4-flash-0731-gguf

Quantized
(42)
this model