Built to run with antirez / ds4 inference engine. Not tested or built for llama cpp. Q4K available with dspark drafter.
- Downloads last month
- -
Hardware compatibility
Log In to add your hardware
16-bit
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐ Ask for provider support
Model tree for sm54/deepseek-v4-flash-0731-gguf
Base model
deepseek-ai/DeepSeek-V4-Flash-0731