GGUF
feature-extraction

Mellum2-12B-A2.5B-Thinking-Dflash GGUF

Original draft model: https://huggingface.co/RedHatAI/Mellum2-12B-A2.5B-Thinking-Dflash

Original target model: https://huggingface.co/JetBrains/Mellum2-12B-A2.5B-Thinking

Converted using: https://github.com/Anbeeld/beellama.cpp

This repository only contains GGUF conversions.

Downloads last month
261
GGUF
Model size
0.8B params
Architecture
dflash-draft
Hardware compatibility
Log In to add your hardware

4-bit

5-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for rpatel622/Mellum2-12B-A2.5B-Thinking-Dflash-GGUF-FIXED

Quantized
(34)
this model