GGUF
feature-extraction
How to use from
Lemonade
Pull the model
# Download Lemonade from https://lemonade-server.ai/
lemonade pull rpatel622/Mellum2-12B-A2.5B-Thinking-Dflash-GGUF-FIXED:
Run and chat with the model
lemonade run user.Mellum2-12B-A2.5B-Thinking-Dflash-GGUF-FIXED-
List all available models
lemonade list
Quick Links

Mellum2-12B-A2.5B-Thinking-Dflash GGUF

Original draft model: https://huggingface.co/RedHatAI/Mellum2-12B-A2.5B-Thinking-Dflash

Original target model: https://huggingface.co/JetBrains/Mellum2-12B-A2.5B-Thinking

Converted using: https://github.com/Anbeeld/beellama.cpp

This repository only contains GGUF conversions.

Downloads last month
261
GGUF
Model size
0.8B params
Architecture
dflash-draft
Hardware compatibility
Log In to add your hardware

4-bit

5-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for rpatel622/Mellum2-12B-A2.5B-Thinking-Dflash-GGUF-FIXED

Quantized
(34)
this model