GGUF
How to use from
Lemonade
Pull the model
# Download Lemonade from https://lemonade-server.ai/
lemonade pull mzayed/gemma-7b-it-q4_k_m:Q4_K_M
Run and chat with the model
lemonade run user.gemma-7b-it-q4_k_m-Q4_K_M
List all available models
lemonade list
Quick Links

Gemma-7B-it GGUF Quantized

Usage

This model can be used with the latest version of llama.cpp and LM Studio >0.2.16.

Downloads last month
29
GGUF
Model size
9B params
Architecture
gemma
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support