--- language: - en - zh - ko - ja license: apache-2.0 base_model: Jackrong/Gemopus-4-26B-A4B-it library_name: mlx tags: - mlx - gemma - gemma4 - instruction-tuned - reasoning - alignment pipeline_tag: text-generation --- # 🦆 zecanard/Gemopus-4-26B-A4B-it-MLX-3bit-mixed_3_6 [This model](https://huggingface.co/zecanard/Gemopus-4-26B-A4B-it-MLX-3bit-mixed_3_6) was converted to MLX from [`Jackrong/Gemopus-4-26B-A4B-it`](https://huggingface.co/Jackrong/Gemopus-4-26B-A4B-it) using `mlx-vlm` version **0.4.4**. Please refer to the [original model card](https://huggingface.co/Jackrong/Gemopus-4-26B-A4B-it) for more details. ## 🌟 Quality Quantized language model with an effective **4.295 bits per weight**. `mlx_vlm.convert --quantize --q-group-size 32 --quant-predicate mixed_3_6` ## 🛠️ Customizations This quant is aware of the current date, and also enables thinking (if available). You may disable this behavior by deleting the following line from the chat template: `{%- set enable_thinking = true %}` You may also need to adjust your environment’s **Reasoning Section Parsing** to recognize `<|channel>thought` as the **Start String**, and `` as the **End String**. ## 🖥️ Use with `mlx` ```bash pip install -U mlx-vlm ``` ```bash mlx_vlm.generate --model zecanard/Gemopus-4-26B-A4B-it-MLX-3bit-mixed_3_6 --max-tokens 100 --temperature 0 --prompt "Describe this image." --image ```