Q3.5-4B-OpusGLM-MAX-0731-ablated-GGUF

Q3.5-4B-OpusGLM-MAX-0731-ablated is a reasoning-capable 4B-parameter language model built on top of Qwen/Qwen3.5-4B. The model was trained through a multi-stage training pipeline using General Purpose GLM and Opus reasoning traces, along with additional high-quality reasoning traces, to improve long-form reasoning, mathematical problem solving, scientific analysis, instruction-following capabilities, and general analytical performance.

This model is an experimental release and may generate unexpected behaviors or reasoning artifacts in certain scenarios.

Model Files

File Name Quant Type File Size File Link
Q3.5-4B-OpusGLM-MAX-0731-ablated.BF16.gguf BF16 8.42 GB Download
Q3.5-4B-OpusGLM-MAX-0731-ablated.F16.gguf F16 8.42 GB Download
Q3.5-4B-OpusGLM-MAX-0731-ablated.Q3_K_L.gguf Q3_K_L 2.42 GB Download
Q3.5-4B-OpusGLM-MAX-0731-ablated.Q3_K_M.gguf Q3_K_M 2.26 GB Download
Q3.5-4B-OpusGLM-MAX-0731-ablated.Q3_K_S.gguf Q3_K_S 2.07 GB Download
Q3.5-4B-OpusGLM-MAX-0731-ablated.Q4_K_M.gguf Q4_K_M 2.71 GB Download
Q3.5-4B-OpusGLM-MAX-0731-ablated.Q4_K_S.gguf Q4_K_S 2.56 GB Download
Q3.5-4B-OpusGLM-MAX-0731-ablated.Q5_K_M.gguf Q5_K_M 3.07 GB Download
Q3.5-4B-OpusGLM-MAX-0731-ablated.Q5_K_S.gguf Q5_K_S 2.99 GB Download
Q3.5-4B-OpusGLM-MAX-0731-ablated.Q6_K.gguf Q6_K 3.46 GB Download
Q3.5-4B-OpusGLM-MAX-0731-ablated.Q8_0.gguf Q8_0 4.48 GB Download
Q3.5-4B-OpusGLM-MAX-0731-ablated.mmproj-bf16.gguf mmproj-bf16 676 MB Download
Q3.5-4B-OpusGLM-MAX-0731-ablated.mmproj-f16.gguf mmproj-f16 676 MB Download
Q3.5-4B-OpusGLM-MAX-0731-ablated.mmproj-q8_0.gguf mmproj-q8_0 367 MB Download

llama.cpp

LLM inference in C/C++ — https://github.com/ggml-org/llama.cpp

Downloads last month
4
GGUF
Model size
4B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

3-bit

4-bit

5-bit

6-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for prithivMLmods/Q3.5-4B-OpusGLM-MAX-0731-ablated-GGUF

Finetuned
Qwen/Qwen3.5-4B
Quantized
(1)
this model

Collection including prithivMLmods/Q3.5-4B-OpusGLM-MAX-0731-ablated-GGUF