h4rm0n1c commited on
Commit
c44d27f
·
verified ·
1 Parent(s): 8378a92

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +4 -0
README.md CHANGED
@@ -15,6 +15,10 @@ tags:
15
 
16
  # Qwen3.5-24B-A3B-REAP-0.32 IQ4_NL GGUF
17
 
 
 
 
 
18
  GGUF quantization of [sandeshrajx/Qwen3.5-24B-A3B-REAP-0.32](https://huggingface.co/sandeshrajx/Qwen3.5-24B-A3B-REAP-0.32) — a REAP-pruned 24B total, **~3B active** MoE model.
19
 
20
  **REAP** (Razor Edge And Pruning, [arxiv:2510.13999](https://arxiv.org/abs/2510.13999)) is a structured pruning technique that reduces the base Qwen3.5 model while preserving capability.
 
15
 
16
  # Qwen3.5-24B-A3B-REAP-0.32 IQ4_NL GGUF
17
 
18
+ Quanter's Note: This model is the parent of every solid performing reasoning distill I've benched out of 40 or so models in the last month, this thing has some solid potential for a model that has been given a partial lobotomy, awfully impressive.
19
+ Perplexity test I devised is Extremely Tough on models, high convergence rate on alien/unusual codebases, plenty of potential for finetunes.
20
+ I felt the base model deserved some attention alongside its author and I'm glad to see people downloading it.
21
+
22
  GGUF quantization of [sandeshrajx/Qwen3.5-24B-A3B-REAP-0.32](https://huggingface.co/sandeshrajx/Qwen3.5-24B-A3B-REAP-0.32) — a REAP-pruned 24B total, **~3B active** MoE model.
23
 
24
  **REAP** (Razor Edge And Pruning, [arxiv:2510.13999](https://arxiv.org/abs/2510.13999)) is a structured pruning technique that reduces the base Qwen3.5 model while preserving capability.