batiai/Qwen3.8-27B-GGUF
Text Generation β’ 27B β’ Updated β’ 2.83k β’ 1
Qwen's newest 27B: vision, 262K context, Korean-verified. Start with IQ4_XS β we measured every tier and Q3_K_M is smaller yet slower. Needs 24GB+.
Note Six quants + vision projector. Per-quant Apple Silicon and CUDA speeds in the card β including what we got wrong the first time.
Note MoE alternative β 3B active, 45 t/s on M4 Max vs 16.6 for a dense 27B. The pick for 16GB Macs or when throughput beats peak quality.
Note Previous generation, same class. Our most-downloaded model β a useful baseline for comparison.