fraQtl/Qwen3.6-35B-A3B-Hi-Fi-GGUF
Text Generation • 35B • Updated • 942 • 4
Calibration-aware Q4-class GGUFs at identical size to leading community quants, with measured KLD receipts.
Note ≈23% lower KLD than the leading public Q4_K_M at identical file size — MoE-aware per-tensor policy
Note Datacenter variant: 1.49× decode via MTP speculative decoding (A100-80GB)
Note 36% closer to the original than the community Q4_K_M at identical size — plus a 17.5%-smaller phone artifact
Note −9% KLD to the original at same size — with the GSM8K loss disclosed up front (honest split)