--- license: apache-2.0 base_model: - zai-org/GLM-4.7 pipeline_tag: text-generation --- ## Model Description A quantization setup used for GLM-4.7: - Weights: NVFP4 - KV cache: FP8 - Tooling: NVIDIA/Model-Optimizer - Deploy with TensorRT-LLM