gemma-4-E4B-it-GGUF / scores /gemma-4-E4B-it-q2_k.arc
eaddario's picture
Generate Perplexity, KLD, ARC, HellaSwag, MMLU, Truthful QA and WinoGrande scores
2c7e1e2 verified
Raw
History Blame Contribute Delete
1.11 kB
llama_model_loader: loaded meta data with 49 key-value pairs and 720 tensors from gemma-4-E4B-it-WIP/gemma-4-E4B-it-Q2_K.gguf (version GGUF V3 (latest))
llama_model_loader: - type f32: 339 tensors
llama_model_loader: - type f16: 1 tensors
llama_model_loader: - type q4_1: 7 tensors
llama_model_loader: - type q2_K: 112 tensors
llama_model_loader: - type q3_K: 36 tensors
llama_model_loader: - type q4_K: 15 tensors
llama_model_loader: - type q5_K: 2 tensors
llama_model_loader: - type iq2_xxs: 94 tensors
llama_model_loader: - type iq2_xs: 38 tensors
llama_model_loader: - type iq3_xxs: 38 tensors
llama_model_loader: - type iq4_xs: 2 tensors
llama_model_loader: - type iq1_m: 36 tensors
print_info: file format = GGUF V3 (latest)
print_info: file type = IQ2_S - 2.5 bpw
print_info: file size = 2.19 GiB (2.50 BPW)
multiple_choice_score: there are 869 tasks in prompt
multiple_choice_score: selecting 750 random tasks from 869 tasks available
multiple_choice_score : calculating TruthfulQA score over 750 tasks.
Final result: 34.8000 +/- 1.7405
Random chance: 25.0083 +/- 1.5824