mradermacher commited on
Commit
1140ff6
·
verified ·
1 Parent(s): f577012

auto-patch README.md

Browse files
Files changed (1) hide show
  1. README.md +118 -0
README.md CHANGED
@@ -1,3 +1,71 @@
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
  <!-- ### quantize_version: 2 -->
2
  <!-- ### output_tensor_quantised: 1 -->
3
  <!-- ### convert_type: hf -->
@@ -7,3 +75,53 @@
7
  <!-- ### quants_skip: -->
8
  <!-- ### skip_mmproj: -->
9
  static quants of https://huggingface.co/ValiantLabs/gemma-4-12B-it-Esper4
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
+ ---
2
+ base_model: ValiantLabs/gemma-4-12B-it-Esper4
3
+ datasets:
4
+ - sequelbox/Mitakihara2-DeepSeek-V4-Pro
5
+ - sequelbox/Tachibana4-DeepSeek-V4-Pro
6
+ - sequelbox/Titanium4-DeepSeek-V4-Pro
7
+ language:
8
+ - en
9
+ library_name: transformers
10
+ license: apache-2.0
11
+ mradermacher:
12
+ readme_rev: 1
13
+ quantized_by: mradermacher
14
+ tags:
15
+ - esper
16
+ - esper-4
17
+ - valiant
18
+ - valiant-labs
19
+ - gemma
20
+ - gemma-4
21
+ - gemma-4-12b
22
+ - gemma-4-12b-it
23
+ - 12b
24
+ - reasoning
25
+ - code
26
+ - code-instruct
27
+ - python
28
+ - typescript
29
+ - javascript
30
+ - java
31
+ - c++
32
+ - c
33
+ - c#
34
+ - rust
35
+ - go
36
+ - haskell
37
+ - dev-ops
38
+ - jenkins
39
+ - terraform
40
+ - ansible
41
+ - docker
42
+ - jenkins
43
+ - kubernetes
44
+ - helm
45
+ - grafana
46
+ - prometheus
47
+ - shell
48
+ - bash
49
+ - azure
50
+ - aws
51
+ - gcp
52
+ - cloud
53
+ - scripting
54
+ - powershell
55
+ - problem-solving
56
+ - architect
57
+ - engineer
58
+ - developer
59
+ - creative
60
+ - analytical
61
+ - expert
62
+ - rationality
63
+ - conversational
64
+ - chat
65
+ - instruct
66
+ ---
67
+ ## About
68
+
69
  <!-- ### quantize_version: 2 -->
70
  <!-- ### output_tensor_quantised: 1 -->
71
  <!-- ### convert_type: hf -->
 
75
  <!-- ### quants_skip: -->
76
  <!-- ### skip_mmproj: -->
77
  static quants of https://huggingface.co/ValiantLabs/gemma-4-12B-it-Esper4
78
+
79
+ <!-- provided-files -->
80
+
81
+ ***For a convenient overview and download list, visit our [model page for this model](https://hf.tst.eu/model#gemma-4-12B-it-Esper4-GGUF).***
82
+
83
+ weighted/imatrix quants seem not to be available (by me) at this time. If they do not show up a week or so after the static ones, I have probably not planned for them. Feel free to request them by opening a Community Discussion.
84
+ ## Usage
85
+
86
+ If you are unsure how to use GGUF files, refer to one of [TheBloke's
87
+ READMEs](https://huggingface.co/TheBloke/KafkaLM-70B-German-V0.1-GGUF) for
88
+ more details, including on how to concatenate multi-part files.
89
+
90
+ ## Provided Quants
91
+
92
+ (sorted by size, not necessarily quality. IQ-quants are often preferable over similar sized non-IQ quants)
93
+
94
+ | Link | Type | Size/GB | Notes |
95
+ |:-----|:-----|--------:|:------|
96
+ | [GGUF](https://huggingface.co/mradermacher/gemma-4-12B-it-Esper4-GGUF/resolve/main/gemma-4-12B-it-Esper4.mmproj-f16.gguf) | mmproj-f16 | 0.2 | multi-modal supplement |
97
+ | [GGUF](https://huggingface.co/mradermacher/gemma-4-12B-it-Esper4-GGUF/resolve/main/gemma-4-12B-it-Esper4.mmproj-Q8_0.gguf) | mmproj-Q8_0 | 0.3 | multi-modal supplement |
98
+ | [GGUF](https://huggingface.co/mradermacher/gemma-4-12B-it-Esper4-GGUF/resolve/main/gemma-4-12B-it-Esper4.Q2_K.gguf) | Q2_K | 4.9 | |
99
+ | [GGUF](https://huggingface.co/mradermacher/gemma-4-12B-it-Esper4-GGUF/resolve/main/gemma-4-12B-it-Esper4.Q3_K_S.gguf) | Q3_K_S | 5.6 | |
100
+ | [GGUF](https://huggingface.co/mradermacher/gemma-4-12B-it-Esper4-GGUF/resolve/main/gemma-4-12B-it-Esper4.Q3_K_M.gguf) | Q3_K_M | 6.2 | lower quality |
101
+ | [GGUF](https://huggingface.co/mradermacher/gemma-4-12B-it-Esper4-GGUF/resolve/main/gemma-4-12B-it-Esper4.Q3_K_L.gguf) | Q3_K_L | 6.7 | |
102
+ | [GGUF](https://huggingface.co/mradermacher/gemma-4-12B-it-Esper4-GGUF/resolve/main/gemma-4-12B-it-Esper4.Q4_K_S.gguf) | Q4_K_S | 7.1 | fast, recommended |
103
+ | [GGUF](https://huggingface.co/mradermacher/gemma-4-12B-it-Esper4-GGUF/resolve/main/gemma-4-12B-it-Esper4.Q4_K_M.gguf) | Q4_K_M | 7.5 | fast, recommended |
104
+ | [GGUF](https://huggingface.co/mradermacher/gemma-4-12B-it-Esper4-GGUF/resolve/main/gemma-4-12B-it-Esper4.Q5_K_S.gguf) | Q5_K_S | 8.4 | |
105
+ | [GGUF](https://huggingface.co/mradermacher/gemma-4-12B-it-Esper4-GGUF/resolve/main/gemma-4-12B-it-Esper4.Q6_K.gguf) | Q6_K | 9.9 | very good quality |
106
+ | [GGUF](https://huggingface.co/mradermacher/gemma-4-12B-it-Esper4-GGUF/resolve/main/gemma-4-12B-it-Esper4.Q8_0.gguf) | Q8_0 | 12.8 | fast, best quality |
107
+
108
+ Here is a handy graph by ikawrakow comparing some lower-quality quant
109
+ types (lower is better):
110
+
111
+ ![image.png](https://www.nethype.de/huggingface_embed/quantpplgraph.png)
112
+
113
+ And here are Artefact2's thoughts on the matter:
114
+ https://gist.github.com/Artefact2/b5f810600771265fc1e39442288e8ec9
115
+
116
+ ## FAQ / Model Request
117
+
118
+ See https://huggingface.co/mradermacher/model_requests for some answers to
119
+ questions you might have and/or if you want some other model quantized.
120
+
121
+ ## Thanks
122
+
123
+ I thank my company, [nethype GmbH](https://www.nethype.de/), for letting
124
+ me use its servers and providing upgrades to my workstation to enable
125
+ this work in my free time.
126
+
127
+ <!-- end -->