mradermacher commited on
Commit
9be428c
·
verified ·
1 Parent(s): cf76cd8

auto-patch README.md

Browse files
Files changed (1) hide show
  1. README.md +9 -1
README.md CHANGED
@@ -5,6 +5,13 @@ language:
5
  library_name: transformers
6
  license: apache-2.0
7
  quantized_by: mradermacher
 
 
 
 
 
 
 
8
  ---
9
  ## About
10
 
@@ -16,7 +23,7 @@ quantized_by: mradermacher
16
  static quants of https://huggingface.co/prithivMLmods/Sombrero-QwQ-32B-Elite9
17
 
18
  <!-- provided-files -->
19
- weighted/imatrix quants seem not to be available (by me) at this time. If they do not show up a week or so after the static ones, I have probably not planned for them. Feel free to request them by opening a Community Discussion.
20
  ## Usage
21
 
22
  If you are unsure how to use GGUF files, refer to one of [TheBloke's
@@ -33,6 +40,7 @@ more details, including on how to concatenate multi-part files.
33
  | [GGUF](https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-GGUF/resolve/main/Sombrero-QwQ-32B-Elite9.Q3_K_S.gguf) | Q3_K_S | 14.5 | |
34
  | [GGUF](https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-GGUF/resolve/main/Sombrero-QwQ-32B-Elite9.Q3_K_M.gguf) | Q3_K_M | 16.0 | lower quality |
35
  | [GGUF](https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-GGUF/resolve/main/Sombrero-QwQ-32B-Elite9.Q3_K_L.gguf) | Q3_K_L | 17.3 | |
 
36
  | [GGUF](https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-GGUF/resolve/main/Sombrero-QwQ-32B-Elite9.Q4_K_S.gguf) | Q4_K_S | 18.9 | fast, recommended |
37
  | [GGUF](https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-GGUF/resolve/main/Sombrero-QwQ-32B-Elite9.Q4_K_M.gguf) | Q4_K_M | 20.0 | fast, recommended |
38
  | [GGUF](https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-GGUF/resolve/main/Sombrero-QwQ-32B-Elite9.Q5_K_S.gguf) | Q5_K_S | 22.7 | |
 
5
  library_name: transformers
6
  license: apache-2.0
7
  quantized_by: mradermacher
8
+ tags:
9
+ - text-generation-inference
10
+ - code
11
+ - StreamlinedMemory
12
+ - Qwen
13
+ - QwQ
14
+ - General-purpose
15
  ---
16
  ## About
17
 
 
23
  static quants of https://huggingface.co/prithivMLmods/Sombrero-QwQ-32B-Elite9
24
 
25
  <!-- provided-files -->
26
+ weighted/imatrix quants are available at https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-i1-GGUF
27
  ## Usage
28
 
29
  If you are unsure how to use GGUF files, refer to one of [TheBloke's
 
40
  | [GGUF](https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-GGUF/resolve/main/Sombrero-QwQ-32B-Elite9.Q3_K_S.gguf) | Q3_K_S | 14.5 | |
41
  | [GGUF](https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-GGUF/resolve/main/Sombrero-QwQ-32B-Elite9.Q3_K_M.gguf) | Q3_K_M | 16.0 | lower quality |
42
  | [GGUF](https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-GGUF/resolve/main/Sombrero-QwQ-32B-Elite9.Q3_K_L.gguf) | Q3_K_L | 17.3 | |
43
+ | [GGUF](https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-GGUF/resolve/main/Sombrero-QwQ-32B-Elite9.IQ4_XS.gguf) | IQ4_XS | 18.0 | |
44
  | [GGUF](https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-GGUF/resolve/main/Sombrero-QwQ-32B-Elite9.Q4_K_S.gguf) | Q4_K_S | 18.9 | fast, recommended |
45
  | [GGUF](https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-GGUF/resolve/main/Sombrero-QwQ-32B-Elite9.Q4_K_M.gguf) | Q4_K_M | 20.0 | fast, recommended |
46
  | [GGUF](https://huggingface.co/mradermacher/Sombrero-QwQ-32B-Elite9-GGUF/resolve/main/Sombrero-QwQ-32B-Elite9.Q5_K_S.gguf) | Q5_K_S | 22.7 | |