Xin Liu commited on
Commit
75a3535
1 Parent(s): e2cd77b

Signed-off-by: Xin Liu <sam@secondstate.io>

Files changed (1) hide show
  1. README.md +14 -13
README.md CHANGED
@@ -25,22 +25,23 @@ tags:
25
  ## Original Model
26
 
27
  [sentence-transformers/all-MiniLM-L6-v2](https://huggingface.co/sentence-transformers/all-MiniLM-L6-v2)
28
- <!--
29
  ## Quantized GGUF Models
30
 
31
  | Name | Quant method | Bits | Size | Use case |
32
  | ---- | ---- | ---- | ---- | ----- |
33
- | [openchat-3.5-0106.Q2_K.gguf](https://huggingface.co/second-state/OpenChat-3.5-0106-GGUF/blob/main/openchat-3.5-0106-Q2_K.gguf) | Q2_K | 2 | 3.08 GB| smallest, significant quality loss - not recommended for most purposes |
34
- | [openchat-3.5-0106.Q3_K_L.gguf](https://huggingface.co/second-state/OpenChat-3.5-0106-GGUF/blob/main/openchat-3.5-0106-Q3_K_L.gguf) | Q3_K_L | 3 | 3.82 GB| small, substantial quality loss |
35
- | [openchat-3.5-0106.Q3_K_M.gguf](https://huggingface.co/second-state/OpenChat-3.5-0106-GGUF/blob/main/openchat-3.5-0106-Q3_K_M.gguf) | Q3_K_M | 3 | 3.52 GB| very small, high quality loss |
36
- | [openchat-3.5-0106.Q3_K_S.gguf](https://huggingface.co/second-state/OpenChat-3.5-0106-GGUF/blob/main/openchat-3.5-0106-Q3_K_S.gguf) | Q3_K_S | 3 | 3.16 GB| very small, high quality loss |
37
- | [openchat-3.5-0106.Q4_0.gguf](https://huggingface.co/second-state/OpenChat-3.5-0106-GGUF/blob/main/openchat-3.5-0106-Q4_0.gguf) | Q4_0 | 4 | 4.11 GB| legacy; small, very high quality loss - prefer using Q3_K_M |
38
- | [openchat-3.5-0106.Q4_K_M.gguf](https://huggingface.co/second-state/OpenChat-3.5-0106-GGUF/blob/main/openchat-3.5-0106-Q4_K_M.gguf) | Q4_K_M | 4 | 4.37 GB| medium, balanced quality - recommended |
39
- | [openchat-3.5-0106.Q4_K_S.gguf](https://huggingface.co/second-state/OpenChat-3.5-0106-GGUF/blob/main/openchat-3.5-0106-Q4_K_S.gguf) | Q4_K_S | 4 | 4.14 GB| small, greater quality loss |
40
- | [openchat-3.5-0106.Q5_0.gguf](https://huggingface.co/second-state/OpenChat-3.5-0106-GGUF/blob/main/openchat-3.5-0106-Q5_0.gguf) | Q5_0 | 5 | 5.00 GB| legacy; medium, balanced quality - prefer using Q4_K_M |
41
- | [openchat-3.5-0106.Q5_K_M.gguf](https://huggingface.co/second-state/OpenChat-3.5-0106-GGUF/blob/main/openchat-3.5-0106-Q5_K_M.gguf) | Q5_K_M | 5 | 5.13 GB| large, very low quality loss - recommended |
42
- | [openchat-3.5-0106.Q5_K_S.gguf](https://huggingface.co/second-state/OpenChat-3.5-0106-GGUF/blob/main/openchat-3.5-0106-Q5_K_S.gguf) | Q5_K_S | 5 | 5.00 GB| large, low quality loss - recommended |
43
- | [openchat-3.5-0106.Q6_K.gguf](https://huggingface.co/second-state/OpenChat-3.5-0106-GGUF/blob/main/openchat-3.5-0106-Q6_K.gguf) | Q6_K | 6 | 5.94 GB| very large, extremely low quality loss |
44
- | [openchat-3.5-0106.Q8_0.gguf](https://huggingface.co/second-state/OpenChat-3.5-0106-GGUF/blob/main/openchat-3.5-0106-Q8_0.gguf) | Q8_0 | 8 | 7.70 GB| very large, extremely low quality loss - not recommended | -->
 
45
 
46
  *Quantized with llama.cpp b2334*
 
25
  ## Original Model
26
 
27
  [sentence-transformers/all-MiniLM-L6-v2](https://huggingface.co/sentence-transformers/all-MiniLM-L6-v2)
28
+
29
  ## Quantized GGUF Models
30
 
31
  | Name | Quant method | Bits | Size | Use case |
32
  | ---- | ---- | ---- | ---- | ----- |
33
+ | [all-MiniLM-L6-v2-Q2_K.gguf](https://huggingface.co/second-state/All-MiniLM-L6-v2-Embedding-GGUF/blob/main/all-MiniLM-L6-v2-Q2_K.gguf) | Q2_K | 2 | 19.2 MB| smallest, significant quality loss - not recommended for most purposes |
34
+ | [all-MiniLM-L6-v2-Q3_K_L.gguf](https://huggingface.co/second-state/All-MiniLM-L6-v2-Embedding-GGUF/blob/main/all-MiniLM-L6-v2-Q3_K_L.gguf) | Q3_K_L | 3 | 20.5 MB| small, substantial quality loss |
35
+ | [all-MiniLM-L6-v2-Q3_K_M.gguf](https://huggingface.co/second-state/All-MiniLM-L6-v2-Embedding-GGUF/blob/main/all-MiniLM-L6-v2-Q3_K_M.gguf) | Q3_K_M | 3 | 19.9 MB| very small, high quality loss |
36
+ | [all-MiniLM-L6-v2-Q3_K_S.gguf](https://huggingface.co/second-state/All-MiniLM-L6-v2-Embedding-GGUF/blob/main/all-MiniLM-L6-v2-Q3_K_S.gguf) | Q3_K_S | 3 | 19.2 MB| very small, high quality loss |
37
+ | [all-MiniLM-L6-v2-Q4_0.gguf](https://huggingface.co/second-state/All-MiniLM-L6-v2-Embedding-GGUF/blob/main/all-MiniLM-L6-v2-Q4_0.gguf) | Q4_0 | 4 | 19.7 MB| legacy; small, very high quality loss - prefer using Q3_K_M |
38
+ | [all-MiniLM-L6-v2-Q4_K_M.gguf](https://huggingface.co/second-state/All-MiniLM-L6-v2-Embedding-GGUF/blob/main/all-MiniLM-L6-v2-Q4_K_M.gguf) | Q4_K_M | 4 | 21 MB| medium, balanced quality - recommended |
39
+ | [all-MiniLM-L6-v2-Q4_K_S.gguf](https://huggingface.co/second-state/All-MiniLM-L6-v2-Embedding-GGUF/blob/main/all-MiniLM-L6-v2-Q4_K_S.gguf) | Q4_K_S | 4 | 20.7 MB| small, greater quality loss |
40
+ | [all-MiniLM-L6-v2-Q5_0.gguf](https://huggingface.co/second-state/All-MiniLM-L6-v2-Embedding-GGUF/blob/main/all-MiniLM-L6-v2-Q5_0.gguf) | Q5_0 | 5 | 21 MB| legacy; medium, balanced quality - prefer using Q4_K_M |
41
+ | [all-MiniLM-L6-v2-Q5_K_M.gguf](https://huggingface.co/second-state/All-MiniLM-L6-v2-Embedding-GGUF/blob/main/all-MiniLM-L6-v2-Q5_K_M.gguf) | Q5_K_M | 5 | 21.7 MB| large, very low quality loss - recommended |
42
+ | [all-MiniLM-L6-v2-Q5_K_S.gguf](https://huggingface.co/second-state/All-MiniLM-L6-v2-Embedding-GGUF/blob/main/all-MiniLM-L6-v2-Q5_K_S.gguf) | Q5_K_S | 5 | 21.5 MB| large, low quality loss - recommended |
43
+ | [all-MiniLM-L6-v2-Q6_K.gguf](https://huggingface.co/second-state/All-MiniLM-L6-v2-Embedding-GGUF/blob/main/all-MiniLM-L6-v2-Q6_K.gguf) | Q6_K | 6 | 24.2 MB| very large, extremely low quality loss |
44
+ | [all-MiniLM-L6-v2-Q8_0.gguf](https://huggingface.co/second-state/All-MiniLM-L6-v2-Embedding-GGUF/blob/main/all-MiniLM-L6-v2-Q8_0.gguf) | Q8_0 | 8 | 25 MB| very large, extremely low quality loss - not recommended |
45
+ | [all-MiniLM-L6-v2-ggml-model-f16.gguf](https://huggingface.co/second-state/All-MiniLM-L6-v2-Embedding-GGUF/blob/main/all-MiniLM-L6-v2-ggml-model-f16.gguf) | Q8_0 | 8 | 45.9 MB| very large, extremely low quality loss - not recommended |
46
 
47
  *Quantized with llama.cpp b2334*