mradermacher
/

MegaDolphin-120b-i1-GGUF

Inference Endpoints

Model card Files Files and versions Community

mradermacher commited on May 6

Commit

0248395

•

1 Parent(s): e063159

auto-patch README.md

Files changed (1) hide show

README.md +0 -1

README.md CHANGED Viewed

@@ -40,7 +40,6 @@ more details, including on how to concatenate multi-part files.
 | [PART 1](https://huggingface.co/mradermacher/MegaDolphin-120b-i1-GGUF/resolve/main/MegaDolphin-120b.i1-Q4_K_S.gguf.split-aa) [PART 2](https://huggingface.co/mradermacher/MegaDolphin-120b-i1-GGUF/resolve/main/MegaDolphin-120b.i1-Q4_K_S.gguf.split-ab) | i1-Q4_K_S | 68.7 | optimal size/speed/quality |
 | [PART 1](https://huggingface.co/mradermacher/MegaDolphin-120b-i1-GGUF/resolve/main/MegaDolphin-120b.i1-Q4_K_M.gguf.split-aa) [PART 2](https://huggingface.co/mradermacher/MegaDolphin-120b-i1-GGUF/resolve/main/MegaDolphin-120b.i1-Q4_K_M.gguf.split-ab) | i1-Q4_K_M | 72.6 | fast, recommended |
 Here is a handy graph by ikawrakow comparing some lower-quality quant
 types (lower is better):

 | [PART 1](https://huggingface.co/mradermacher/MegaDolphin-120b-i1-GGUF/resolve/main/MegaDolphin-120b.i1-Q4_K_S.gguf.split-aa) [PART 2](https://huggingface.co/mradermacher/MegaDolphin-120b-i1-GGUF/resolve/main/MegaDolphin-120b.i1-Q4_K_S.gguf.split-ab) | i1-Q4_K_S | 68.7 | optimal size/speed/quality |
 | [PART 1](https://huggingface.co/mradermacher/MegaDolphin-120b-i1-GGUF/resolve/main/MegaDolphin-120b.i1-Q4_K_M.gguf.split-aa) [PART 2](https://huggingface.co/mradermacher/MegaDolphin-120b-i1-GGUF/resolve/main/MegaDolphin-120b.i1-Q4_K_M.gguf.split-ab) | i1-Q4_K_M | 72.6 | fast, recommended |
 Here is a handy graph by ikawrakow comparing some lower-quality quant
 types (lower is better):