Upload folder using huggingface_hub
Browse files- .gitattributes +7 -0
- Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q2_K.gguf +3 -0
- Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q3_K_S.gguf +3 -0
- Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q4_K_M.gguf +3 -0
- Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q4_K_S.gguf +3 -0
- Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q5_K_S.gguf +3 -0
- Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q6_K.gguf +3 -0
- Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q8_0.gguf +3 -0
- README.md +47 -0
.gitattributes
CHANGED
@@ -33,3 +33,10 @@ saved_model/**/* filter=lfs diff=lfs merge=lfs -text
|
|
33 |
*.zip filter=lfs diff=lfs merge=lfs -text
|
34 |
*.zst filter=lfs diff=lfs merge=lfs -text
|
35 |
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
33 |
*.zip filter=lfs diff=lfs merge=lfs -text
|
34 |
*.zst filter=lfs diff=lfs merge=lfs -text
|
35 |
*tfevents* filter=lfs diff=lfs merge=lfs -text
|
36 |
+
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q2_K.gguf filter=lfs diff=lfs merge=lfs -text
|
37 |
+
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
38 |
+
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
|
39 |
+
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q4_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
40 |
+
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q5_K_S.gguf filter=lfs diff=lfs merge=lfs -text
|
41 |
+
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
|
42 |
+
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q8_0.gguf filter=lfs diff=lfs merge=lfs -text
|
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q2_K.gguf
ADDED
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
1 |
+
version https://git-lfs.github.com/spec/v1
|
2 |
+
oid sha256:984f18444729a32a86b5a28c27bc7b98772d0ffbf156eb9e488c59a4df1948f7
|
3 |
+
size 482142720
|
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q3_K_S.gguf
ADDED
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
1 |
+
version https://git-lfs.github.com/spec/v1
|
2 |
+
oid sha256:b8f7b29cf5af23e2508cf622c6e9b1922b4f827d6a2aaeae6f747aab64fd9cc9
|
3 |
+
size 499341824
|
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q4_K_M.gguf
ADDED
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
1 |
+
version https://git-lfs.github.com/spec/v1
|
2 |
+
oid sha256:1e3f36e1225cfada320caab6214b54b5e78615c4064ff7a045341dff4d06b3e1
|
3 |
+
size 667814400
|
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q4_K_S.gguf
ADDED
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
1 |
+
version https://git-lfs.github.com/spec/v1
|
2 |
+
oid sha256:a4967adbaa3394127dee1a36e28cee6a19e98ae2782ecb26f0d531bb2668059f
|
3 |
+
size 642755072
|
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q5_K_S.gguf
ADDED
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
1 |
+
version https://git-lfs.github.com/spec/v1
|
2 |
+
oid sha256:d62ad751a4fd2e1b7d7b88318cdc3d5958adcf29c67fcd0e82b17f568ff0161f
|
3 |
+
size 766028288
|
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q6_K.gguf
ADDED
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
1 |
+
version https://git-lfs.github.com/spec/v1
|
2 |
+
oid sha256:3d6faa0a959b9abaa643b076d50267f624fb1458cfdf480f13ec30712c696f5b
|
3 |
+
size 903412224
|
Doctor-Shotgun-TinyLlama-1.1B-32k-Instruct.Q8_0.gguf
ADDED
@@ -0,0 +1,3 @@
|
|
|
|
|
|
|
|
|
1 |
+
version https://git-lfs.github.com/spec/v1
|
2 |
+
oid sha256:a2c38935ab95c0dcf68aa26dfa69f6b9c1ae664537fc21fb827860556e3b9585
|
3 |
+
size 1169807872
|
README.md
ADDED
@@ -0,0 +1,47 @@
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
1 |
+
|
2 |
+
---
|
3 |
+
license: openrail
|
4 |
+
pipeline_tag: text-generation
|
5 |
+
library_name: transformers
|
6 |
+
language:
|
7 |
+
- zh
|
8 |
+
- en
|
9 |
+
---
|
10 |
+
|
11 |
+
|
12 |
+
## Original model card
|
13 |
+
|
14 |
+
Buy me a coffee if you like this project ;)
|
15 |
+
<a href="https://www.buymeacoffee.com/s3nh"><img src="https://www.buymeacoffee.com/assets/img/guidelines/download-assets-sm-1.svg" alt=""></a>
|
16 |
+
|
17 |
+
#### Description
|
18 |
+
|
19 |
+
GGUF Format model files for [This project](https://huggingface.co/Doctor-Shotgun/TinyLlama-1.1B-32k-Instruct).
|
20 |
+
|
21 |
+
### GGUF Specs
|
22 |
+
|
23 |
+
GGUF is a format based on the existing GGJT, but makes a few changes to the format to make it more extensible and easier to use. The following features are desired:
|
24 |
+
|
25 |
+
Single-file deployment: they can be easily distributed and loaded, and do not require any external files for additional information.
|
26 |
+
Extensible: new features can be added to GGML-based executors/new information can be added to GGUF models without breaking compatibility with existing models.
|
27 |
+
mmap compatibility: models can be loaded using mmap for fast loading and saving.
|
28 |
+
Easy to use: models can be easily loaded and saved using a small amount of code, with no need for external libraries, regardless of the language used.
|
29 |
+
Full information: all information needed to load a model is contained in the model file, and no additional information needs to be provided by the user.
|
30 |
+
The key difference between GGJT and GGUF is the use of a key-value structure for the hyperparameters (now referred to as metadata), rather than a list of untyped values.
|
31 |
+
This allows for new metadata to be added without breaking compatibility with existing models, and to annotate the model with additional information that may be useful for
|
32 |
+
inference or for identifying the model.
|
33 |
+
|
34 |
+
### Perplexity params
|
35 |
+
|
36 |
+
Model Measure Q2_K Q3_K_S Q3_K_M Q3_K_L Q4_0 Q4_1 Q4_K_S Q4_K_M Q5_0 Q5_1 Q5_K_S Q5_K_M Q6_K Q8_0 F16
|
37 |
+
7B perplexity 6.7764 6.4571 6.1503 6.0869 6.1565 6.0912 6.0215 5.9601 5.9862 5.9481 5.9419 5.9208 5.9110 5.9070 5.9066
|
38 |
+
13B perplexity 5.8545 5.6033 5.4498 5.4063 5.3860 5.3608 5.3404 5.3002 5.2856 5.2706 5.2785 5.2638 5.2568 5.2548 5.2543
|
39 |
+
|
40 |
+
|
41 |
+
|
42 |
+
### inference
|
43 |
+
|
44 |
+
|
45 |
+
TODO
|
46 |
+
|
47 |
+
# Original model card
|