Upload folder using huggingface_hub

Browse files

Files changed (7) hide show

.gitattributes +5 -0
README.md +45 -0
utena-7b-v3.Q3_K_S.gguf +3 -0
utena-7b-v3.Q4_K_M.gguf +3 -0
utena-7b-v3.Q5_K_M.gguf +3 -0
utena-7b-v3.Q6_K.gguf +3 -0
utena-7b-v3.Q8_0.gguf +3 -0

.gitattributes CHANGED Viewed

@@ -33,3 +33,8 @@ saved_model/**/* filter=lfs diff=lfs merge=lfs -text
 *.zip filter=lfs diff=lfs merge=lfs -text
 *.zst filter=lfs diff=lfs merge=lfs -text
 *tfevents* filter=lfs diff=lfs merge=lfs -text

 *.zip filter=lfs diff=lfs merge=lfs -text
 *.zst filter=lfs diff=lfs merge=lfs -text
 *tfevents* filter=lfs diff=lfs merge=lfs -text
+utena-7b-v3.Q3_K_S.gguf filter=lfs diff=lfs merge=lfs -text
+utena-7b-v3.Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
+utena-7b-v3.Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
+utena-7b-v3.Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
+utena-7b-v3.Q8_0.gguf filter=lfs diff=lfs merge=lfs -text

README.md ADDED Viewed

	@@ -0,0 +1,45 @@

+---
+license: openrail
+pipeline_tag: text-generation
+library_name: transformers
+language:
+- zh
+- en
+---
+## Original model card
+Buy me a coffee if you like this project ;)
+<a href="https://www.buymeacoffee.com/s3nh"><img src="https://www.buymeacoffee.com/assets/img/guidelines/download-assets-sm-1.svg" alt=""></a>
+#### Description
+GGUF Format model files for [This project](https://huggingface.co/AI-B/UTENA-7B-V3).
+### GGUF Specs
+GGUF is a format based on the existing GGJT, but makes a few changes to the format to make it more extensible and easier to use. The following features are desired:
+Single-file deployment: they can be easily distributed and loaded, and do not require any external files for additional information.
+Extensible: new features can be added to GGML-based executors/new information can be added to GGUF models without breaking compatibility with existing models.
+mmap compatibility: models can be loaded using mmap for fast loading and saving.
+Easy to use: models can be easily loaded and saved using a small amount of code, with no need for external libraries, regardless of the language used.
+Full information: all information needed to load a model is contained in the model file, and no additional information needs to be provided by the user.
+The key difference between GGJT and GGUF is the use of a key-value structure for the hyperparameters (now referred to as metadata), rather than a list of untyped values.
+This allows for new metadata to be added without breaking compatibility with existing models, and to annotate the model with additional information that may be useful for
+inference or for identifying the model.
+### inference
+ User: Tell me story about what is an quantization and what do we need to build.
+ArtificialLinguist: The term "quantization" refers to the process of converting a continuous signal, such as sound or an image, into a discrete set of values. In digital signal processing (DSP), this is necessary because computers only understand discrete values rather than continuous signals. Quantization involves mapping a range of input values to a finite set of output values, which results in some loss of information due to the rounding off of values to the nearest quantized value.
+Building a quantizer requires specifying the input signal (e.g., audio or image data), the range of possible input
+# Original model card

utena-7b-v3.Q3_K_S.gguf ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:7504f1a4d66c8a0cac5d04733527780c77b2bbb103b3fa942ef33dae7c28a034
+size 3164567200

utena-7b-v3.Q4_K_M.gguf ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:ee172587b46d1854802823aabad190ebfe175af07875489c79877d82be252698
+size 4368438944

utena-7b-v3.Q5_K_M.gguf ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:1a1da110bcb2dc068271a8a01e4b64860bcadefb4c31d74b398336d2f3737f2c
+size 5131409056

utena-7b-v3.Q6_K.gguf ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:cea1d343f1808ed8b03cd857fb542f3866be886d0cec72b45f2654f983742a9b
+size 5942064800

utena-7b-v3.Q8_0.gguf ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:0f05b9a359e85ddd801cc0f33dcaa0ea230b87d6afef95004aaeec96931477d1
+size 7695857312