Upload folder using huggingface_hub

Browse files

Files changed (7) hide show

.gitattributes +4 -0
README.md +35 -0
qwen1.5-4b.fp16.bin +3 -0
qwen1.Q4_K_M.gguf +3 -0
qwen1.Q5_K_M.gguf +3 -0
qwen1.Q6_K.gguf +3 -0
qwen1.Q8_0.gguf +3 -0

.gitattributes CHANGED Viewed

@@ -33,3 +33,7 @@ saved_model/**/* filter=lfs diff=lfs merge=lfs -text
 *.zip filter=lfs diff=lfs merge=lfs -text
 *.zst filter=lfs diff=lfs merge=lfs -text
 *tfevents* filter=lfs diff=lfs merge=lfs -text

 *.zip filter=lfs diff=lfs merge=lfs -text
 *.zst filter=lfs diff=lfs merge=lfs -text
 *tfevents* filter=lfs diff=lfs merge=lfs -text
+qwen1.Q4_K_M.gguf filter=lfs diff=lfs merge=lfs -text
+qwen1.Q5_K_M.gguf filter=lfs diff=lfs merge=lfs -text
+qwen1.Q6_K.gguf filter=lfs diff=lfs merge=lfs -text
+qwen1.Q8_0.gguf filter=lfs diff=lfs merge=lfs -text

README.md ADDED Viewed

	@@ -0,0 +1,35 @@

+Thanks to @s3nh for the great quantization notebook code.
+---
+license: openrail
+pipeline_tag: text-generation
+library_name: transformers
+language:
+- en
+---
+## Original model card
+Buy @s3nh a coffee if you like this project ;)
+<a href="https://www.buymeacoffee.com/s3nh"><img src="https://www.buymeacoffee.com/assets/img/guidelines/download-assets-sm-1.svg" alt=""></a>
+#### Description
+GGUF Format model files for [This project](https://huggingface.co/Qwen/Qwen1.5-4B).
+### GGUF Specs
+GGUF is a format based on the existing GGJT, but makes a few changes to the format to make it more extensible and easier to use. The following features are desired:
+Single-file deployment: they can be easily distributed and loaded, and do not require any external files for additional information.
+Extensible: new features can be added to GGML-based executors/new information can be added to GGUF models without breaking compatibility with existing models.
+mmap compatibility: models can be loaded using mmap for fast loading and saving.
+Easy to use: models can be easily loaded and saved using a small amount of code, with no need for external libraries, regardless of the language used.
+Full information: all information needed to load a model is contained in the model file, and no additional information needs to be provided by the user.
+The key difference between GGJT and GGUF is the use of a key-value structure for the hyperparameters (now referred to as metadata), rather than a list of untyped values.
+This allows for new metadata to be added without breaking compatibility with existing models, and to annotate the model with additional information that may be useful for
+inference or for identifying the model.
+# Original model card

qwen1.5-4b.fp16.bin ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:f6c530192fcf357c2838ed79c69c0ed3ad8a3372e0c067b03549fa0998fb3df5
+size 7905600288

qwen1.Q4_K_M.gguf ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:9bce85800c14630df72a66e26606a1c7ec249c71436a63b54faa463c34bc2c82
+size 2452992352

qwen1.Q5_K_M.gguf ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:3fdfc48edc485a8090cf45536057f238c685df27f5f3d1e0cc46d2f2d32c00f0
+size 2837483872

qwen1.Q6_K.gguf ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:1f60f348b30db359448e2a7edafafb356f5898d9f92f16d04fd79d8becec920e
+size 3825197

qwen1.Q8_0.gguf ADDED Viewed

	@@ -0,0 +1,3 @@

+version https://git-lfs.github.com/spec/v1
+oid sha256:105d7cb68c4828c961f5848eb8302206a94c5444575ac04219c1cb3196e49d75
+size 128761856