Yi-9B-200K-GGUF / README.md
bartowski's picture
Llamacpp quants
5b750d1 verified
metadata
license: other
license_name: yi-license
license_link: LICENSE
widget:
  - example_title: Yi-34B-Chat
    text: hi
    output:
      text: ' Hello! How can I assist you today?'
  - example_title: Yi-34B
    text: >-
      There's a place where time stands still. A place of breath taking wonder,
      but also
    output:
      text: >2-
         an eerie sense that something is just not right…
        Between the two worlds lies The Forgotten Kingdom - home to creatures
        long since thought extinct and ancient magic so strong it defies belief!
        Only here can you find what has been lost for centuries: An Elixir Of
        Life which will restore youth and vitality if only those who seek its
        power are brave enough to face up against all manner of dangers lurking
        in this mysterious land! But beware; some say there may even exist
        powerful entities beyond our comprehension whose intentions towards
        humanity remain unclear at best ---- they might want nothing more than
        destruction itself rather then anything else from their quest after
        immortality (and maybe someone should tell them about modern medicine)?
        In any event though  one thing remains true regardless : whether or not
        success comes easy depends entirely upon how much effort we put into
        conquering whatever challenges lie ahead along with having faith deep
        down inside ourselves too ;) So let’s get started now shall We?
pipeline_tag: text-generation
quantized_by: bartowski

Llamacpp Quantizations of Yi-9B-200K

Using llama.cpp release b2440 for quantization.

Original model: https://huggingface.co/01-ai/Yi-9B-200K

Download a file (not the whole branch) from below:

Filename Quant type File Size Description
Yi-9B-200K-Q8_0.gguf Q8_0 9.38GB Extremely high quality, generally unneeded but max available quant.
Yi-9B-200K-Q6_K.gguf Q6_K 7.24GB Very high quality, near perfect, recommended.
Yi-9B-200K-Q5_K_M.gguf Q5_K_M 6.25GB High quality, very usable.
Yi-9B-200K-Q5_K_S.gguf Q5_K_S 6.10GB High quality, very usable.
Yi-9B-200K-Q5_0.gguf Q5_0 6.10GB High quality, older format, generally not recommended.
Yi-9B-200K-Q4_K_M.gguf Q4_K_M 5.32GB Good quality, similar to 4.25 bpw.
Yi-9B-200K-Q4_K_S.gguf Q4_K_S 5.07GB Slightly lower quality with small space savings.
Yi-9B-200K-Q4_0.gguf Q4_0 5.03GB Decent quality, older format, generally not recommended.
Yi-9B-200K-Q3_K_L.gguf Q3_K_L 4.69GB Lower quality but usable, good for low RAM availability.
Yi-9B-200K-Q3_K_M.gguf Q3_K_M 4.32GB Even lower quality.
Yi-9B-200K-Q3_K_S.gguf Q3_K_S 3.89GB Low quality, not recommended.
Yi-9B-200K-Q2_K.gguf Q2_K 3.35GB Extremely low quality, not recommended.

Want to support my work? Visit my ko-fi page here: https://ko-fi.com/bartowski