Edit model card
YAML Metadata Warning: empty or missing yaml metadata in repo card (https://huggingface.co/docs/hub/model-cards#model-card-metadata)

Quantization made by Richard Erkhov.

Github

Discord

Request more models

Yi-1.5-34B - GGUF

Name Quant method Size
Yi-1.5-34B.Q2_K.gguf Q2_K 11.94GB
Yi-1.5-34B.IQ3_XS.gguf IQ3_XS 13.26GB
Yi-1.5-34B.IQ3_S.gguf IQ3_S 13.99GB
Yi-1.5-34B.Q3_K_S.gguf Q3_K_S 13.93GB
Yi-1.5-34B.IQ3_M.gguf IQ3_M 14.5GB
Yi-1.5-34B.Q3_K.gguf Q3_K 15.51GB
Yi-1.5-34B.Q3_K_M.gguf Q3_K_M 15.51GB
Yi-1.5-34B.Q3_K_L.gguf Q3_K_L 16.89GB
Yi-1.5-34B.IQ4_XS.gguf IQ4_XS 17.36GB
Yi-1.5-34B.Q4_0.gguf Q4_0 18.13GB
Yi-1.5-34B.IQ4_NL.gguf IQ4_NL 18.3GB
Yi-1.5-34B.Q4_K_S.gguf Q4_K_S 18.25GB
Yi-1.5-34B.Q4_K.gguf Q4_K 19.24GB
Yi-1.5-34B.Q4_K_M.gguf Q4_K_M 19.24GB
Yi-1.5-34B.Q4_1.gguf Q4_1 20.1GB
Yi-1.5-34B.Q5_0.gguf Q5_0 22.08GB
Yi-1.5-34B.Q5_K_S.gguf Q5_K_S 22.08GB
Yi-1.5-34B.Q5_K.gguf Q5_K 22.65GB
Yi-1.5-34B.Q5_K_M.gguf Q5_K_M 22.65GB
Yi-1.5-34B.Q5_1.gguf Q5_1 24.05GB
Yi-1.5-34B.Q6_K.gguf Q6_K 26.28GB
Yi-1.5-34B.Q8_0.gguf Q8_0 34.03GB

Original model description:

license: apache-2.0

πŸ™ GitHub β€’ πŸ‘Ύ Discord β€’ 🐀 Twitter β€’ πŸ’¬ WeChat
πŸ“ Paper β€’ πŸ’ͺ Tech Blog β€’ πŸ™Œ FAQ β€’ πŸ“— Learning Hub

Intro

Yi-1.5 is an upgraded version of Yi. It is continuously pre-trained on Yi with a high-quality corpus of 500B tokens and fine-tuned on 3M diverse fine-tuning samples.

Compared with Yi, Yi-1.5 delivers stronger performance in coding, math, reasoning, and instruction-following capability, while still maintaining excellent capabilities in language understanding, commonsense reasoning, and reading comprehension.

Model Context Length Pre-trained Tokens
Yi-1.5 4K, 16K, 32K 3.6T

Models

Benchmarks

  • Chat models

    Yi-1.5-34B-Chat is on par with or excels beyond larger models in most benchmarks.

    image/png

    Yi-1.5-9B-Chat is the top performer among similarly sized open-source models.

    image/png

  • Base models

    Yi-1.5-34B is on par with or excels beyond larger models in some benchmarks.

    image/png

    Yi-1.5-9B is the top performer among similarly sized open-source models.

    image/png

Quick Start

For getting up and running with Yi-1.5 models quickly, see README.

Downloads last month
21
GGUF
Model size
34.4B params
Architecture
llama

2-bit

3-bit

4-bit

5-bit

6-bit

8-bit

Inference API
Unable to determine this model's library. Check the docs .