wwe180's picture
Update README.md
5b6fbce verified
metadata
base_model:
  - wwe180/Llama3-18B-lingyang-v1
library_name: transformers
tags:
  - mergekit
  - merge
  - Llama3
license:
  - other

After simple testing, the effect is good, stronger than llama-3-8b!

merge

This is a merge of pre-trained language models created using mergekit.

💻 Usage

!pip install -qU transformers accelerate

from transformers import AutoTokenizer
import transformers
import torch

model = "Llama3-18B-lingyang-v1"
messages = [{"role": "user", "content": "What is a large language model?"}]

tokenizer = AutoTokenizer.from_pretrained(model)
prompt = tokenizer.apply_chat_template(messages, tokenize=False, add_generation_prompt=True)
pipeline = transformers.pipeline(
    "text-generation",
    model=model,
    torch_dtype=torch.float16,
    device_map="auto",
)

Statement:

Llama3-18B-lingyang-v1 does not represent the views and positions of the model developers We will not be liable for any problems arising from the use of the Llama3-18B-lingyang-v1 open Source model, including but not limited to data security issues, risk of public opinion, or any risks and problems arising from the misdirection, misuse, dissemination or misuse of the model.