gemma-7b-zephyr-sft / README.md
tcapelle's picture
Update README.md
6ac82af verified
|
raw
history blame
1 kB
metadata
library_name: transformers
datasets:
  - HuggingFaceH4/ultrachat_200k
base_model: google/gemma-7b

Visualize in Weights & Biases

Gemma 7B Zephyr SFT

The Zephyr SFT recipe applied on top of Gemma 7B

Model description

  • Model type: A 8.5B parameter GPT-like model fine-tuned on a mix of publicly available, synthetic datasets.
  • Language(s) (NLP): Primarily English
  • Finetuned from model: google/gemma-7b

Recipe

We trained using the alignment handbook recipe and logging to W&B

Visit the W&B workspace here

Compute provided by Lambda Labs - 8xA100 80GB node