BEE-spoke-data
/

smol_llama-101M-GQA

Text Generation

text-generation-inference

Model card Files Files and versions

pszemraj commited on Nov 7, 2023

Commit

5948652

·

1 Parent(s): 5c6d1f0

Update README.md

Files changed (1) hide show

README.md +4 -1

README.md CHANGED Viewed

@@ -73,7 +73,10 @@ A small 101M param (total) decoder model. This is the first version of the model
 **This checkpoint** is the 'raw' pre-trained model and has not been tuned to a more specific task. **It should be fine-tuned** before use in most cases.
-- smaller 81M parameter checkpoint with in/out embeddings tied: [here](https://huggingface.co/BEE-spoke-data/smol_llama-81M-tied)
 - For the chat version of this model, please [see here](https://youtu.be/dQw4w9WgXcQ?si=3ePIqrY1dw94KMu4)
 ---

 **This checkpoint** is the 'raw' pre-trained model and has not been tuned to a more specific task. **It should be fine-tuned** before use in most cases.
+### Checkpoints  & Links
+- _smol_-er 81M parameter checkpoint with in/out embeddings tied: [here](https://huggingface.co/BEE-spoke-data/smol_llama-81M-tied)
+- Fine-tuned on `pypi` to generate Python code - [link](https://huggingface.co/BEE-spoke-data/smol_llama-101M-GQA-python)
 - For the chat version of this model, please [see here](https://youtu.be/dQw4w9WgXcQ?si=3ePIqrY1dw94KMu4)
 ---