Text Generation
Transformers
Safetensors
English
llama
smol_llama
llama2
Inference Endpoints
text-generation-inference
pszemraj commited on
Commit
f49a072
1 Parent(s): be538b7

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +6 -1
README.md CHANGED
@@ -65,4 +65,9 @@ A small 101M param (total) decoder model. This is the first version of the model
65
 
66
  - 768 hidden size, 6 layers
67
  - GQA (24 heads, 8 key-value), context length 1024
68
- - train-from-scratch
 
 
 
 
 
 
65
 
66
  - 768 hidden size, 6 layers
67
  - GQA (24 heads, 8 key-value), context length 1024
68
+ - train-from-scratch
69
+
70
+ For the chat version of this model, please [see here](https://youtu.be/dQw4w9WgXcQ?si=3ePIqrY1dw94KMu4)
71
+
72
+ ---
73
+