Llama-3.2-1B-dal-overtrained

Over fine-tune of quantised LLaMa 3.2 1B on this dataset (12 epochs).

As with other overtrained models, temperature above 0.5-0.7 and high values of sampling parameters will result in the model producing garbage or the dataset source material regardless of current context.

Framework versions

  • PEFT 0.13.0
  • Trained with unsloth on Colab.
Downloads last month
22
GGUF
Model size
1B params
Architecture
llama
Hardware compatibility
Log In to add your hardware

4-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support