knatware/SmolLM2-135M

SmolLM2-135M

Model Description

This model was produced with the Parameterised Efficiency (PEFT / LoRA) method (PASTA framework) starting from the base model HuggingFaceTB/SmolLM2-135M, fine-tuned on a sample of the tatsu-lab/alpaca dataset.

  • Task type: instruction
  • Library: peft
  • Generated by: the PASTA fine-tuning Colab notebook generator

Intended Uses & Limitations

This model was fine-tuned on a small sample for demonstration purposes. It has not been evaluated at scale and should not be used in production or safety-critical settings without further training, evaluation, and review. Behaviour is inherited from the base model and the (small) fine-tuning sample, and may reflect biases present in either.

Training Procedure

Hyperparameters

Hyperparameter Value
Method Parameterised Efficiency (PEFT / LoRA)
Base model HuggingFaceTB/SmolLM2-135M
Dataset tatsu-lab/alpaca (train[:200])
Epochs 1
Batch size 4
Learning rate 0.0005
Max steps 20
LoRA rank 4
LoRA alpha 8

Framework versions

See the !pip install cell in the training notebook for the exact package set used.

How to Get Started

from transformers import pipeline

gen = pipeline("text-generation", model="knatware/SmolLM2-135M")
gen("Your prompt here")

Testing

Before being pushed, this model was tested locally with a sample inference call, and was re-loaded and tested again directly from the Hub after pushing to confirm the upload was complete and usable.


© Knatware Technology UK. Developed by Kayode Okosi.

Downloads last month
18
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for knatware/SmolLM2-135M

Adapter
(24)
this model

Dataset used to train knatware/SmolLM2-135M