fdfdfddsfdfdfd/llama-nano-tiny-sft

Small language model trained with a two-phase pipeline: streaming pre-training and SFT.

  • Final Training Stage: N/A
  • Final Loss: 0.0000
  • Pre-training Steps: 1000
  • SFT Epochs: 1
  • Framework: PyTorch
  • Architecture: Llama
Downloads last month
9
Safetensors
Model size
107k params
Tensor type
F32
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support