fdfdfddsfdfdfd/llama-nano-tiny-cpu-fast-sft

Small language model trained with a two-phase pipeline: streaming pre-training and SFT.

  • Final Training Stage: Pre-training
  • Final Loss: 0.0000
  • Pre-training Steps: 1000
  • SFT Steps: 200
  • Framework: PyTorch
  • Architecture: Llama
Downloads last month
17
Safetensors
Model size
6.47M params
Tensor type
F32
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for fdfdfddsfdfdfd/llama-nano-tiny-cpu-fast-sft

Quantizations
2 models