FosCap Qwen3.6-35B-A3B Stage 1

Stage 1 full fine-tune for the FosCap three-stage fossil-image understanding and captioning pipeline. It was trained on foscap_stage1.

Training summary

  • Learning rate: 1e-5
  • Epochs: 1
  • Seed: 42
  • Effective training batch size: 64
  • Framework: LLaMA-Factory / Transformers

The next-stage checkpoint is FosCap Stage 2.

Downloads last month
2
Safetensors
Model size
665k params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for LockOnN/FosCap-Qwen3.6-35B-A3B-Stage1

Finetuned
(208)
this model
Finetunes
1 model

Collection including LockOnN/FosCap-Qwen3.6-35B-A3B-Stage1