FosCap Qwen3.6-35B-A3B Stage 2

Stage 2 full fine-tune for the FosCap three-stage fossil-image understanding and captioning pipeline. It continues from Stage 1 and was trained on foscap_stage2.

Training summary

  • Learning rate: 1e-5
  • Epochs: 1
  • Seed: 42
  • Effective training batch size: 64
  • Framework: LLaMA-Factory / Transformers

The final checkpoint is FosCap Stage 3.

Downloads last month
1
Safetensors
Model size
665k params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for LockOnN/FosCap-Qwen3.6-35B-A3B-Stage2

Finetuned
(1)
this model
Finetunes
1 model

Collection including LockOnN/FosCap-Qwen3.6-35B-A3B-Stage2