FosCap Qwen3.6-35B-A3B Stage 3

Final Stage 3 full fine-tune for the FosCap three-stage fossil-image understanding and captioning pipeline. It continues from Stage 2 and was trained on foscap_stage3_new_visual_expert.

Training summary

  • Learning rate: 1e-5
  • Epochs: 3
  • Seed: 668
  • Effective training batch size: 64
  • Framework: LLaMA-Factory / Transformers

Only the final checkpoint is published in this repository; the redundant intermediate checkpoint-45 training snapshot is intentionally omitted.

Downloads last month
-
Safetensors
Model size
665k params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for LockOnN/FosCap-Qwen3.6-35B-A3B-Stage3

Collection including LockOnN/FosCap-Qwen3.6-35B-A3B-Stage3