AnchorSR-InternVL3.5-8B-SFT

Full-parameter supervised fine-tuning of InternVL3.5-8B-HF on AnchorSR v3.

Training provenance

  • Run: internvl3_5_8b_sft_20260908_120038_665121
  • Training examples: 32,000; validation examples: 8,000.
  • Completed epochs: 2.
  • Optimizer updates: 2000 / 2000.
  • Actual global batch size for this completed run: 32.
  • Precision / distributed optimizer: BF16 / DeepSpeed ZeRO-2.
  • Video sampling: uniform 16 frames.
  • Numeric-answer loss share: 0.2.
  • Final validation loss: 0.14299454.

The validation loss is a training metric, not a task benchmark score; losses across different model tokenizers are not directly comparable.

Contents

Final model weights, tokenizer, processor and generation configuration are included. Intermediate checkpoints and optimizer states are not included. This export supports model loading and inference; it is not an exact optimizer resume checkpoint. Preserve the supplied chat template and multimodal processor.

This model is intended for research on spatial reasoning. Downstream benchmark accuracy and general deployment behavior have not been established by these training-completion checks.

Downloads last month
-
Safetensors
Model size
9B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Dataset used to train AnchorSR/AnchorSR-InternVL3.5-8B-SFT