Janus-Pro-7B ScienceQA SFT checkpoint

This private checkpoint is the stage-1 ScienceQA supervised-fine-tuning result used by the Janus thesis reproduction in KFCCrazzzyThursday/Janus_pro_finetune.

  • Base model: deepseek-ai/Janus-Pro-7B
  • Training view: 5,678 ScienceQA image-only examples with official solutions
  • Training: full-parameter BF16, 267 optimizer steps, four GPUs
  • Intended next stage: TQA GRPO

The thesis did not report the stage-1 SFT hyperparameters, so this checkpoint uses the documented reproduction assumptions. See configs/paper.yaml and reproducibility/paper_audit.md in the code repository before interpreting or redistributing the model.

This repository contains model artifacts only. It intentionally contains no dataset images, API keys, judge responses, or training logs.

Downloads last month
25
Safetensors
Model size
7B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Billyshears/Janus_pro_finetune

Finetuned
(37)
this model