Vision-OPD Qwen2.5-VL-7B RandomDrop5 1 Epoch

Qwen2.5-VL-7B-Instruct trained for one epoch with Vision-OPD using random 5% visual-token retention for the student and full visual tokens for the teacher. This is the final step-64 checkpoint.

The checkpoint can run with full visual tokens using standard Transformers. For 5% visual-token inference, use the pruning-aware serving code in prune-opd.

Training data: Vision-OPD-6K.

Downloads last month
17
Safetensors
Model size
8B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for zhuqiang/Vision-OPD-Qwen2.5-VL-7B-Instruct-RandomDrop5-1Epoch

Finetuned
(1176)
this model