Vision-OPD Qwen3.5-4B RandomDrop5 Step30

Qwen3.5-4B trained with Vision-OPD using random 5% visual-token retention for the student and full visual tokens for the teacher. This is the 30-step proof-of-concept checkpoint.

The checkpoint can run with full visual tokens using standard Transformers. For 5% visual-token inference, use the pruning-aware serving code in prune-opd.

Training data: Vision-OPD-6K.

Downloads last month
12
Safetensors
Model size
5B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for zhuqiang/Vision-OPD-Qwen3.5-4B-RandomDrop5-Step30

Finetuned
Qwen/Qwen3.5-4B
Finetuned
(496)
this model