Compressing visual tokens in vision-language models: 3x more requests per GPU on Qwen2-VL
• 1
None defined yet.
You can modify this app directly by editing index.html in the Files and versions tab.
Also don't forget to check the Spaces documentation.