SigLIP-SO400M β Renesas X5H
SigLIP vs. CLIP evaluation results, from the original paper β as shown on the official google/siglip-so400m-patch14-384 model card
π§ Coming soon. This model is being optimized and validated for Renesas hardware β no fixed release date yet.
Introduction
SigLIP-SO400M is Google's Sigmoid-Loss vision-language embedding model (the "shape-optimized" 400M-parameter variant), commonly used as a vision encoder for zero-shot image classification and retrieval. Renesas is preparing an optimized deployment of this model for the R-Car Gen5 platform.
- Model Architecture: SigLIP vision/text dual encoder (vision tower only is the focus for on-device deployment).
- Source Model: google/siglip-so400m-patch14-384
This page is a placeholder. Content is provisional and subject to change before the model is fully published.
Model tree for Renesas/SigLIP-SO400M
Base model
google/siglip-so400m-patch14-384