VectraYX-Vision-1B: A Sub-2B Spanish/LATAM Cybersecurity Vision-Language Model with Structured Visual Reasoning and Native Tool Use
Paper • 2608.08477 • Published • 7
This repository contains the model described in the paper VectraYX-Vision-1B: A Sub-2B Spanish/LATAM Cybersecurity Vision-Language Model with Structured Visual Reasoning and Native Tool Use.
VectraYX-Vision-1B is a sub-2B vision-language model specialized for Spanish/LATAM cybersecurity imagery. It couples a frozen SigLIP-so400m encoder to a 1.04B Spanish/LATAM security decoder via an MLP, and is designed to answer in Spanish, emit structured reasoning via native <|think|> tokens, and invoke tools via the Model Context Protocol (<|tool_call|>).