PulseAugur
EN
LIVE 12:27:33

New Spanish-language cybersecurity vision-language model released

Researchers have developed VectraYX-Vision-1B, a compact vision-language model designed for Spanish and Latin American cybersecurity imagery. This model, under 2 billion parameters, integrates a SigLIP-SO400M encoder with a specialized decoder, enabling it to process cybersecurity UIs and respond in Spanish. It features structured visual reasoning through native tokens and tool invocation capabilities, with potential for air-gapped deployment via llama.cpp's LLaVA format. Initial testing revealed challenges with visual grounding, but the team is addressing these issues and has released code, configurations, and model weights to facilitate further research on its architectural components. AI

IMPACT This model could enable more accessible cybersecurity analysis in Spanish and Latin American regions, particularly in air-gapped environments.

RANK_REASON The item describes a new model release and associated research paper detailing its architecture and preliminary findings. [lever_c_demoted from research: ic=1 ai=1.0]

Read on Hugging Face Daily Papers →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

New Spanish-language cybersecurity vision-language model released

COVERAGE [1]

  1. Hugging Face Daily Papers TIER_1 English(EN) ·

    VectraYX-Vision-1B: A Sub-2B Spanish/LATAM Cybersecurity Vision-Language Model with Structured Visual Reasoning and Native Tool Use

    We present VectraYX-Vision-1B, a sub-2B vision-language model (VLM) for Spanish/LATAM cybersecurity imagery, coupling a frozen SigLIP-so400m encoder to a 1.04B Spanish/LATAM security decoder via an MLP. To our knowledge, it is the first sub-2B VLM specialized for cyber UI (IDA, G…