PulseAugur
EN
LIVE 18:02:00

Mistral AI releases small, customizable safety model Shieldstral

Mistral AI has released Shieldstral, a new 3-billion parameter safety model that can evaluate AI inputs and outputs for violations. This model operates by answering natural language yes-or-no questions, allowing users to define their own safety criteria at runtime. Despite its small size, Shieldstral demonstrates performance comparable to models seven times larger on certain benchmarks and can be run locally. AI

IMPACT Offers a more efficient and customizable approach to AI safety checks, potentially enabling broader local deployment.

RANK_REASON Release of a new, smaller model focused on safety capabilities. [lever_c_demoted from research: ic=1 ai=1.0]

Read on The Decoder →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

Mistral AI releases small, customizable safety model Shieldstral

COVERAGE [1]

  1. The Decoder TIER_1 English(EN) · Jonathan Kemper ·

    Mistral's open model Shieldstral matches much larger safety models at a fraction of the size

    <p><img alt="" class="attachment-full size-full wp-post-image" height="1152" src="https://the-decoder.com/wp-content/uploads/2026/07/mistral_ai-3.png" style="height: auto; margin-bottom: 10px;" width="2048" /></p> <p> Mistral's new 3B Shieldstral model checks AI inputs and output…