PulseAugur
EN
LIVE 03:55:00
Deutsch(DE) Mistral veröffentlicht Shieldstral, ein 3B Open-Weight Safety-Modell, das feste Kategorien durch natürliche Sprache ersetzt. Es schlägt rechenintensivere Klassi

Mistral AI releases Shieldstral 1.0 3B, an adaptable open-weights safety classifier

Mistral AI has launched Shieldstral 1.0 3B, an open-weights safety classifier designed for policy adaptability. Unlike traditional models that rely on fixed harm categories, Shieldstral uses natural language questions to perform content moderation, allowing operators to define custom policies at inference time. This approach enables the model to achieve strong performance on both text and multimodal safety benchmarks, matching larger models while running efficiently on a single GPU with 16GB of VRAM. AI

IMPACT Enables flexible, on-premise content moderation for diverse applications, potentially reducing reliance on third-party vendors.

RANK_REASON Frontier-lab model release with system card.

Read on Mastodon — mastodon.social →

AI-generated summary · Google Gemini · from 5 sources. How we write summaries →

Mistral AI releases Shieldstral 1.0 3B, an adaptable open-weights safety classifier

COVERAGE [5]

  1. MarkTechPost TIER_1 English(EN) · Asif Razzaq ·

    Mistral AI Releases Shieldstral 1.0 3B: An Open-Weights Policy-Adaptive Multimodal Safety Classifier Matching Models 7× Its Size

    <p>Mistral AI has released Shieldstral 1.0 3B, an open-weights, policy-adaptive multimodal safety classifier that frames content moderation as a single yes/no question instead of a fixed harm taxonomy. Operators supply the policy as a plain-language query at inference time and ge…

  2. Mastodon — fosstodon.org TIER_1 English(EN) · [email protected] ·

    Mistral AI has released Shieldstral 1.0 3B, an open-weights safety classifier that adapts to different policies at runtime. Rather than using a fixed harm taxon

    Mistral AI has released Shieldstral 1.0 3B, an open-weights safety classifier that adapts to different policies at runtime. Rather than using a fixed harm taxonomy, it treats content moderation as a yes/no question. Achieves 84.9% F1 on text safety benchmarks while running on a s…

  3. Mastodon — mastodon.social TIER_1 Italiano(IT) · [email protected] ·

    🧠 # Mistral # AI has released # Shieldstral 1.0 3B, an open-weight model designed for content moderation and safety. 👉 Details: https://ww

    🧠 # Mistral # AI ha rilasciato # Shieldstral 1.0 3B, un modello open-weight progettato per la moderazione e la sicurezza dei contenuti. 👉 I dettagli: https://www. linkedin.com/posts/alessiopoma ro_mistral-ai-shieldstral-share-7492449329505337344-xahm/ ___ ✉️ 𝗦𝗲 𝘃𝘂𝗼𝗶 𝗿𝗶𝗺𝗮𝗻𝗲𝗿𝗲 𝗮𝗴𝗴𝗶…

  4. Mastodon — mastodon.social TIER_1 English(EN) · sipirtu ·

    Mistral's Shieldstral is a 3B open model that performs safety checks via natural language questions, matching larger models in benchmarks. Operators can define

    Mistral's Shieldstral is a 3B open model that performs safety checks via natural language questions, matching larger models in benchmarks. Operators can define their own criteria and run it locally. Source: The Decoder AI https:// the-decoder.com/mistrals-open- model-shieldstral-…

  5. Mastodon — mastodon.social TIER_1 Deutsch(DE) · aisyndicate ·

    Mistral releases Shieldstral, a 3B open-weight safety model that replaces hard categories with natural language. It outperforms compute-intensive classi

    Mistral veröffentlicht Shieldstral, ein 3B Open-Weight Safety-Modell, das feste Kategorien durch natürliche Sprache ersetzt. Es schlägt rechenintensivere Klassifikatoren bei gleicher Benchmark-Leistung – relevant für lokale Inferenz. https:// the-decoder.de/mistrals-shield stral-…