PulseAugur
实时 21:28:00
English(EN) Shieldstral: Why a 3B Guard Model Ties a 20B One

Mistral AI发布Shieldstral,一个灵活的3B Guard模型

Mistral AI发布了Shieldstral,一个拥有30亿参数的Guard模型,用于分类内容策略违规。与LlamaGuard和ShieldGemma等先前模型不同,Shieldstral的策略嵌入在提示中而非模型权重中,允许在不重新训练的情况下进行动态调整。这种方法使Shieldstral能够作为二元分类器运行,根据“是”或“否”标记的logits输出0到1之间的分数,与固定标签相比,在设置审核阈值方面提供了更大的灵活性。 AI

影响 通过允许在不重新训练的情况下动态调整策略,为内容审核提供了一种更灵活且具成本效益的方法。

排序理由 前沿实验室(Mistral AI)发布新模型。[lever_c_demoted from frontier_release: ic=1 ai=1.0]

在 dev.to — LLM tag 阅读 →

AI 生成摘要 · Google Gemini · 来自 1 个来源。 我们如何撰写摘要 →

Mistral AI发布Shieldstral,一个灵活的3B Guard模型

报道来源 [1]

  1. dev.to — LLM tag TIER_1 English(EN) · AI Coding Patterns ·

    Shieldstral: Why a 3B Guard Model Ties a 20B One

    <p>Some teams are paying for inference on a 20B model that reasons out loud for several hundred tokens just to decide whether a comment breaks their content policy. On August 4th Mistral released Shieldstral, a 3B classifier that answers the same question with a single token <sup…