PulseAugur
EN
LIVE 14:17:43
ENTITY ShieldGemma

ShieldGemma

PulseAugur coverage of ShieldGemma — every cluster mentioning ShieldGemma across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
2
8 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
0
4 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

2 day(s) with sentiment data

RECENT · PAGE 1/1 · 8 TOTAL
  1. SIGNIFICANT · CL_186551 ·

    Mistral AI releases Shieldstral, a flexible 3B guard model

    Mistral AI has released Shieldstral, a 3 billion parameter guard model designed to classify content policy violations. Unlike previous models like LlamaGuard and ShieldGemma, Shieldstral's policy is embedded within the …

  2. SIGNIFICANT · CL_182799 ·

    Mistral AI unveils policy-adaptive multimodal safety model Shieldstral

    Mistral AI has released Shieldstral 1.0 3B, a new multimodal safety classifier designed for efficient content moderation. Unlike traditional models that predict fixed categories, Shieldstral adapts to natural language s…

  3. RESEARCH · CL_109527 ·

    Encoder classifiers offer cost-effective LLM safety evaluation, study finds

    A new research paper explores the effectiveness of encoder classifiers, specifically from the ModernBERT family, as a cost-efficient alternative to LLM-based judges for evaluating the safety of large language model outp…

  4. RESEARCH · CL_84362 ·

    New system detects distributional shift in AI safety classifiers

    Researchers have developed a new online system designed to monitor distributional shift in deployed AI safety classifiers. This system uses sequential statistics to detect when a classifier's performance degrades due to…

  5. TOOL · CL_79753 ·

    AI safety judges trained with curriculum for improved rubric consistency

    Researchers have developed a new training strategy for AI safety judges, aiming to improve their consistency and reliability. The strategy involves using dynamic rubrics generated from prompt-response-label triples to e…

  6. TOOL · CL_38995 ·

    GLiNER Guard unifies LLM safety and PII detection in single pass

    A new system called GLiNER Guard (GLiGuard) has been developed to streamline safety moderation and PII detection for large language models. This unified encoder collapses multiple classifiers and NER models into a singl…

  7. TOOL · CL_30372 ·

    Fastino Labs open-sources GLiGuard safety model

    Fastino Labs has released GLiGuard, an open-source safety moderation model designed to be significantly faster and more efficient than existing solutions. Unlike traditional decoder-only models that generate responses t…

  8. RESEARCH · CL_01364 ·

    Google releases Gemma 2 2B, ShieldGemma, and Gemma Scope

    Google has announced updates to its Gemma family of models, including the release of Gemma 2 2B. This new iteration is designed for efficiency and accessibility, aiming to empower developers with powerful yet lightweigh…