PulseAugur
EN
LIVE 07:11:57
ENTITY AgentHarm

AgentHarm

PulseAugur coverage of AgentHarm — every cluster mentioning AgentHarm across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
1
3 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
1
3 over 90d
TIER MIX · 90D
TOPICS
SENTIMENT · 30D

1 day(s) with sentiment data

RECENT · PAGE 1/1 · 3 TOTAL
  1. TOOL · CL_169815 ·

    New prompt injection detection techniques leverage cross-domain methods

    Researchers have developed seven novel techniques for detecting prompt injection attacks, moving beyond traditional pattern matching and fine-tuned transformer classifiers. These new methods draw inspiration from divers…

  2. TOOL · CL_74391 ·

    New framework guides LLM agents to revise plans, improving safety

    Researchers have developed TRIAD, a new framework for LLM agents that integrates guardrails to improve safety and utility. Unlike traditional guardrails that simply block unsafe actions, TRIAD provides feedback to guide…

  3. TOOL · CL_32708 ·

    New framework LiSA enhances AI guardrails with sparse failure data

    Researchers have developed LiSA (Lifelong Safety Adaptation), a new framework designed to improve AI guardrails by learning from sparse and noisy failure data. LiSA uses structured memory to generalize from individual i…