AgentDyn
PulseAugur coverage of AgentDyn — every cluster mentioning AgentDyn across labs, papers, and developer communities, ranked by signal.
-
StepGuard system enhances AI agent safety with step-level control
Researchers have developed StepGuard, a novel system designed to monitor and control the actions of AI agents at a step-by-step level, rather than just evaluating completed trajectories. This approach aims to prevent se…
-
New frameworks LongGuard and StepGuard enhance LLM safety guardrails
Researchers have developed two new frameworks, LongGuard and StepGuard, to address safety failures in large language models (LLMs). LongGuard focuses on analyzing and mitigating failures in long-context guardrails, prop…
-
New framework studies backdoor decontamination in LLM agents
Researchers have developed a framework to study how LLM agents can be decontaminated from hidden backdoors installed during fine-tuning. Their experiments show that introducing a known backdoor and then unlearning it ca…