PulseAugur
EN
LIVE 10:10:17
ENTITY AgentDojo

AgentDojo

PulseAugur coverage of AgentDojo — every cluster mentioning AgentDojo across labs, papers, and developer communities, ranked by signal.

Show in brief
Total · 30d
6
14 over 90d
Releases · 30d
0
0 over 90d
Papers · 30d
5
12 over 90d
TIER MIX · 90D
TOPICS
RELATIONSHIPS
SENTIMENT · 30D

4 day(s) with sentiment data

RECENT · PAGE 1/1 · 14 TOTAL
  1. TOOL · CL_183116 ·

    New defense probes detect and mitigate indirect prompt injection in LLMs

    Researchers have developed a method to detect indirect prompt injection (IPI) attacks in agentic large language models (LLMs). By training simple linear probes on the models' internal states, they can predict IPI exposu…

  2. TOOL · CL_185542 ·

    New PIMiner system automates LLM prompt injection red-teaming

    Researchers have developed PIMiner, an agentic system designed for automated prompt injection red-teaming of large language models. Unlike existing methods that often struggle with generalization, PIMiner builds a trans…

  3. TOOL · CL_169815 ·

    New prompt injection detection techniques leverage cross-domain methods

    Researchers have developed seven novel techniques for detecting prompt injection attacks, moving beyond traditional pattern matching and fine-tuned transformer classifiers. These new methods draw inspiration from divers…

  4. RESEARCH · CL_167426 ·

    New method enhances LLM agent safety by stratifying risk in tool calls

    Researchers have developed a new method called role-stratified per-field conformal risk control to enhance the safety of language-model agents. This technique calibrates risk budgets separately for different semantic ro…

  5. TOOL · CL_160935 ·

    New RL framework PISmith tests and breaks prompt injection defenses

    Researchers have developed PISmith, a novel reinforcement learning (RL) framework designed to rigorously test the effectiveness of prompt injection defenses in large language models (LLMs). The framework trains an attac…

  6. RESEARCH · CL_158657 ·

    New methods compress LLM agent context for improved security and efficiency

    Researchers have developed new methods for compressing context in large language model (LLM) agents to improve efficiency and security. One approach, "Twin Agent," separates agents into an "Explore Agent" for untrusted …

  7. TOOL · CL_156360 ·

    New pipeline hardens agentic AI apps against data leaks

    Researchers have developed a new pre-deployment pipeline designed to prevent data leakage and tool misuse in agentic applications. This pipeline scans, hardens, and validates agentic systems by analyzing prompt template…

  8. TOOL · CL_116442 ·

    Prompt optimization may weaken LLM adversarial robustness, new benchmark suggests

    A new benchmark has been developed to investigate whether prompt optimization techniques for Large Language Models (LLMs) weaken their robustness against adversarial attacks, specifically prompt injection. Initial findi…

  9. TOOL · CL_70446 ·

    LLM attack benchmarks cover less than 25% of threat landscape

    Researchers have developed a new framework to audit the coverage of benchmarks designed to test Large Language Model (LLM) attacks. This framework, based on a taxonomy of over 500 inference-time attacks, reveals that cu…

  10. TOOL · CL_53868 ·

    New Protocol Enables LLMs to Safely Control Small Devices

    Researchers have introduced the Device Context Protocol (DCP), a new architecture designed to enable large language models (LLMs) to safely control constrained devices. DCP is significantly more lightweight than existin…

  11. TOOL · CL_50472 ·

    Arc Gate offers solution to OpenAI's 'unfixable' prompt injection vulnerability

    OpenAI has stated that prompt injection in browser agents is an unfixable structural vulnerability at the model level. However, a new architectural solution called Arc Gate has demonstrated significant success in mitiga…

  12. TOOL · CL_32688 ·

    LLM attack benchmarks show significant gaps in security coverage

    Researchers have developed a new framework to audit the coverage of LLM attack benchmarks, revealing significant gaps in current evaluations. Their analysis of six public benchmarks showed they collectively cover less t…

  13. RESEARCH · CL_16489 ·

    New attack exploits LLM agent relays, bypassing alignment defenses

    Researchers have identified a new vulnerability in LLM agent architectures that use Bring-Your-Own-Key (BYOK) systems. These architectures route LLM traffic through third-party relays, creating an integrity gap where a …

  14. RESEARCH · CL_99526 ·

    New benchmarks and safety methods emerge for advanced LLM agents

    New research explores the development and evaluation of AI agents, focusing on their ability to navigate complex environments and adhere to policies. StarDojo benchmarks agent performance in open-ended simulations like …