At Pwn2Own Ireland 2026, researchers successfully exploited a novel argument-injection vulnerability in OpenAI Codex, a cloud-based AI coding agent. This exploit bypassed traditional prompt-injection defenses by targeting the arguments passed to the agent's tools rather than the model's reasoning itself. The vulnerability highlights a critical gap in current AI security, as many tools focus on inbound prompt filtering and fail to inspect the data exchanged between an AI agent and its execution environment. AI
IMPACT Highlights a new class of AI security vulnerabilities targeting agentic tool execution, potentially requiring new defense mechanisms beyond prompt filtering.
RANK_REASON Security tool Sentinel's blog post details a vulnerability found in OpenAI Codex at Pwn2Own, focusing on the tool's security capabilities.
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →