PulseAugur
EN
LIVE 21:46:44

AI Reasoning Traces Vulnerable to Cross-Model Decryption Attacks

Researchers have discovered a vulnerability in the API ecosystems of Anthropic, OpenAI, and Google that allows for the extraction of supposedly hidden reasoning traces. By replaying encrypted reasoning blocks into weaker compatible models, these models act as decryption oracles, revealing the opaque traces as readable plaintext. A scan of public agent trajectories uncovered personal information and credentials within these decoded blocks, highlighting a potential security risk for sensitive data embedded in AI reasoning outputs. AI

IMPACT Highlights a potential security flaw in how AI reasoning data is handled, urging caution for sensitive information.

RANK_REASON Academic paper detailing a novel attack vector on AI reasoning outputs. [lever_c_demoted from research: ic=1 ai=1.0]

Read on dev.to — Anthropic tag →

AI-generated summary · Google Gemini · from 1 sources. How we write summaries →

AI Reasoning Traces Vulnerable to Cross-Model Decryption Attacks

COVERAGE [1]

  1. dev.to — Anthropic tag TIER_1 English(EN) · Simon Paxton ·

    Anthropic, OpenAI and Google Reasoning Locks Met Their Own Spare Keys

    <p>Researchers reported on August 10, 2026, that they could extract supposedly hidden reasoning traces from <a href="https://arxiv.org/abs/2608.09867" rel="noopener noreferrer">Anthropic, OpenAI, and Google API ecosystems</a> by replaying encrypted reasoning blocks into weaker co…