2 min read · Mastodon
Encrypted Reasoning Traces in Proprietary LLM APIs Expose Credentials
Researchers discovered that encrypted reasoning blocks returned by proprietary LLM APIs can be replayed in jailbroken weaker models to extract raw thinking traces verbatim. Analysis of public agent trajectories revealed hundreds of exposed API keys and credentials.
Why it matters. Audit your agent trajectories and sanitize prompts immediately to prevent sensitive credentials from leaking through encrypted reasoning payloads.