Researchers Show Hidden Chain-of-Thought Can Be Extracted From Encrypted API Reasoning Blocks
A red-team paper shows encrypted reasoning traces from frontier LLM APIs can be replayed into a weaker model to recover the original hidden chain-of-thought in plaintext.