Researchers demonstrate a method to extract plaintext reasoning traces from encrypted frontier LLM outputs. The technique exploits how APIs return encrypted chain-of-thought blocks that can be cross-played into weaker sibling models to bypass protections. This vulnerability exposes proprietary reasoning data from major providers like OpenAI, Anthropic, and Google.
Opening Kapyn…