Researchers decode "encrypted" chain-of-thought reasoning from Anthropic, OpenAI, and Google models
•12 sources
Questions this post answers
What is cross-model replay and how does it extract chain-of-thought reasoning from AI models?
Cross-model replay is a technique that pulls reasoning traces out of one AI model and reuses them across other models. It does not break any underlying cryptographic primitive — it sidesteps the question entirely. Applied to frontier models from Anthropic, OpenAI, and Google, it demonstrates that the supposed encryption protecting chain-of-thought reasoning traces was largely ineffective, meaning distillation may have been possible without any cryptographic attack. Developers building on top of frontier LLMs track research like this on daily.dev to stay ahead of security implications.
198 Impressions