AIWired6h ago

Researchers find that feeding a frontier model's encrypted reasoning

Researchers find that feeding a frontier model's encrypted reasoning traces to a weaker model from the same provider can make it output the traces in plaintext

Researchers find that feeding a frontier model's encrypted reasoning

Researchers devised a way to extract “reasoning traces” from Claude, GPT, and Gemini. What they found, they say …

Read full article

Source: Wired · Opens in new tab