Wired · Will Knight ·

Researchers find that feeding a frontier model's encrypted reasoning traces to a weaker model from the same provider can make it output the traces in plaintext

Researchers devised a way to extract "reasoning traces" from Claude, GPT, and Gemini. What they found, they say …

Researchers find that feeding a frontier model's encrypted reasoning traces to a weaker model from the same provider can make it output the traces in plaintext

Lead Source

How this story grew

Coverage · 0 Discussion · 0
Aug 11Aug 13

More

arXiv.org: arXiv.org
Simon Willison's Weblog: Simon Willison's Weblog
The Hacker News: The Hacker News
WinBuzzer: WinBuzzer
Digit: Digit
The American Bazaar: The American Bazaar
Latent.Space: Latent.Space
The Neuron: The Neuron
Wccftech: Wccftech

Discussion

Related stories