Stolen LLM reasoning: researchers show vulnerable models can reuse encrypted reasoning across OpenAI, Anthrophic, and Google; weaker models can reproduce the reasoning verbatim
Read the original at old.reddit.com→If you haven't checked the paper: https://arxiv.org/abs/2608.09867 TLDR: the authors show that you can swap out the "encrypted" reasoning of the biggest model, like Opus, Sol, and put them into weaker model with less...
Original headline: "Stolen LLM Reasoning: How come OpenAI, Anthrophic, Google have the same vulnerabilities?"
Coverage timeline
- Aug 13, 12:13 UTC r/LocalLLaMA lead source Stolen LLM Reasoning: How come OpenAI, Anthrophic, Google have the same vulnerabilities?