Stealing Reasoning Traces from Proprietary LLM APIs
This paper reveals a vulnerability in how large language model providers store and return their models' step-by-step reasoning, allowing attackers to extract sensitive information and potentially inject malicious code. Practitioners should care because this vulnerability can be exploited to steal proprietary models' reasoning and private data, as well as inject malicious payloads into public models.