RESEARCHInvestigateNEXT 12 MONTHS
From Leaky Thoughts to Private Reasoning: Controlling What LRMs Say to Themselves
arXiv cs.CL — Computation and Language
Factual evidence
What the source reports
Research demonstrates that reasoning model traces leak sensitive user data via prompt injection and tests instruction-following controls.
Open source