RESEARCHInvestigateNEXT 12 MONTHS
The Knowing-Saying Gap: When Probes See Errors that Confidence Misses
arXiv cs.LG — Machine Learning
Factual evidence
What the source reports
Research shows linear probes detecting LLM context corruption fail to accurately predict final answer correctness in multi-step reasoning.
Open source