RESEARCHInvestigateNEXT 12 MONTHS
When Context Returns: Toward Robust Internalization in On-Policy Distillation
arXiv cs.LG — Machine Learning
Factual evidence
What the source reports
Researchers identify a vulnerability where reintroducing privileged context to distilled student models degrades inference performance.
Open source