RESEARCHInvestigateNEXT 12 MONTHS
LOCKS: Page-Local Compact Key Summaries for Efficient Long-Context Decoding
arXiv cs.LG — Machine Learning
Factual evidence
What the source reports
Researchers propose LOCKS, a KV cache compression method using page-local spectral summaries to reduce long-context inference overhead.
Open source