RESEARCHInvestigateNOW
Sharding Prevents LLM Oversight Failures and Adversarial Exploitation
arXiv cs.LG — Machine Learning
Factual evidence
What the source reports
Research demonstrates that LLM judge reliability falls as the number of verdicts per call increases, recommending a 'sharding' architecture.
Open source