RESEARCHInvestigateNEXT 12 MONTHS
Improving Reproducibility in Evaluation through Multi-Level Annotator Modeling
arXiv cs.LG — Machine Learning
Factual evidence
What the source reports
Researchers propose a multi-level annotator modeling framework to address the reproducibility crisis in subjective human evaluations of LLMs.
Open source