OneBench
MARGIN: Runtime Confidence Calibration for Multi-Agent Foundation Model Coordination | OneBench: AI Insights