OneBench
ROGUE: Evaluating Corrigibility Failures in Frontier Computer-Use Agents | OneBench: AI Insights