OneBench
Single Canonical Prompts Underestimate LLM Safety's Surface-Form Sensitivity | OneBench: AI Insights