RESEARCHInvestigateNEXT 12 MONTHS
ToolHazard: Scaling Adversarial Environments for Security Evaluation and Alignment of LLM-based Agents
arXiv cs.CL — Computation and Language
Factual evidence
What the source reports
ToolHazard creates scalable adversarial environments to evaluate LLM agent vulnerability to indirect prompt injections.
Open source