RESEARCHInvestigateNOW
StepJack: Benchmarking Computer-Use Agent Safety Against Multi-Step Indirect Prompt Injection
arXiv cs.CL — Computation and Language
Factual evidence
What the source reports
Researchers introduce StepJack, a benchmark proving computer-use agents are vulnerable to multi-step indirect prompt injection attacks.
Open source