RESEARCHInvestigateNEXT 12 MONTHS
Toward Understanding the Transferability of Adversarial Suffixes in Large Language Models
arXiv cs.CL — Computation and Language
Factual evidence
What the source reports
Researchers analyze why discrete optimization-based adversarial suffixes (jailbreaks) transfer successfully across different LLMs.
Open source