RESEARCHInvestigateNOW
Is Vibe Coding Safe? Benchmarking Vulnerability of Agent-Generated Code in Real-World Tasks
arXiv cs.CL — Computation and Language
Factual evidence
What the source reports
Researchers benchmarked LLM agent-generated code security across 186 real-world software engineering tasks using the SUSVIBES benchmark.
Open source