RESEARCHInvestigateNEXT 12 MONTHS
TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectories
arXiv cs.LG — Machine Learning
Factual evidence
What the source reports
TraceSafe-Bench introduces a framework to benchmark LLM safety guardrails across multi-step, intermediate tool-calling execution traces.
Open source