Reimagining service delivery in the agentic era with Google Public Sector
cloud.google.com·19h ago
TL;DR
Existing benchmarks for evaluating the safety of AI agents against decomposition attacks were insufficient. DECOMPBENCH, a new benchmarking tool, was developed to assess agent safety in these scenarios.
✦ Why It Matters
Engineers can use DECOMPBENCH to better evaluate and improve the safety of AI agents against decomposition attacks.
Key Takeaways
How It Works
DECOMPBENCH employs a decomposition-by-design principle, allowing harmful tasks to be broken down into smaller, benign subtasks. This is achieved through a graphical framework that simulates realistic workflows, enabling the evaluation of agent responses to both monolithic and decomposed tasks.
Related