papersSEP 10 04:00 UTC
Black-Box Red Teaming of Agentic AI: A Taxonomy-Driven Framework for Automated Risk Discovery
A new paper presents an automated black-box red teaming framework designed to uncover security risks in agentic AI systems. It uses a structured risk taxonomy to guide systematic testing, addressing the shortfalls of standard single-turn evaluations. The work targets agentic setups where models process untrusted inputs, invoke tools with real permissions, and act autonomously.