Black-Box Red Teaming of Agentic AI: A Taxonomy-Driven Framework for Automated Risk Discovery
arXiv cs.AIen
arXiv cs.AI
AI Global WirearXiv:2609.09647v1 Announce Type: new Abstract: Agentic systems are rapidly moving to production, where they read untrusted inputs, call tools with real permissions, and act autonomously, expanding the security surface beyond chat-only models. Yet standard evaluations remain single-turn and fail to capture multi-step agent vulnerabilities. We present a systematic black-box framework for risk-aware agent evaluation requiring only basic system descriptions. Our approach introduces: (1) a seven-domain taxonomy mapping observable behaviors to risk categories, (2) fully automated SAGE-RT red teaming producing 120 adversarial scenarios per domain, and (3) human-validated evaluation using LLM judge
This is a short summary published by AI Global Wire. The full article is owned and hosted by arXiv cs.AI — open it there to read it in full.
Read the full story at arXiv cs.AI- Verktyg
- Forskning
- Agenter
- Företag
Related AI news
- Meta share price surges after personal AI agent Muse releaseEconomic Times Tech · September 10, 2026
- Exclusive: Bynario raises €2.1m to tackle cybersecurity’s AI slop problemSifted · September 10, 2026
- Donnerstag: Apples neue iPhones auch aufklappbar, KI-Agenten weiter ungezügeltheise online – KI · September 10, 2026
- Generative AI a new tool in Mali's information war: studyEconomic Times Tech · September 10, 2026
- Sources: Alibaba is set to lead a $300M round in AI model testing startup UniPat AI at a $2.5B valuation; UniPat founder Li Kuan worked at Alibaba's Tongyi Lab (Bloomberg)Techmeme · September 10, 2026
- OpenDiscoveryTrace: Process Traces for Evaluating AI Scientist WorkflowsarXiv cs.AI · September 10, 2026