Beyond Predictable Paths: Redefining AI Security Incident Reporting for Agents

📅 2026-09-21
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
本文针对AI代理安全事件报告问题,通过专家意见识别必要信息,并提出高效记录事件和评估漏洞泛化等研究方向。
📝 Abstract
AI agents are being deployed rapidly, accompanied by a growing number of AI-specific attacks and corresponding incidents. As incident reporting becomes increasingly important for legal compliance, governance, accountability, and security; current frameworks must be adapted to the unique characteristics of AI agents. In this paper, two editorial authors compare AI systems and AI agents and, drawing on input from 23 experts in academia and industry, identify the information required for reporting incidents where the security of AI agents is harmed. %involving AI agents. Potential reporting elements include, for example, agent memory and memory accesses, actual and potential levels of autonomy, and tool usage. Based on these findings, we identify several open research questions, including how to efficiently record incidents and how to determine whether vulnerabilities and incidents generalize. Expert feedback also highlighted potential reporting weaknesses, such as risks of data leakage and attacks targeting the reporting infrastructure itself, creating additional research needs. Lastly, we summarize privacy requirements and outline research directions for the secure and trustworthy deployment of AI agents.
Problem

Research questions and friction points this paper is trying to address.

AI agents
incident reporting
security
vulnerabilities
data leakage
Innovation

Methods, ideas, or system contributions that make the work stand out.

AI Security Incident Reporting
Autonomy Levels
Agent Memory
🔎 Similar Papers
No similar papers found.
A
Anastasia Pustozerova
SBA Research, Austria
Eugene Bagdasarian
Eugene Bagdasarian
UMass Amherst, Google
ML SecurityML PrivacyContextual IntegrityTrustworthy AIAI Safety
Luca Beurer-Kellner
Luca Beurer-Kellner
ETH Zürich
Battista Biggio
Battista Biggio
Professor of Computer Engineering, University of Cagliari, Italy
Adversarial Machine LearningAI SecurityMachine LearningComputer Security
N
Nico Ebert
ZHAW, Switzerland
David Filip
David Filip
ISO/IEC JTC 1/SC 42, Huawei, Ireland
M
Marc Fischer
Snyk, Switzerland
H
Heather Frase
Veraitech and Virginia Tech, US
D
David Hofer
Snyk, Switzerland
J
Juliane Hoffmann
FAU Erlangen-Nürnberg, Germany
Daphne Ippolito
Daphne Ippolito
Carnegie Mellon University
natural language processing
Somesh Jha
Somesh Jha
Lubar Chair of Computer Science, University of Wisconsin
Trustworthy Machine LearningSecurityFormal methodsProgramming Languages
S
Sean McGregor
Responsible AI collaborative, US
Esfandiar Mohammadi
Esfandiar Mohammadi
Universität zu Lübeck
Differential PrivacyPrivacy-Preserving Machine LearningAnonymous Communication
L
Luca Nannini
Trustora Digital, Spain
Cristina Nita-Rotaru
Cristina Nita-Rotaru
Professor, Khoury College of Computer Science, Northeastern University
network securitydistributed systemsbyzantine-resiliencetrustworthy AI
Alina Oprea
Alina Oprea
Northeastern University
Computer SecurityAdversarial Machine LearningAI Security
K
Kevin Paeth
UL Research Institutes, US
Andrew Paverd
Andrew Paverd
Microsoft
SecurityPrivacy
Jonathan Petit
Jonathan Petit
Qualcomm
Computer ScienceVehicular NetworksDistributed SystemsSecurityPrivacy
Andreas Rauber
Andreas Rauber
TU Wien / Vienna University of Technology
Machine LearningData ScienceInformation RetrievalMusic Information RetrievalDigital Preservation
Christian Riess
Christian Riess
Friedrich-Alexander-University of Erlangen-Nuremberg
Digital ForensicsMultimedia SecurityMachine LearningImage Processing
J
John Sotiropoulos
Deep Cyber/OW ASP GenAI Security Project, United Kingdom
A
Andreas Wespi
IBM Research Europe–Zurich, Switzerland
Kathrin Grosse
Kathrin Grosse
IBM Research
AI Security (in practice)ML Security (in practice)