Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security
迈向可信的自主AI:安全性、鲁棒性、隐私与系统安全的全面综述
机构 * Faculty of Engineering, Department of Computer Science and Engineering, The Chinese University of Hong Kong(香港中文大学工程学院、计算机科学与工程系) ; Artificial Intelligence Innovation and Incubation Institute, Fudan University(复旦大学人工智能创新与孵化院) ; Shanghai Academy of AI for Science(上海人工智能科学研究院)
专题命中 安全评测 :trustworthy(title,abstract);safety(title,abstract);alignment(abstract);分类 cs.CL、cs.AI
AI总结 本文综述了自主AI系统在安全鲁棒性与隐私系统安全两个核心维度的风险来源、阶段缓解策略及统一评估指标,并讨论了开放挑战。
Comments 36 pages, 4 figures. Survey/review article on trustworthy agentic AI. Published in Academia AI and Applications, 2026
Journal ref Academia AI and Applications, vol. 2, 2026