Data Privacy and Trustworthy Machine Learning
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.LG
Comments Copyright ©2022, IEEE
Journal ref Published in: IEEE Security & Privacy ( Volume: 20, Issue: 5, Sept.-Oct. 2022)
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.LG
Comments Copyright ©2022, IEEE
Journal ref Published in: IEEE Security & Privacy ( Volume: 20, Issue: 5, Sept.-Oct. 2022)
专题命中 安全评测 :safety(title,abstract);分类 cs.LG
Comments Principles of Distribution Shift (PODS) Workshop at ICML 2022, 4 pages, 2 figures
专题命中 安全评测 :safety(title,abstract);分类 cs.AI
Comments As presented at the Embedded World Conference, Nuremberg, 2022
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.LG
Comments Accepted to 39th International Conference on Machine Learning, Workshop on Healthcare AI and COVID-19
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.LG
Comments In proceedings of the 11th Bulk Power Systems Dynamics and Control Symposium (IREP 2022), July 25-30, 2022, Banff, Canada. 21 pages, 12 figures, 5 tables
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.CY
专题命中 安全评测 :safety(title,abstract);分类 cs.AI
Journal ref AISafety, Jul 2022, Vienne, Austria
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.CY
专题命中 安全评测 :alignment(title,abstract);分类 cs.LG
Comments to appear at KDD 2022, the software package is at https://github.com/jyanln/AlignReg. arXiv admin note: text overlap with arXiv:2011.13052
专题命中 安全评测 :alignment(title,abstract);分类 cs.AI
专题命中 安全评测 :alignment(title,abstract);分类 cs.AI
Comments to appear in VLDB Journal, 2022
专题命中 安全评测 :alignment(title);trustworthy(abstract);分类 cs.LG
Comments 17 pages, 10 figures. Published in CHI 2022. For more details, see http://shared-interest.csail.mit.edu
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.LG
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.LG
Comments 20 pages (58 pages pre-print), 6 figures
Journal ref Journal of Information Security and Applications 65 (2022) 103121
专题命中 安全评测 :safety(title,abstract);分类 cs.AI
Comments 8 pages
专题命中 安全评测 :safety(title,abstract);分类 cs.LG
专题命中 安全评测 :safety(title,abstract);分类 cs.LG
Comments Accepted Scitech, AI for Space
专题命中 安全评测 :alignment(title,abstract);分类 cs.CL
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.AI
专题命中 安全评测 :safety(title,abstract);分类 cs.AI
Journal ref 26th IEEE Pacific Rim International Symposium on Dependable Computing (PRDC 2021), IEEE, Dec 2021, Perth, Australia
专题命中 安全评测 :safety(title,abstract);分类 cs.LG
Comments Paper accepted at IEEE AI Test'21
专题命中 安全评测 :safety(title,abstract);分类 cs.LG
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.LG
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.LG
Journal ref IEEE Transactions on Emerging Topics in Computational Intelligence 0 (2021) 1-12
专题命中 安全评测 :alignment(title,abstract);分类 cs.CL
Comments EACL 2021
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.LG
Comments Accepted in 2021 KDD
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.LG
专题命中 安全评测 :trustworthy(title,abstract);分类 cs.AI
Comments 9 pages, 1 figure, preprint. Accepted in RCIS 2021
专题命中 安全评测 :alignment(title);trustworthy(abstract);分类 cs.LG
Comments ECCV 2020, Code at https://github.com/nupurkmr9/Attributional-Robustness
专题命中 安全评测 :safety(title,abstract);分类 cs.LG
Comments Second (revised) version for public access