On Fairness of Unified Multimodal Large Language Model for Image Generation
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY
Comments 25 pages, 1 figure
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
Journal ref Computing Conference 2025 (upcoming)
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.CY
Comments ICLR 2025
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY
Comments 30 pages, 6 tables, 14 figures
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY
Comments Presented at CLAIRvoyant (ConventicLE on Artificial Intelligence Regulation) Workshop 2024
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY
Comments This is a preprint of paper that has been accepted for Publication at 2024 IEEE International Conference on Trust, Privacy and Security in Intelligent Systems, and Applications
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.CY
Comments Di Jin and Xing Liu contributed equally to this work
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.LG
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.LG
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
专题命中 AI治理与伦理 :safety(abstract);分类 cs.CL、cs.CY
Comments 10 Page, 3 Figures. Accepted in: (i) ICML'24: LLMs & Cognition Workshop (Non-archival; OpenReview: https://openreview.net/forum?id=63C9YSc77p) (ii) EMNLP'24 : NLP for Science Workshop (Archival; ACL Anthology: https://aclanthology.org/2024.nlp4science-1.22/)
Journal ref Proceedings of the 1st Workshop on NLP for Science (NLP4Science), EMNLP 2024
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
Comments Accepted to NeurlIPS 2024, SoLaR workshop
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
Comments EMNLP 2024
Journal ref Proc. EMNLP (2024) 9004-9018
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY
Journal ref Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society (2024), 7(1), 438-450
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
Comments Accepted by KDD'25 ADS Track
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY
Comments Submitted for review to the Proceedings of the IEEE
专题命中 AI治理与伦理 :safety(abstract);分类 cs.CL、cs.AI
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.LG
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.LG
Comments NeurIPS 2024
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.LG
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
Comments Under Review
专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI
专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY
Comments 36 pages, 11 figures and tables
专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY
Comments Accepted as a workshop paper at MILCOM 2024, 8 pages