arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 1832 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. AI治理与伦理 1832 篇

2308.09979 2023-08-22 cs.CY cs.AI 62%

Artificial Intelligence across Europe: A Study on Awareness, Attitude and Trust

Teresa Scantamburlo, Atia Cortés, Francesca Foffano, Cristian Barrué, Veronica Distefano, Long Pham, Alessandro Fabris

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.03198 2023-07-14 cs.CY cs.AI 62%

A multilevel framework for AI governance

Hyesun Choung, Prabu David, John S. Seberger

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY

Comments This paper has been accepted for publication and is forthcoming in The Global and Digital Governance Handbook. Cite as: Choung, H., David, P., & Seberger, J.S. (2023). A multilevel framework for AI governance. The Global and Digital Governance Handbook. Routledge, Taylor & Francis Group

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.01800 2023-06-06 cs.CY cs.AI 62%

The ethical ambiguity of AI data enrichment: Measuring gaps in research ethics norms and practices

Will Hawkins, Brent Mittelstadt

专题命中 AI治理与伦理 :RLHF(abstract);分类 cs.AI、cs.CY

Comments 10 pages

Journal ref 2023 ACM Conference on Fairness, Accountability, and Transparency

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.08486 2023-05-30 cs.CV cs.AI cs.LG 62%

Scalar Invariant Networks with Zero Bias

Chuqin Geng, Xiaojie Xu, Haolin Ye, Xujie Si

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.LG

Comments 22 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.17137 2023-05-30 cs.AI cs.LG 62%

Integrating Generative Artificial Intelligence in Intelligent Vehicle Systems

Lukas Stappen, Jeremy Dillmann, Serena Striegel, Hans-Jörg Vögel, Nicolas Flores-Herr, Björn W. Schuller

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.LG

Comments under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.11215 2023-04-25 cs.CY cs.AI 62%

ChatGPT: More than a Weapon of Mass Deception, Ethical challenges and responses from the Human-Centered Artificial Intelligence (HCAI) perspective

Alejo Jose G. Sison, Marco Tulio Daza, Roberto Gozalo-Brizuela, Eduardo C. Garrido-Merchán

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.03843 2023-03-07 cs.LG cs.CY 62%

Counterfactual Fairness Is Basically Demographic Parity

Lucas Rosenblatt, R. Teal Witter

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.CY、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.05442 2023-02-13 cs.CV cs.AI cs.LG 62%

Scaling Vision Transformers to 22 Billion Parameters

Mostafa Dehghani, Josip Djolonga, Basil Mustafa, Piotr Padlewski, Jonathan Heek, Justin Gilmer, Andreas Steiner, Mathilde Caron, Robert Geirhos, Ibrahim Alabdulmohsin, Rodolphe Jenatton, Lucas Beyer, Michael Tschannen, Anurag Arnab, Xiao Wang, Carlos Riquelme, Matthias Minderer, Joan Puigcerver, Utku Evci, Manoj Kumar, Sjoerd van Steenkiste, Gamaleldin F. Elsayed, Aravindh Mahendran, Fisher Yu, Avital Oliver, Fantine Huot, Jasmijn Bastings, Mark Patrick Collier, Alexey Gritsenko, Vighnesh Birodkar, Cristina Vasconcelos, Yi Tay, Thomas Mensink, Alexander Kolesnikov, Filip Pavetić, Dustin Tran, Thomas Kipf, Mario Lučić, Xiaohua Zhai, Daniel Keysers, Jeremiah Harmsen, Neil Houlsby

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.02323 2023-02-07 cs.LG cs.AI stat.ML 62%

Improving Fair Training under Correlation Shifts

Yuji Roh, Kangwook Lee, Steven Euijong Whang, Changho Suh

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.00479 2023-01-18 cs.CR cs.AI cs.CL stat.ML 62%

The Design Principle of Blockchain: An Initiative for the SoK of SoKs

Luyao Sunshine Zhang

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.08704 2022-12-01 cs.LG cs.CY 62%

Accurate Fairness: Improving Individual Fairness without Trading Accuracy

Xuran Li, Peng Wu, Jing Su

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.CY、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.03274 2022-11-24 cs.LG cs.AI 62%

TCNL: Transparent and Controllable Network Learning Via Embedding Human-Guided Concepts

Zhihao Wang, Chuang Zhu

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.15289 2022-10-28 cs.CY cs.AI 62%

On the Efficiency of Ethics as a Governing Tool for Artificial Intelligence

Nicholas Kluge Corrêa, Nythamar De Oliveira, Diogo Massmann

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.14975 2022-10-28 cs.CL cs.LG 62%

MABEL: Attenuating Gender Bias using Textual Entailment Data

Jacqueline He, Mengzhou Xia, Christiane Fellbaum, Danqi Chen

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.LG

Comments Accepted to EMNLP 2022. Code and models are publicly available at https://github.com/princeton-nlp/mabel

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.02667 2022-10-07 cs.AI cs.CY 62%

A Human Rights-Based Approach to Responsible AI

Vinodkumar Prabhakaran, Margaret Mitchell, Timnit Gebru, Iason Gabriel

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.CY

Comments Presented as a (non-archival) poster at the 2022 ACM Conference on Equity and Access in Algorithms, Mechanisms, and Optimization or (EAAMO '22)

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.11770 2022-09-27 cs.HC cs.AI cs.LG 62%

Toward Smart Doors: A Position Paper

Luigi Capogrosso, Geri Skenderi, Federico Girella, Franco Fummi, Marco Cristani

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.LG

Comments 2nd International Workshop on Industrial Machine Learning @ ICPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.12645 2022-08-29 cs.CY cs.AI 62%

The Brussels Effect and Artificial Intelligence: How EU regulation will impact the global AI market

Charlotte Siegmann, Markus Anderljung

专题命中 AI治理与伦理 :trustworthy(abstract);分类 cs.AI、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.07635 2022-08-19 cs.AI cs.CY 62%

AI Ethics Issues in Real World: Evidence from AI Incident Database

Mengyi Wei, Zhixuan Zhou

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY

Comments 56th Hawaii International Conference on System Sciences (HICSS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.03721 2022-06-14 cs.CY cs.AI 62%

Demystifying the Draft EU Artificial Intelligence Act

Michael Veale, Frederik Zuiderveen Borgesius

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY

Comments 16 pages, 1 table

Journal ref Computer Law Review International (2021), 22(4) 97-112

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.00088 2022-06-14 cs.CY cs.GT cs.LG 62%

AI for Social Impact: Learning and Planning in the Data-to-Deployment Pipeline

Andrew Perrault, Fei Fang, Arunesh Sinha, Milind Tambe

专题命中 AI治理与伦理 :safety(abstract);分类 cs.CY、cs.LG

Comments AI Magazine, Winter 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.03220 2022-06-08 cs.CY cs.AI 62%

A Transparency Index Framework for AI in Education

Muhammad Ali Chaudhry, Mutlu Cukurova, Rose Luckin

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY

Comments 17 pages, 4 Figures, 2 Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.10785 2022-05-24 cs.CY cs.AI 62%

Responsible Artificial Intelligence -- from Principles to Practice

Virginia Dignum

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.CY

Comments This paper is a curated version of my keynote at the Web Conference 2022. arXiv admin note: substantial text overlap with arXiv:2202.07446

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.04279 2022-05-10 cs.CY cs.AI 62%

Aligned with Whom? Direct and social goals for AI systems

Anton Korinek, Avital Balwit

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.AI、cs.CY

Comments Prepared for the Oxford Handbook of AI Governance (23 pages, 2 figures)

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.02894 2022-05-09 cs.HC cs.AI cs.CL 62%

Interactive Model Cards: A Human-Centered Approach to Model Documentation

Anamaria Crisan, Margaret Drouhard, Jesse Vig, Nazneen Rajani

专题命中 AI治理与伦理 :safety(abstract);分类 cs.CL、cs.AI

Comments To appear at ACM FAccT'22

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.09110 2021-12-07 cs.CY cs.AI 62%

Developing Future Human-Centered Smart Cities: Critical Analysis of Smart City Security, Interpretability, and Ethical Challenges

Kashif Ahmad, Majdi Maabreh, Mohamed Ghaly, Khalil Khan, Junaid Qadir, Ala Al-Fuqaha

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY

Comments 34 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.12146 2021-10-20 cs.CY cs.AI 62%

Avoiding Negative Side Effects due to Incomplete Knowledge of AI Systems

Sandhya Saisubramanian, Shlomo Zilberstein, Ece Kamar

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY

Comments 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.02972 2021-10-15 cs.LG cs.CY 62%

Empirical observation of negligible fairness-accuracy trade-offs in machine learning for public policy

Kit T. Rodolfa, Hemank Lamba, Rayid Ghani

专题命中 AI治理与伦理 :safety(abstract);分类 cs.CY、cs.LG

Comments 40 pages, 4 figures, 2 tables, 7 supplementary figures, 4 supplementary tables; revised to improve clarity and discussion

Journal ref Nat Mach Intell 3, 896-904 (2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.02203 2021-10-05 cs.CY cs.LG 62%

Accuracy-Efficiency Trade-Offs and Accountability in Distributed ML Systems

A. Feder Cooper, Karen Levy, Christopher De Sa

专题命中 AI治理与伦理 :safety(abstract);分类 cs.CY、cs.LG

Journal ref Equity and Access in Algorithms, Mechanisms, and Optimization (EAAMO 2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.02928 2021-09-15 cs.CY cs.AI 62%

Toward a Rational and Ethical Sociotechnical System of Autonomous Vehicles: A Novel Application of Multi-Criteria Decision Analysis

Veljko Dubljević, George F. List, Jovan Milojevich, Nirav Ajmeri, William Bauer, Munindar P. Singh, Eleni Bardaka, Thomas Birkland, Charles Edwards, Roger Mayer, Ioan Muntean, Thomas Powers, Hesham Rakha, Vance Ricks, M. Shoaib Samandar

专题命中 AI治理与伦理 :safety(abstract);分类 cs.AI、cs.CY

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.05704 2021-09-15 cs.CL cs.AI 62%

Mitigating Language-Dependent Ethnic Bias in BERT

Jaimeen Ahn, Alice Oh

专题命中 AI治理与伦理 :alignment(abstract);分类 cs.CL、cs.AI

Comments 17 pages including references and appendix. To appear in EMNLP 2021 (camera-ready ver.)

详情

展开后加载摘要…

URL PDF HTML 收藏