arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 1836 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. AI治理与伦理 1836 篇

2408.16700 2024-08-30 cs.CV 50%

GradBias: Unveiling Word Influence on Bias in Text-to-Image Generative Models

Moreno D'Incà, Elia Peruzzo, Massimiliano Mancini, Xingqian Xu, Humphrey Shi, Nicu Sebe

专题命中 AI治理与伦理 :safety(abstract)

Comments Under review. Code: https://github.com/Moreno98/GradBias

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.14023 2024-07-22 cs.SE 50%

Towards Extracting Ethical Concerns-related Software Requirements from App Reviews

Aakash Sorathiya, Gouri Ginde

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.11323 2024-07-01 cs.IR 50%

Transparency, Privacy, and Fairness in Recommender Systems

Dominik Kowald

专题命中 AI治理与伦理 :trustworthy(abstract)

Comments Habilitation (post-doctoral thesis) at Graz University of Technology for the scientific subject "Applied Computer Science" (accepted in June 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.13912 2024-06-21 cs.CV 50%

From Descriptive Richness to Bias: Unveiling the Dark Side of Generative Image Caption Enrichment

Yusuke Hirota, Ryo Hachiuma, Chao-Han Huck Yang, Yuta Nakashima

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.04314 2024-06-04 cs.CV 50%

GPT4SGG: Synthesizing Scene Graphs from Holistic and Region-specific Narratives

Zuyao Chen, Jinlin Wu, Zhen Lei, Zhaoxiang Zhang, Changwen Chen

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.16430 2024-05-28 cs.RO cs.MA 50%

GAMEOPT+: Improving Fuel Efficiency in Unregulated Heterogeneous Traffic Intersections via Optimal Multi-agent Cooperative Control

Nilesh Suriyarachchi, Rohan Chandra, Arya Anantula, John S. Baras, Dinesh Manocha

专题命中 AI治理与伦理 :safety(abstract)

Comments Journal Version

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.12558 2024-04-22 cs.HC 50%

Just Like Me: The Role of Opinions and Personal Experiences in The Perception of Explanations in Subjective Decision-Making

Sharon Ferguson, Paula Akemi Aoyagui, Young-Ho Kim, Anastasia Kuzminykh

专题命中 AI治理与伦理 :trustworthy(abstract)

Comments Presented at the Trust and Reliance in Evolving Human-AI Workflows (TREW) Workshop at CHI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06777 2024-04-11 cs.NI 50%

Responsible Federated Learning in Smart Transportation: Outlooks and Challenges

Xiaowen Huang, Tao Huang, Shushi Gu, Shuguang Zhao, Guanglin Zhang

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.14504 2024-03-05 cs.HC 50%

People's Perceptions Toward Bias and Related Concepts in Large Language Models: A Systematic Review

Lu Wang, Max Song, Rezvaneh Rezapour, Bum Chul Kwon, Jina Huh-Yoo

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.05332 2024-02-07 cs.CV 50%

Gender Stereotyping Impact in Facial Expression Recognition

Iris Dominguez-Catena, Daniel Paternain, Mikel Galar

专题命中 AI治理与伦理 :safety(abstract)

Comments Presented at SoGood 2022, The 7th Workshop on Data Science for Social Good, held in conjunction with ECML PKDD 2022, in September 2022, at Grenoble, France

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.12861 2024-02-06 cs.MA cs.RO math.OC 50%

Hierarchical Control for Head-to-Head Autonomous Racing

Rishabh Saumil Thakkar, Aryaman Singh Samyal, David Fridovich-Keil, Zhe Xu, Ufuk Topcu

专题命中 AI治理与伦理 :safety(abstract)

Journal ref Field Robotics, 4, 46-69 (2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18420 2024-01-31 cs.CV 50%

TeG-DG: Textually Guided Domain Generalization for Face Anti-Spoofing

Lianrui Mu, Jianhong Bai, Xiaoxuan He, Jiangnan Ye, Xiaoyu Liang, Yuchen Yang, Jiedong Zhuang, Haoji Hu

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.10861 2023-11-21 econ.GN q-fin.EC 50%

First, Do No Harm: Algorithms, AI, and Digital Product Liability

Marc J. Pfeiffer

专题命中 AI治理与伦理 :safety(abstract)

Comments 22 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.15003 2023-11-14 cs.MA 50%

Feasible Action-Space Reduction as a Metric of Causal Responsibility in Multi-Agent Spatial Interactions

Ashwin George, Luciano Cavalcante Siebert, David Abbink, Arkady Zgonnikov

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.06219 2023-10-11 cs.SE 50%

Runtime Monitoring of Human-centric Requirements in Machine Learning Components: A Model-driven Engineering Approach

Hira Naveed

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.10650 2023-07-21 cs.IR 50%

Language-Enhanced Session-Based Recommendation with Decoupled Contrastive Learning

Zhipeng Zhang, Piao Tong, Yingwei Ma, Qiao Liu, Xujiang Liu, Xu Luo

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.11465 2023-05-22 cs.MA cs.RO 50%

Counterfactual Fairness Filter for Fair-Delay Multi-Robot Navigation

Hikaru Asano, Ryo Yonetani, Mai Nishimura, Tadashi Kozuno

专题命中 AI治理与伦理 :safety(abstract)

Comments To appear in the International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.09319 2023-05-17 cs.IR 50%

Fairness and Diversity in Information Access Systems

Lorenzo Porcaro, Carlos Castillo, Emilia Gómez, João Vinagre

专题命中 AI治理与伦理 :trustworthy(abstract)

Comments Presented at the European Workshop on Algorithmic Fairness (EWAF'23) Winterthur, Switzerland, June 7-9, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.09243 2023-05-17 cs.SI 50%

LogDoctor: an open and decentralized worker-centered solution for occupational management in healthcare

Sami Barrit, Alexandre Niset

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.09218 2023-04-20 cs.CV 50%

Generative models improve fairness of medical classifiers under distribution shifts

Ira Ktena, Olivia Wiles, Isabela Albuquerque, Sylvestre-Alvise Rebuffi, Ryutaro Tanno, Abhijit Guha Roy, Shekoofeh Azizi, Danielle Belgrave, Pushmeet Kohli, Alan Karthikesalingam, Taylan Cemgil, Sven Gowal

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.11360 2023-02-24 cs.IR 50%

Commonality in Recommender Systems: Evaluating Recommender Systems to Enhance Cultural Citizenship

Andres Ferraro, Gustavo Ferreira, Fernando Diaz, Georgina Born

专题命中 AI治理与伦理 :alignment(abstract)

Comments extended version of "Measuring Commonality in Recommendation of Cultural Content: Recommender Systems to Enhance Cultural Citizenship", published at RecSys 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.11134 2022-11-28 cs.CV 50%

Open Vocabulary Object Detection with Proposal Mining and Prediction Equalization

Peixian Chen, Kekai Sheng, Mengdan Zhang, Mingbao Lin, Yunhang Shen, Shaohui Lin, Bo Ren, Ke Li

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.01285 2022-08-03 cs.MA cs.GT 50%

Evaluating Inter-Operator Cooperation Scenarios to Save Radio Access Network Energy

Xavier Marjou, Tangui Le Gléau, Vincent Messié, Benoit Radier, Tayeb Lemlouma, Gaël Fromentoux

专题命中 AI治理与伦理 :safety(abstract)

Comments 5 pages, 2 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.08938 2022-06-22 q-fin.RM 50%

Baseline validation of a bias-mitigated loan screening model based on the European Banking Authority's trust elements of Big Data & Advanced Analytics applications using Artificial Intelligence

Alessandro Danovi, Marzio Roma, Davide Meloni, Stefano Olgiati, Fernando Metelli

专题命中 AI治理与伦理 :trustworthy(abstract)

Comments 13 pages, 4 tables, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.07555 2022-06-16 cs.HC 50%

Respect as a Lens for the Design of AI Systems

William Seymour, Max Van Kleek, Reuben Binns, Dave Murray-Rust

专题命中 AI治理与伦理 :safety(abstract)

Comments To appear in the Proceedings of the 2022 AAAI/ACM Conference on AI, Ethics, and Society (AIES '22)

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.09731 2022-05-20 cs.CV 50%

Towards Unified Keyframe Propagation Models

Patrick Esser, Peter Michael, Soumyadip Sengupta

专题命中 AI治理与伦理 :alignment(abstract)

Comments CVPRW 2022 - AI for Content Creation Workshop. Code at https://github.com/runwayml/guided-inpainting

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.11572 2022-03-21 cs.RO 50%

GAMEOPT: Optimal Real-time Multi-Agent Planning and Control for Dynamic Intersections

Nilesh Suriyarachchi, Rohan Chandra, John S. Baras, Dinesh Manocha

专题命中 AI治理与伦理 :safety(abstract)

Comments Submitted to ITSC 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.01121 2021-12-03 cs.CV 50%

"Just Drive": Colour Bias Mitigation for Semantic Segmentation in the Context of Urban Driving

Jack Stelling, Amir Atapour-Abarghouei

专题命中 AI治理与伦理 :safety(abstract)

Comments 2021 IEEE International Conference on Big Data (IEEE BigData 2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.03635 2021-11-08 cs.CV 50%

BBC-Oxford British Sign Language Dataset

Samuel Albanie, Gül Varol, Liliane Momeni, Hannah Bull, Triantafyllos Afouras, Himel Chowdhury, Neil Fox, Bencie Woll, Rob Cooper, Andrew McParland, Andrew Zisserman

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.02818 2021-11-02 cs.HC 50%

Why? Why not? When? Visual Explanations of Agent Behavior in Reinforcement Learning

Aditi Mishra, Utkarsh Soni, Jinbin Huang, Chris Bryan

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏