arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 1836 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. AI治理与伦理 1836 篇

2511.19886 2025-11-26 cs.CR cs.CV 50%

Frequency Bias Matters: Diving into Robust and Generalized Deep Image Forgery Detection

频率偏差至关重要:深入探讨鲁棒且通用的深度图像伪造检测

Chi Liu, Tianqing Zhu, Wanlei Zhou, Wei Zhao

机构 * Faculty of Data Science, City University of Macau(数据科学学院,澳门城市大学) Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences(深圳先进技术研究所,中国科学院)

专题命中 AI治理与伦理 :alignment(abstract)

AI总结 本文从频率视角分析深度图像伪造检测器的通用性和鲁棒性问题,提出两步频率对齐方法以提升检测可靠性并增强反伪造能力。

Comments Accepted for publication in IEEE Transactions on Dependable and Secure Computing

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15859 2025-11-21 cs.SE 50%

RE for AI in Practice: Managing Data Annotation Requirements for AI Autonomous Driving Systems

在实践中为AI管理数据标注需求:为AI自主驾驶系统管理数据标注要求

Hina Saeeda, Mazen Mohamad, Eric Knauss, Jennifer Horkoff, Ali Nouri

专题命中 AI治理与伦理 :safety(abstract)

AI总结 本研究通过实证分析,探讨了自动驾驶系统中数据标注需求的定义、挑战及最佳实践,为提升AI系统可靠性提供指导。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.22896 2025-11-17 physics.soc-ph 50%

Modelling vehicle and pedestrian collective dynamics: Challenges and advances

Antoine Tordeux, Cécile Appert-Rolland, Alexandre Nicolas, Armin Seyfried, Denis Ullmo

专题命中 AI治理与伦理 :safety(abstract)

Comments 20 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14686 2025-11-12 cs.CV 50%

From Semantics, Scene to Instance-awareness: Distilling Foundation Model for Grounded Open-vocabulary Situation Recognition

Chen Cai, Tianyi Liu, Jianjun Gao, Wenyang Liu, Kejun Wu, Ruoyu Wang, Yi Wang, Soo Chin Liew

机构 * National University of Singapore(新加坡国立大学) Nanyang Technological University(南洋理工大学) Huazhong University of Science and Technology(华中科技大学) The Hong Kong Polytechnic University(香港理工大学)

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15857 2025-11-11 cs.SI 50%

Investigating Prosocial Behavior Theory in LLM Agents under Policy-Induced Inequities

Yujia Zhou, Hexi Wang, Qingyao Ai, Zhen Wu, Yiqun Liu

专题命中 AI治理与伦理 :alignment(abstract)

Comments Accepted by AAAI2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14886 2025-11-04 cs.CV 50%

Surgical Scene Understanding in the Era of Foundation AI Models: A Comprehensive Review

Ufaq Khan, Umair Nawaz, Adnan Qayyum, Shazad Ashraf, Yutong Xie, Muhammad Haris Khan, Muhammad Bilal, Junaid Qadir

机构 * MBZ University of AI(MBZ人工智能大学) Hamad Bin Khalifa University(哈马德·本·卡西姆大学) Birmingham City University(伯明翰城市大学) Qatar University(卡塔尔大学) University Hospitals Birmingham and University of Birmingham(伯明翰大学医院和伯明翰大学)

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00026 2025-11-04 cs.RO 50%

Gen AI in Automotive: Applications, Challenges, and Opportunities with a Case study on In-Vehicle Experience

Chaitanya Shinde, Divya Garikapati

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.08878 2025-10-23 cs.HC 50%

Knowledge Prompting: How Knowledge Engineers Use Large Language Models

Elisavet Koutsiana, Johanna Walker, Michelle Nwachukwu, Bohui Zhang, Albert Meroño-Peñuela, Elena Simperl

专题命中 AI治理与伦理 :trustworthy(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17842 2025-10-22 cs.SE cs.HC 50%

Vibe Coding: Toward an AI-Native Paradigm for Semantic and Intent-Driven Programming

Vinay Bamil

专题命中 AI治理与伦理 :alignment(abstract)

Comments 10 pages, 1 figure, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17119 2025-10-21 cs.HC 50%

Design Framework for Conversational Agent in Couple relationships: A Systematic Review

Soyoung Jung, Sung Park

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.07314 2025-10-21 cs.IR 50%

Learnable Item Tokenization for Generative Recommendation

Wenjie Wang, Honghui Bao, Xinyu Lin, Jizhi Zhang, Yongqi Li, Fuli Feng, See-Kiong Ng, Tat-Seng Chua

专题命中 AI治理与伦理 :alignment(abstract)

Comments Accepted by CIKM 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.18837 2025-10-14 cs.SI 50%

Sentiment and Social Signals in the Climate Crisis: A Survey on Analyzing Social Media Responses to Extreme Weather Events

Pouya Shaeri, Yasaman Mohammadpour, Alimohammad Beigi, Ariane Middel

专题命中 AI治理与伦理 :alignment(abstract)

Comments Accepted and Published in SBP-BRiMS 2025. 18th International Conference on Social Computing, Behavioral-Cultural Modeling & Prediction and Behavior Representation in Modeling and Simulation

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06556 2025-10-09 econ.TH 50%

Artificial Intelligence in Port Logistics: A Bibliometric Analysis of Technological Integration and Research Dynamics

Abdelhafid Khazzar, Yassine Sekaki, Yasser Lachhab, Said El-marzouki

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06306 2025-10-09 cs.HC 50%

"Grillz on a hijabi": Intersectional Identities in Fostering Critical AI Literacy

Jaemarie Solyst, Chloe Fong, Faisal Nurdin, Rotem Landesman, R. Benjamin Shapiro

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06222 2025-10-09 cs.HC econ.GN q-fin.EC 50%

Inducing State Anxiety in LLM Agents Reproduces Human-Like Biases in Consumer Decision-Making

Ziv Ben-Zion, Zohar Elyoseph, Tobias Spiller, Teddy Lazebnik

专题命中 AI治理与伦理 :safety(abstract)

Comments Manuscript Main Text - 20 pages, including 3 Figures and 1 Table. Supplementary Materials - 10 pages, including 4 Supplemental Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.04038 2025-10-07 eess.SY cs.SY 50%

Distributed MPC-based Coordination of Traffic Perimeter and Signal Control: A Lexicographic Optimization Approach

Viet Hoang Pham, Hyo-Sung Ahn

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25963 2025-10-01 cs.CV 50%

Self-Supervised Anatomical Consistency Learning for Vision-Grounded Medical Report Generation

Longzhen Yang, Zhangkai Ni, Ying Wen, Yihang Liu, Lianghua He, Heng Tao Shen

机构 * Tongji University(同济大学) East China Normal University(华东师范大学) Shanghai Eye Disease Prevention and Treatment Center(上海眼病防治中心)

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22424 2025-09-29 q-bio.OT 50%

Desiderata for a biomedical knowledge network: opportunities, challenges and future Directions

Chunlei Wu, Hongfang Liu, Jason Flannick, Mark A. Musen, Andrew I. Su, Lawrence Hunter, Thomas M. Powers, Cathy H. Wu

专题命中 AI治理与伦理 :trustworthy(abstract)

Comments 6 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06700 2025-09-09 cs.NI 50%

Sovereign AI for 6G: Towards the Future of AI-Native Networks

Swarna Bindu Chetty, David Grace, Simon Saunders, Paul Harris, Eirini Eleni Tsiropoulou, Tony Quek, Hamed Ahmadi

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18600 2025-08-27 cs.GT cs.MA econ.GN q-fin.EC 50%

Bias-Adjusted LLM Agents for Human-Like Decision-Making via Behavioral Economics

Ayato Kitadai, Yusuke Fukasawa, Nariaki Nishino

专题命中 AI治理与伦理 :alignment(abstract)

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10286 2025-08-19 cs.HC 50%

Artificial Emotion: A Survey of Theories and Debates on Realising Emotion in Artificial Intelligence

Yupei Li, Qiyang Sun, Michelle Schlicher, Yee Wen Lim, Björn W. Schuller

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05281 2025-08-08 cs.HC 50%

A Methodological Framework and Questionnaire for Investigating Perceived Algorithmic Fairness

Ahmed Abdal Shafi Rasel, Ahmed Mustafa Amlan, Tasmim Shajahan Mim, Tanvir Hasan

专题命中 AI治理与伦理 :trustworthy(abstract)

Comments 34 pages, Submitted for review

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18328 2025-07-25 cs.NI 50%

Enhanced Velocity-Adaptive Scheme: Joint Fair Access and Age of Information Optimization in Vehicular Networks

Xiao Xu, Qiong Wu, Pingyi Fan, Kezhi Wang, Nan Cheng, Wen Chen, Khaled B. Letaief

专题命中 AI治理与伦理 :safety(abstract)

Comments This paper has been submitted to IEEE TMC

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21898 2025-07-09 cs.HC 50%

Bias, Accuracy, and Trust: Gender-Diverse Perspectives on Large Language Models

Aimen Gaba, Emily Wall, Tejas Ramkumar Babu, Yuriy Brun, Kyle Hall, Cindy Xiong Bearfield

专题命中 AI治理与伦理 :trustworthy(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05046 2025-07-08 cs.HC 50%

What Shapes User Trust in ChatGPT? A Mixed-Methods Study of User Attributes, Trust Dimensions, Task Context, and Societal Perceptions among University Students

Kadija Bouyzourn, Alexandra Birch

专题命中 AI治理与伦理 :alignment(abstract)

Comments 25 pages, 11 tables, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04534 2025-07-08 cs.SI 50%

Simulating User Watch-Time to Investigate Bias in YouTube Shorts Recommendations

Selimhan Dagtas, Mert Can Cakmak, Nitin Agarwal

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01069 2025-07-03 cs.CE cs.SE 50%

Agentic AI in Product Management: A Co-Evolutionary Model

Nishant A. Parikh

专题命中 AI治理与伦理 :alignment(abstract)

Comments 41 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.16889 2025-07-02 cs.CV 50%

Beyond Diagnostic Performance: Revealing and Quantifying Ethical Risks in Pathology Foundation Models

Weiping Lin, Shen Liu, Runchen Zhu, Yixuan Lin, Baoshun Wang, Liansheng Wang

专题命中 AI治理与伦理 :trustworthy(abstract)

Comments 33 pages,5 figure,23 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19081 2025-06-25 cond-mat.soft 50%

Emergent collective dynamics from motile photokinetic organisms

J. Morales, P. Munoz, D. Noto, H. N Ulloa, F. Guzman-Lastra

专题命中 AI治理与伦理 :alignment(abstract)

Comments 11 pages 5 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03409 2025-06-19 cs.CR 50%

Technical Options for Flexible Hardware-Enabled Guarantees

James Petrie, Onni Aarne

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏