arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 1836 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. AI治理与伦理 1836 篇

2506.14166 2025-06-18 cs.HC 50%

Affective-CARA: A Knowledge Graph Driven Framework for Culturally Adaptive Emotional Intelligence in HCI

Nirodya Pussadeniya, Bahareh Nakisa, Mohmmad Naim Rastgoo

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07766 2025-06-11 cs.CR cs.SE econ.TH 50%

Realigning Incentives to Build Better Software: a Holistic Approach to Vendor Accountability

Gergely Biczók, Sasha Romanosky, Mingyan Liu

专题命中 AI治理与伦理 :safety(abstract)

Comments accepted to WEIS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17937 2025-05-29 cs.HC 50%

Survival Games: Human-LLM Strategic Showdowns under Severe Resource Scarcity

Zhihong Chen, Yiqian Yang, Jinzhao Zhou, Qiang Zhang, Chin-Teng Lin, Yiqun Duan

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19895 2025-05-27 cs.CV 50%

Underwater Diffusion Attention Network with Contrastive Language-Image Joint Learning for Underwater Image Enhancement

Afrah Shaahid, Muzammil Behzad

机构 * King Fahd University of Petroleum and Minerals(国王法赫德石油与矿物大学)

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19045 2025-05-27 econ.GN q-fin.EC 50%

A General Theory of Growth, Employment, and Technological Change: Experiential Matrix Theory and the Transition from GDP to Humanist Experiential Growth in the Age of Artificial Intelligence

Christian Callaghan

专题命中 AI治理与伦理 :alignment(abstract)

Comments 57 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04752 2025-05-14 cs.IR 50%

Investigating Popularity Bias Amplification in Recommender Systems Employed in the Entertainment Domain

Dominik Kowald

专题命中 AI治理与伦理 :trustworthy(abstract)

Comments Accepted at EWAF'25, summarizes fairness and popularity bias research presented in Dr. Kowald's habilitation: https://domkowald.github.io/documents/others/2024habilitation_recsys.pdf

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07377 2025-05-13 cs.IR 50%

Process-Supervised LLM Recommenders via Flow-guided Tuning

Chongming Gao, Mengyao Gao, Chenxiao Fan, Shuai Yuan, Wentao Shi, Xiangnan He

专题命中 AI治理与伦理 :alignment(abstract)

Comments Accepted by SIGIR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05596 2025-05-12 econ.GN q-fin.EC 50%

Multi-level Governance, Smart Meter Adoption, and Utilities' Energy Efficiency Savings in the U.S

Yue Gao, Jing Zhang

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07787 2025-04-11 cs.SE 50%

Fairness Mediator: Neutralize Stereotype Associations to Mitigate Bias in Large Language Models

Yisong Xiao, Aishan Liu, Siyuan Liang, Xianglong Liu, Dacheng Tao

专题命中 AI治理与伦理 :alignment(abstract)

Comments Accepted by ISSTA 2025.20 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07610 2025-04-11 cs.MA 50%

What Contributes to Affective Polarization in Networked Online Environments? Evidence from an Agent-Based Model

Narayani Vedam, Subhayan Mukerjee, Prasanta Bhattacharya

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.17671 2025-04-04 cs.CV 50%

A Bias-Free Training Paradigm for More General AI-generated Image Detection

Fabrizio Guillaro, Giada Zingarini, Ben Usman, Avneesh Sud, Davide Cozzolino, Luisa Verdoliva

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21165 2025-03-28 eess.SY cs.AR cs.SY 50%

Extending Silicon Lifetime: A Review of Design Techniques for Reliable Integrated Circuits

Shaik Jani Babu, Fan Hu, Linyu Zhu, Sonal Singhal, Xinfei Guo

专题命中 AI治理与伦理 :safety(abstract)

Comments This work is under review by ACM

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.09230 2025-03-24 cs.CV 50%

Surgical Text-to-Image Generation

Chinedu Innocent Nwoye, Rupak Bose, Kareem Elgohary, Lorenzo Arboit, Giorgio Carlino, Joël L. Lavanchy, Pietro Mascagni, Nicolas Padoy

专题命中 AI治理与伦理 :alignment(abstract)

Comments 13 pages, 13 figures, 3 tables, published in Pattern Recognition Letters 2025, project page at https://camma-public.github.io/endogen/

Journal ref Pattern Recognition Letters, Volume 190, April 2025, Pages 73-80

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15949 2025-03-21 cs.CV 50%

CausalCLIPSeg: Unlocking CLIP's Potential in Referring Medical Image Segmentation with Causal Intervention

Yaxiong Chen, Minghong Wei, Zixuan Zheng, Jingliang Hu, Yilei Shi, Shengwu Xiong, Xiao Xiang Zhu, Lichao Mou

专题命中 AI治理与伦理 :alignment(abstract)

Comments MICCAI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10831 2025-03-20 cs.CV 50%

Low-Biased General Annotated Dataset Generation

Dengyang Jiang, Haoyu Wang, Lei Zhang, Wei Wei, Guang Dai, Mengmeng Wang, Jingdong Wang, Yanning Zhang

专题命中 AI治理与伦理 :alignment(abstract)

Comments CVPR2025 Accepted Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09024 2025-03-13 cs.RO cs.SY eess.IV eess.SY 50%

Traffic Regulation-aware Path Planning with Regulation Databases and Vision-Language Models

Xu Han, Zhiwen Wu, Xin Xia, Jiaqi Ma

专题命中 AI治理与伦理 :safety(abstract)

Comments 7 pages, 7 figures, submitted to ICRA

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.11464 2025-03-13 cs.CV 50%

High-Quality Mask Tuning Matters for Open-Vocabulary Segmentation

Quan-Sheng Zeng, Yunheng Li, Daquan Zhou, Guanbin Li, Qibin Hou, Ming-Ming Cheng

专题命中 AI治理与伦理 :alignment(abstract)

Comments Revised version according to comments from reviewers of ICLR2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.03270 2025-03-06 cs.CV cs.CR 50%

Reduced Spatial Dependency for More General Video-level Deepfake Detection

Beilin Chu, Xuan Xu, Yufei Zhang, Weike You, Linna Zhou

专题命中 AI治理与伦理 :safety(abstract)

Comments 5 pages, 2 figures. Accepted to ICASSP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19822 2025-02-28 cs.HC 50%

Empowering Social Service with AI: Insights from a Participatory Design Study with Practitioners

Yugin Tan, Kai Xin Soh, Renwen Zhang, Jungup Lee, Han Meng, Biswadeep Sen, Yi-Chieh Lee

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.11752 2025-02-20 cs.CV 50%

Are generative models fair? A study of racial bias in dermatological image generation

Miguel López-Pérez, Søren Hauberg, Aasa Feragen

专题命中 AI治理与伦理 :trustworthy(abstract)

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.11024 2025-02-18 cs.CV 50%

TPCap: Unlocking Zero-Shot Image Captioning with Trigger-Augmented and Multi-Modal Purification Modules

Ruoyu Zhang, Lulu Wang, Yi He, Tongling Pan, Zhengtao Yu, Yingna Li

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.13459 2025-02-18 cs.CV 50%

Adapting Multi-modal Large Language Model to Concept Drift From Pre-training Onwards

Xiaoyu Yang, Jie Lu, En Yu

专题命中 AI治理与伦理 :alignment(abstract)

Comments ICLR 2025 Poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.12167 2025-02-11 econ.GN math.OC q-fin.EC 50%

A Principal-Agent Model for Optimal Incentives in Renewable Investments

René Aïd, Annika Kemper, Nizar Touzi

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.03549 2025-02-11 cs.CV 50%

Kronecker Mask and Interpretive Prompts are Language-Action Video Learners

Jingyi Yang, Zitong Yu, Xiuming Ni, Jia He, Hui Li

专题命中 AI治理与伦理 :alignment(abstract)

Comments Accepted to ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.11548 2025-01-06 cs.IR 50%

A PLMs based protein retrieval framework

Yuxuan Wu, Xiao Yi, Yang Tan, Huiqun Yu, Guisheng Fan, Gaowei Zheng

专题命中 AI治理与伦理 :alignment(abstract)

Comments 16 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.11938 2024-12-17 eess.IV cs.CV 50%

Are the Latent Representations of Foundation Models for Pathology Invariant to Rotation?

Matouš Elphick, Samra Turajlic, Guang Yang

专题命中 AI治理与伦理 :alignment(abstract)

Comments Samra Turajlic and Guang Yang are joint last authors

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.01007 2024-11-05 cs.HC 50%

When Two Wrongs Don't Make a Right" -- Examining Confirmation Bias and the Role of Time Pressure During Human-AI Collaboration in Computational Pathology

Emely Rosbach, Jonas Ammeling, Sebastian Krügel, Angelika Kießig, Alexis Fritz, Jonathan Ganz, Chloé Puget, Taryn Donovan, Andrea Klang, Maximilian C. Köller, Pompei Bolfa, Marco Tecilla, Daniela Denk, Matti Kiupel, Georgios Paraschou, Mun Keong Kok, Alexander F. H. Haake, Ronald R. de Krijger, Andreas F. -P. Sonnen, Tanit Kasantikul, Gerry M. Dorrestein, Rebecca C. Smedley, Nikolas Stathonikos, Matthias Uhl, Christof A. Bertram, Andreas Riener, Marc Aubreville

专题命中 AI治理与伦理 :alignment(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.14129 2024-10-30 cs.CV cs.MM 50%

Mining Generalized Features for Detecting AI-Manipulated Fake Faces

Yang Yu, Rongrong Ni, Yao Zhao

专题命中 AI治理与伦理 :alignment(abstract)

Comments 14 pages, 9 figures. This work has been submitted to the IEEE for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.04508 2024-09-10 cs.HC 50%

Toward LLM-Powered Social Robots for Supporting Sensitive Disclosures of Stigmatized Health Conditions

Alemitu Bezabih, Shadi Nourriz, C. Estelle Smith

专题命中 AI治理与伦理 :safety(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.02433 2024-09-05 cs.SE 50%

From Literature to Practice: Exploring Fairness Testing Tools for the Software Industry Adoption

Thanh Nguyen, Luiz Fernando de Lima, Maria Teresa Badassarre, Ronnie de Souza Santos

专题命中 AI治理与伦理 :trustworthy(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏