Green Shielding: A User-Centric Approach Towards Trustworthy AI
绿色防护:面向可信AI的用户导向方法
Aaron J. Li, Nicolas Sanchez, Hao Huang, Ruijiang Dong, Jaskaran Bains, Katrin Jaradeh, Zhen Xiang, Bo Li, Feng Liu, Aaron Kornblith, Bin Yu
机构
*
University of California, Berkeley(加州大学伯克利分校)
;
University of Melbourne(墨尔本大学)
;
University of California, San Francisco(加州大学旧金山分校)
;
University of Georgia(佐治亚大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
Austin R. Ellis-Mohr, Max Hartman, Lav R. Varshney
机构
*
Department of Electrical and Computer Engineering, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校电子与计算机工程系)
;
AI Innovation Institute, Stony Brook University(石溪大学人工智能创新研究所)
h-MINT: Modeling Pocket-Ligand Binding with Hierarchical Molecular Interaction Network
h-MINT:基于分层分子相互作用网络的口袋-配体结合建模
Yanru Qu, Yijie Zhang, Wenjuan Tan, Xiangzhe Kong, Xiangxin Zhou, Chaoran Cheng, Mathieu Blanchette, Jiaxuan You, Ge Liu
机构
*
Department of Computer Science, UIUC(伊利诺伊大学厄巴纳-香槟分校计算机科学系)
;
DOE Center for Advanced Bioenergy and Bioproducts Innovation, UIUC(美国能源部先进生物能源与生物产品创新中心,伊利诺伊大学厄巴纳-香槟分校)
;
School of Computer Science, McGill University(麦吉尔大学计算机科学系)
;
MILA-Québec AI Institute(魁北克AI研究所)
;
Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)
;
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究 institute)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
A Tale of Two Variances: When Single-Seed Benchmarks Fail in Bayesian Deep Learning
双方差的故事:当单种子基准在贝叶斯深度学习中失效时
Qishi Zhan, Minxuan Hu, Liang He, Guansu Wang, Jiaxin Liu
机构
*
Department of Mathematical and Statistical Sciences(数学与统计科学系)
;
Marquette University(马凯特大学)
;
Cornell Ann S. Bowers College of Computing and Information Science(康奈尔大学安·S·博斯计算与信息科学学院)
;
Cornell University(康奈尔大学)
;
Physics Science and Engineering(物理科学与工程)
;
Tongji University(同济大学)
;
School of Computing and Information Systems(计算与信息科学学院)
;
The University of Melbourne(墨尔本大学)
;
Siebel School of Computing and Data Science(希贝尔计算与数据科学学院)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
Unstable Rankings in Bayesian Deep Learning Evaluation
贝叶斯深度学习评估中的不稳定性
Qishi Zhan, Minxuan Hu, Guansu Wang, Jiaxin Liu, Liang He
机构
*
Department of Mathematical and Statistical Sciences(数学与统计学系)
;
Marquette University(马凯特大学)
;
Cornell University(康奈尔大学)
;
The University of Melbourne(墨尔本大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Tongji University(同济大学)
Learning to Conceal Risk: Controllable Multi-turn Red Teaming for LLMs in the Financial Domain
学习隐藏风险:面向金融领域的可控多轮红队测试框架
Gang Cheng, Haibo Jin, Wenbin Zhang, Haohan Wang, Jun Zhuang
机构
*
Bloomberg(彭博社)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Florida International University(佛罗里达国际大学)
;
Boise State University(博伊西州立大学)
CommentsAccepted for ACL'26 (Main). TL;DR: We propose a controllable multi-turn risk-concealed red-teaming framework, CoRT, that progressively conceals surface-level risk while exploiting regulatory-violating behaviors on a proposed new benchmark, FinRisk-Bench