Decoupled Alignment for Robust Plug-and-Play Adaptation
用于鲁棒即插即用适应的解耦对齐
Haozheng Luo, Jiahao Yu, Wenxin Zhang, Jialong Li, Chenghao Qiu, Yimin Wang, Eric Hanchen Jiang, Jerry Yao-Chieh Hu, Yan Chen, Binghui Wang, Xinyu Xing, Han Liu
机构
*
Northwestern University(西北大学)
;
New York University Abu Dhabi(纽约大学阿布扎克分校)
;
Stanford University(斯坦福大学)
;
Texas A&M University(德克萨斯农工大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
Illinois Institute of Technology(伊利诺伊理工学院)
CommentsRevised to correct the Acknowledgments section. Previous versions inadvertently included acknowledgments of NSF and NIH awards that did not support this work. Those funding acknowledgments have been removed. The technical content, results, and conclusions are unchanged
Safe-SAIL: Towards a Fine-grained Safety Landscape of Large Language Models via Sparse Autoencoder Interpretation Framework
Safe-SAIL: 通过稀疏自编码解释框架构建大语言模型的细粒度安全景观
Jiaqi Weng, Han Zheng, Hanyu Zhang, Ej Zhou, Qinqin He, Jialing Tao, Hui Xue, Zhixuan Chu, Xiting Wang
机构
*
Alibaba Group(阿里巴巴集团)
;
The State Key Laboratory of Blockchain and Data Security, Zhejiang University(浙江大学区块链与数据安全国家重点实验室)
;
Language Technology Lab, University of Cambridge(剑桥大学语言技术实验室)
;
Renmin University of China(中国人民大学)
Code-MUE: Measuring Code LLMs' Uncertainty through Execution-based Semantic Interaction Graphs
Code-MUE:通过基于执行的语义交互图测量代码语言模型的不确定性
Xiaoning Ren, Yinxing Xue, Lei Ma, Yuheng Huang
机构
*
Xi’an Jiaotong University(西安交通大学)
;
Institute of AI for Industries, Chinese Academy of Sciences(中国科学院人工智能产业研究院)
;
The University of Tokyo(东京大学)
;
University of Alberta(阿尔伯塔大学)
Perception-Aligned AI Outputs: End-to-End Visual Prediction for Uncertainty Communication in Clinical Decision-Making
感知对齐的人工智能输出:临床决策中用于不确定性通信的端到端视觉预测
Mohammad Eslami, Solale Tabarestani, Saber Kazeminasab, Ehsan Adeli, Glyn Elwyn, Tobias Elze, Mengyu Wang, Nazlee Zebardast, Lucia Sobrin, Nassir Navab, Daniel Shu Wei Ting, Malek Adjouadi
机构
*
Harvard Ophthalmology AI Lab(哈佛眼科人工智能实验室)
;
Schepens Eye Research Institute of Massachusetts Eye and Ear(马萨诸塞眼耳医院施佩恩眼科研究所)
;
Harvard Medical School(哈佛医学院)
;
Center for Advanced Technology and Education(先进教育技术中心)
;
Florida International University(佛罗里达国际大学)
;
Dartmouth Institute for Health Policy and Clinical Practice(达特茅斯健康政策与临床实践研究所)
;
Dartmouth College(达特茅斯学院)
;
Computer Aided Medical Procedures(医学辅助程序)
;
Technical University of Munich(慕尼黑技术大学)
;
Singapore Eye Research Institute(新加坡眼科研究所)
;
Singapore National Eye Centre(新加坡国家眼科中心)
;
Department of Ophthalmology, Byers Eye Institute, Stanford University(眼科部门,比尔斯眼科研究所,斯坦福大学)
Closing the AI Trust Gap: The Case for Independent Certification for Trustworthy AI
缩小人工智能信任差距:可信人工智能独立认证的案例
Trisevgeni Papakonstantinou, Cansu Canca, Farah Nanji, Waheedullah Pardess, Jen Weedon, Jasmijn Remmers, Eliza Krigman, Matthew Ball, Yalda Daryani, Kiran Iqbal, Francielle Vargas, María Llorente Sánchez, Joe Humphreys, Fendi Tsim, Kelly Fitzpatrick, Jeff Dunn, Catherine Feldman
机构
*
University College London(伦敦大学学院)
;
AI Ethics Lab(人工智能伦理实验室)
;
University of Cambridge(剑桥大学)
;
Columbia University(哥伦比亚大学)
;
MATS
;
Tech with Intention(技术与意图)
;
University of Southern California(南加州大学)
;
Kyushu University(九州大学)
;
University of Chile(智利大学)
;
BehSci Meets AI(行为科学与人工智能)
;
Digital Trust Council(数字信任委员会)
Verbalizable Representations Form a Global Workspace in Language Models
语言模型中的可言语化表征形成全局工作空间
Wes Gurnee, Nicholas Sofroniew, Adam Pearce, Mateusz Piotrowski, Isaac Kauvar, Runjin Chen, Anna Soligo, Paul Bogdan, Euan Ong, Rowan Wang, Ben Thompson, David Abrahams, Subhash Kantamneni, Emmanuel Ameisen, Joshua Batson, Jack Lindsey
MGDT: MLLM-Guided Diffusion Transformer with Relation-Adaptive Mixture-of-Experts for Multimodal Knowledge Graph Completion
MGDT:具有关系自适应专家混合的MLLM引导扩散变压器用于多模态知识图谱补全
Xu Hou, Meiyu Liang, Wei Huang, Yawen Li, Zhe Xue, Wu Liu, Guanhua Ye, Lei Shi, Kangkang Lu
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Zhejiang University(浙江大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Communication University of China(中国传媒大学)
A Transportable Threshold-Based Framework for Interpretable Classification of Medical Data
一种基于阈值的可移植框架用于医学数据的可解释分类
Antony Garcia, Adrian Noriega, Gabrielle Britton, Xinming Huang
机构
*
Worcester Polytechnic Institute(伍斯特理工学院)
;
Centro de Vacunación e Investigación (CEVAXIN)(疫苗接种与研究中心(CEVAXIN))
;
Universidad Tecnológica de Panamá(巴拿马技术大学)
;
Massachusetts Institute of Technology(麻省理工学院)
;
Broad Institute of MIT and Harvard(麻省理工学院和哈佛大学布罗德研究所)
;
McGill University(麦吉尔大学)
;
The Montreal Neurological Hospital-Institute(蒙特利尔神经学医院研究所)
Can We Trust Item Response Theory for AI Evaluation?
我们能信任项目反应理论进行人工智能评估吗?
Han Jiang, Sunbeom Kwon, Jinwen Luo, Ziang Xiao, Susu Zhang
机构
*
Johns Hopkins University(约翰·霍普金斯大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
University of California, Los Angeles(加利福尼亚大学洛杉矶分校)