Estimating Tail Risks in Language Model Output Distributions
语言模型输出分布中的尾部风险估计
Rico Angell, Raghav Singhal, Zachary Horvitz, Zhou Yu, Rajesh Ranganath, Kathleen McKeown, He He
机构
*
Columbia University(哥伦比亚大学)
;
Department of Computer Science, New York University(纽约大学计算机科学系)
;
Center for Data Science, New York University(纽约大学数据科学中心)
Expert-Choice Routing Enables Adaptive Computation in Diffusion Language Models
专家选择路由使扩散语言模型实现自适应计算
Shuibai Zhang, Caspian Zhuang, Chihan Cui, Zhihan Yang, Fred Zhangzhi Peng, Yanxin Zhang, Haoyue Bai, Zack Jia, Yang Zhou, Guanhua Chen, Ming Liu
机构
*
University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
;
Scitix
;
Cornell University(康奈尔大学)
;
Duke University(杜克大学)
;
UC Davis(加州大学戴维斯分校)
;
Southern University of Science and Technology(南方科技大学)
Atlas 2 -- Foundation models for clinical deployment
Atlas 2 -- 临床应用的基础模型
Maximilian Alber, Timo Milbich, Alexandra Carpen-Amarie, Stephan Tietz, Jonas Dippel, Lukas Muttenthaler, Beatriz Perez Cancer, Alessandro Benetti, Panos Korfiatis, Elias Eulig, Jérôme Lüscher, Jiasen Wu, Sayed Abid Hashimi, Gabriel Dernbach, Simon Schallenberg, Neelay Shah, Moritz Krügener, Aniruddh Jammoria, Jake Matras, Patrick Duffy, Matt Redlon, Philipp Jurmeister, David Horst, Lukas Ruff, Klaus-Robert Müller, Frederick Klauschen, Andrew Norgan
机构
*
Aignostics, Germany(德国Aignostics公司)
;
Department of Laboratory Medicine and Pathology, Mayo Clinic, Rochester, MN, US(美国梅奥诊所实验室医学与病理学部门)
;
Department of Radiology, Mayo Clinic, Rochester MN, US(美国梅奥诊所放射学部门)
;
Department of Information Technology, Mayo Clinic, Rochester MN, US(美国梅奥诊所信息技术部门)
;
Mayo Clinic, Rochester MN, US(美国梅奥诊所)
;
Digital Pathology, Mayo Clinic, Rochester MN, US(美国梅奥诊所数字病理学部门)
;
Machine Learning Group, Technische Universität Berlin, Germany(德国柏林技术大学机器学习小组)
;
BIFOLD – Berlin Institute for the Foundations of Learning and Data, Germany(德国柏林学习与数据基础研究所)
;
Department of Artificial Intelligence, Korea University, Republic of Korea(韩国韩国大学人工智能系)
;
Max-Planck Institute for Informatics, Germany(德国马克斯·普朗克信息研究所)
;
German Cancer Research Center (DKFZ) & German Cancer Consortium (DKTK), Berlin & Munich Partner Sites, Germany(德国癌症研究中心(DKFZ)及德国癌症联盟(DKTK)柏林与慕尼黑合作站点)
;
Institute of Pathology, Ludwig-Maximilians-Universität München, Germany(德国慕尼黑路德维希-马克西米利安大学病理学研究所)
;
Institute of Pathology, Charité – Universitätsmedizin Berlin, Germany(德国柏林夏里特大学医学中心病理学研究所)
;
Bavarian Cancer Research Center (BZKF), Germany(德国巴伐利亚癌症研究中心(BZKF))
;
Helmholtz Munich, Germany(德国海德堡-慕尼黑亥姆霍兹中心)
;
Technical University Munich, Germany(德国慕尼黑技术大学)
Entropy-Aware On-Policy Distillation of Language Models
熵感知的在线策略蒸馏语言模型
Woogyeol Jin, Taywon Min, Yongjin Yang, Dennis Wei, Yi Zhou, Swanand Ravindra Kadhe, Nathalie Baracaldo, Kimin Lee
机构
*
IBM Research, San Jose, CA, USA(IBM研究院,旧金山,加州,美国)
;
University of Toronto, Ontario, Canada(多伦多大学,安大略,加拿大)
;
Vector Institute, Ontario, Canada(向量研究所,安大略,加拿大)
MirrorCheck: Efficient Adversarial Defense for Vision-Language Models
MirrorCheck: 视觉-语言模型的高效对抗防御
Samar Fares, Klea Ziu, Toluwani Aremu, Nikita Durasov, Martin Takáč, Pascal Fua, Ivan Laptev, Karthik Nandakumar
机构
*
Mohamed Bin Zayed University of Artificial Intelligence(莫扎伊德大学人工智能大学)
;
NVIDIA
;
École Polytechnique Fédérale de Lausanne(洛桑联邦理工学院)
;
Michigan State University(密歇根州立大学)