GraphAllocBench: A Flexible Benchmark for Preference-Conditioned Multi-Objective Policy Learning
GraphAllocBench:用于偏好条件多目标策略学习的灵活基准测试
Zhiheng Jiang, Yunzhe Wang, Ryan Marr, Ellen Novoseller, Benjamin T. Files, Volkan Ustun
机构
*
University of California, Los Angeles(加州大学洛杉矶分校)
;
USC Institute for Creative Technologies(美国南加州大学创意技术研究所)
;
U.S. Army DEVCOM Army Research Laboratory(美国陆军 DEVCOM 军事研究实验室)
Rating the Pitch, Not the Product: User Evaluations of LLMs Reflect Expectations More Than Performance
评价的是宣传话术,而非产品本身:用户对大语言模型的评估更多反映期望而非性能
Robert Morabito, Tyler McDonald, Charitra Viswanath, Angel Hsing-Chi Hwang, Susanne Gaube, Jad Kabbara, Ali Emami
机构
*
Brock University(布鲁克大学)
;
Emory University(埃默里大学)
;
University of Southern California(南加州大学)
;
University College London(伦敦大学学院)
;
Massachusetts Institute of Technology(麻省理工学院)
机构
*
Thomas Lord Department of Computer Science, University of Southern California, USA(汤姆斯·劳德计算机科学系,南加州大学,美国)
;
Signal Analysis and Interpretation Laboratory (SAIL), University of Southern California, USA(信号分析与解释实验室(SAIL),南加州大学,美国)
Efficient Cross-Validation for Sparse Linear Regression
稀疏线性回归的高效交叉验证
Ryan Cory-Wright, Andrés Gómez
机构
*
Department of Analytics, Marketing and Operations, Imperial Business School, London, UK(分析、营销与运营系,帝国商务学院,伦敦,英国)
;
Department of Industrial and Systems Engineering, Viterbi School of Engineering, University of Southern California, CA(工业与系统工程系,维特比工程学院,美国南加州大学,CA)
Accuracy, Uncertainty, and Adaptability of Automatic Myocardial ASL Segmentation using Deep CNN
自动心脏灌注成像中深度CNN的准确性、不确定性和适应性研究
Hung P. Do, Yi Guo, Andrew J. Yoon, Krishna S. Nayak
机构
*
Ming Hsieh Department of Electrical and Computer Engineering, Viterbi School of Engineering, University of Southern California(明希德电气与计算机工程系,维特比工程学院,南加州大学)
;
Long Beach Memorial Medical Center, University of California Irvine(长滩纪念医疗中心,加州大学伊万尼耶分校)
机构
*
University of California, Los Angeles(加州大学洛杉矶分校)
;
University of Southern California(南加州大学)
;
DeerLab LLC(DeerLab有限责任公司)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Clemson University(克莱姆森大学)
;
Google(谷歌)
;
San Jose State University(圣何塞州立大学)
;
University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
Staleness-Learning Rate Scaling Laws for Asynchronous RLHF
异步RLHF的陈旧度-学习率缩放定律
Jingwei Song, Haofeng Xu, Jie Xiao, Chengke Bao, Jingwei Shi, Pengbin Feng, Weixun Wang, Yuhang Han, Chuan Wu, Linfeng Zhang, Bill Shi
机构
*
The University of Hong Kong(香港大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Gradient
;
University of Southern California(南加州大学)
;
The Hong Kong Polytechnic University(香港理工大学)
机构
*
Department of Electrical Engineering and Computer Sciences University of California Berkeley(加州大学伯克利分校电气工程与计算机科学系)
;
University of Southern California(南加州大学)
HERO: Improving the Reliability and Sensitivity of Generative Model Evaluation Using Historical Data
HERO: 利用历史数据提升生成模型评估的可靠性与敏感性
Xinrui Ruan, Zhenyu Zhao, Waverly Wei, Yueshan Zhang, Zeyu Zheng, Sui Huang, Jingshen Wang
机构
*
Division of Biostatistics, University of California, Berkeley(加州大学伯克利分校生物统计学系)
;
Roblox Corporation(Roblox公司)
;
Department of Data Sciences and Operations, University of Southern California(南加州大学数据科学与运营系)
;
School of Mathematical Sciences, Nankai University(南开大学数学科学学院)
;
Department of Industrial Engineering and Operations Research, University of California, Berkeley(加州大学伯克利分校工业工程与运筹学系)
机构
*
Columbia University(哥伦比亚大学)
;
University of California, Berkeley(加州大学伯克利分校)
;
School of Information Systems and Management, Carnegie Mellon University(信息系统与管理学院,卡内基梅隆大学)
;
Department of Mathematics, University of Southern California(数学系,南加州大学)
;
University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校)
;
Computer Science Department, UC San Diego(计算机科学系,UCSD)