Evaluating AI Alignment in LLMs: Output Analysis of Value Priorities Across 75 Models with Human Benchmarking
评估大语言模型中的AI对齐:通过75个模型的人类基准测试分析价值优先级
Gabriel Rongyang Lau, Wei Yan Low, Seow Min Koh, Fiona Fui-Hoon Nah, Andree Hartanto
机构
*
School of Social Sciences, Nanyang Technological University(南洋理工大学社会科学学院)
;
Interdisciplinary Graduate Programme, Nanyang Technological University(南洋理工大学跨学科研究生项目)
;
Faculty of Arts and Social Sciences, National University of Singapore(新加坡国立大学人文与社会科学学院)
;
School of Computing and Information Systems, Singapore Management University(新加坡管理学院计算与信息学院)
;
School of Social Sciences, Singapore Management University(新加坡管理学院社会科学学院)
Estimating Item Difficulty with Large Language Models as Experts
利用大语言模型作为专家估算项目难度
Diana Kolesnikova, Kirill Fedyanin, Abe D. Hofman, Matthieu J. S. Brinkhuis, Maria Bolsinova
机构
*
Department of Methodology and Statistics, Tilburg University(蒂尔堡大学方法学与统计学系)
;
Smart Business Technologies(智能商务技术公司)
;
Department of Psychological Methods, University of Amsterdam(阿姆斯特丹大学心理方法系)
;
Prowise Learn, Amsterdam(Prowise Learn公司,阿姆斯特丹)
;
Department of Information and Computing Sciences, Utrecht University(乌得勒支大学信息与计算科学系)
Testable and Actionable Calibration for Full Swap Regret
可检验且可操作的全面交换懊悔校准
Konstantina Bairaktari, Lunjia Hu, Huy L. Nguyen, Jonathan Ullman
机构
*
Department of Computer Science, Aarhus University(阿arhus大学计算机科学系)
;
Khoury College of Computer Sciences, Northeastern University(东北大学计算机科学学院)
;
Northeastern University(东北大学)