Probing Ethical Framework Representations in Large Language Models: Structure, Entanglement, and Methodological Challenges
在大型语言模型中探测伦理框架表示:结构、纠缠与方法学挑战
Weilun Xu, Alexander Rusnak, Frederic Kaplan
机构
*
School of Computer and Communication Sciences, École Polytechnique Fédérale de Lausanne(瑞士联邦理工学院计算机与通信科学学院)
;
Digital Humanities Lab, École Polytechnique Fédérale de Lausanne(瑞士联邦理工学院数字人文实验室)
专题命中
知识编辑与模型理解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
UGID: Unified Graph Isomorphism for Debiasing Large Language Models
UGID: 用于去偏大型语言模型的统一图同构
Zikang Ding, Junchi Yao, Junhao Li, Yi Zhang, Wenbo Jiang, Hongbo Liu, Lijie Hu
机构
*
University of Electronic Science and Technology of China(电子科技大学)
;
Mohamed bin Zayed University of Artificial Intelligence(Mohamed bin Zayed人工智能大学)
;
South China University of Technology(华南理工大学)
专题命中
知识编辑与模型理解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
The System Hallucination Scale (SHS): A Minimal yet Effective Human-Centered Instrument for Evaluating Hallucination-Related Behavior in Large Language Models
系统幻觉量表(SHS):一种最小但有效的以人类为中心的评估工具,用于评估大语言模型中的幻觉相关行为
Heimo Müller, Dominik Steiger, Markus Plass, Andreas Holzinger
机构
*
Machine Learning and Information Science Group, Medical University of Graz(格拉茨医科大学机器学习与信息科学组)
;
Human Machine Mind Cooperation, Graz, Austria(格拉茨人类机智合作中心)
;
MIDATA Cooperative, Zurich, Switzerland(苏黎世MIDATA合作组织)
;
Human-Centered AI Lab, BOKU University Vienna(维也纳BOKU大学人本AI实验室)
专题命中
知识编辑与模型理解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
Mechanistic Indicators of Understanding in Large Language Models
大语言模型中理解机制的指示器
Pierre Beckmann, Matthieu Queloz
机构
*
École Polytechnique Fédérale de Lausanne (EPFL)(联邦理工学院)
;
Idiap Research Institute(Idiap研究所)
;
University of Bern, Department of Philosophy(伯尔尼大学哲学系)
专题命中
知识编辑与模型理解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
GOFAI meets Generative AI: Development of Expert Systems by means of Large Language Models
GOFAI与生成式AI的交汇:通过大语言模型开发专家系统
Eduardo C. Garrido-Merchán, Cristina Puente
机构
*
Quantitative Methods Department, Comillas Pontifical University Institute of Research in Technology (IIT)(定量方法部门,科利马斯天主教大学技术研究所)
;
Computer Science Department, ICAI School of Engineering, Comillas Pontifical University(计算机科学系,ICAI工程学院,科利马斯天主教大学)
专题命中
知识编辑与模型理解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
FactCorrector: A Graph-Inspired Approach to Long-Form Factuality Correction of Large Language Models
FactCorrector:一种基于图的长文本事实性校正方法
Javier Carnerero-Cano, Massimiliano Pronesti, Radu Marinescu, Tigran Tchrakian, James Barry, Jasmina Gajcin, Yufang Hou, Alessandra Pascale, Elizabeth Daly
机构
*
IBM Research Europe - Ireland(IBM欧洲研究院-爱尔兰)
专题命中
知识编辑与模型理解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
On the Convergence of Moral Self-Correction in Large Language Models
Guangliang Liu, Haitao Mao, Bochuan Cao, Zhiyu Xue, Xitong Zhang, Rongrong Wang, Kristen Marie Johnson
机构
*
Michigan State University(密歇根州立大学)
;
Amazon(亚马逊)
;
Pennsylvania State University(宾夕法尼亚州立大学)
;
University of California, Santa Barbara(加州大学圣塔芭芭拉分校)
专题命中
知识编辑与模型理解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.LG
机构
*
Department of Electrical and Electronic Engineering, The Hong Kong Polytechnic University(香港理工大学电子与电气工程系)
;
School of Electronics and Information, Northwestern Polytechnical University(西北工业大学电子与信息学院)
专题命中
知识编辑与模型理解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI