The Labyrinth and the Thread: Rethinking Regularizations in Sequential Knowledge Editing for Large Language Models
迷宫与线索:重新思考大语言模型顺序知识编辑中的正则化方法
Zheng Wang, Kaixuan Zhang, Wanfang Chen, Jingwen Zhang, Xiaonan Lu
机构
*
Bosch Center for Artificial Intelligence (BCAI)(博世人工智能中心(BCAI))
;
Bosch (China) Investment Ltd.(博世(中国)投资有限公司)
;
School of Statistics, East China Normal University(东华大学统计学院)
专题命中
知识编辑与模型理解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
Innovation: An Almost Characterization of Hallucination
创新:幻觉的几乎刻画
Nishant P. Das, Piyush Srivastava
机构
*
School of Technology and Computer Science, Tata Institute of Fundamental Research, Mumbai, Maharashtra - 400 005, India(技术与计算机科学学院,塔塔基础研究机构,孟买,马哈拉施特拉邦 - 400 005, 印度)
专题命中
知识编辑与模型理解
:LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
Mechanistic Interpretability of Antibody Language Models Using SAEs
使用 SAE 对抗体语言模型的机制可解释性研究
Rebonto Haque, Oliver M. Turnbull, Anisha Parsan, Nithin Parsan, John J. Yang, Anna L. Beukenhorst, Charlotte M. Deane
机构
*
Department of Statistics, University of Oxford, UK(英国牛津大学统计系)
;
Reticular, San Francisco, USA(美国旧金山Reticular公司)
;
EECS, MIT, Cambridge MA, USA(美国麻省理工学院电子工程与计算机科学系)
;
Leyden Laboratories BV, Leiden, The Netherlands(荷兰莱顿实验室)
Commentsv3: 15 pages; corrected author list and affiliations in the main text; minor text changes; updated steering results following minor code changes; conclusions and findings remain unchanged; included link to data and code in the Data Availability section
机构
*
McGill University(麦吉尔大学)
;
Mila - Quebec AI Institute(魁北克人工智能研究所)
;
University of Cambridge(剑桥大学)
;
MBZUAI - Mohamed bin Zayed University of Artificial Intelligence(MBZUAI - 摩苏尔·本·扎耶德人工智能大学)
;
University of Toronto(多伦多大学)
;
Salesforce
专题命中
知识编辑与模型理解
:LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
机构
*
School of Intelligence Science and Technology(智能科学与技术学院)
;
State Key Laboratory for Novel Software Technology(新型软件技术国家重点实验室)
;
Nanjing University(南京大学)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
Probing Cultural Awareness in LLMs: A Case Study of Cross-Culture Aesthetic Stylistics
探究大型语言模型的文化意识:跨文化美学文体学案例研究
Jiashuo Wang, Fenggang Yu, Jian Wang, Chak Tou Leong, Xiaoyu Shen, Chunpu Xu, Jiawen Duan, Wenjie Li, Johan F. Hoorn
机构
*
Department of Computing, Hong Kong Polytechnic University(香港理工大学计算机系)
;
Institute of Digital Twin, Eastern Institute of Technology, Ningbo(宁波东部技术研究所数字孪生研究所)
;
Department of Language Science and Technology, Hong Kong Polytechnic University(香港理工大学语言科学与技术系)
;
School of Design, Hong Kong Polytechnic University(香港理工大学设计学院)
;
Research Institute for Quantum Technology, Hong Kong Polytechnic University(香港理工大学量子技术研究所)
;
Department of Communication Science, Vrije Universiteit Amsterdam(阿姆斯特丹自由大学传播科学系)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.CL
机构
*
University of Electronic Science and Technology of China(电子科技大学)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)
;
Shandong University(山东大学)
;
Tongji University(同济大学)
;
City University of Hong Kong(香港城市大学)
BioFact-MoE: Biologically Factorized Mixture of Experts for Vision-Language Prognostic Modeling in Hepatocellular Carcinoma
BioFact-MoE:基于生物学因子分解的混合专家模型用于肝细胞癌的视觉-语言预后建模
Junlin Yang, Tian Yu, Nicha C. Dvornek, Yuexi Du, Peiyu Duan, Annabella Shewarega, Lawrence H. Staib, James S. Duncan, Julius Chapiro
机构
*
Department of Radiology \& Biomedical Imaging, Department of Biomedical Engineering, Department of Electrical Engineering, Department of Statistics \& Data Science Yale University, New Haven, CT, 06510, USA