Mechanistic Indicators of Understanding in Large Language Models
大语言模型中理解机制的指示器
Pierre Beckmann, Matthieu Queloz
机构
*
École Polytechnique Fédérale de Lausanne (EPFL)(联邦理工学院)
;
Idiap Research Institute(Idiap研究所)
;
University of Bern, Department of Philosophy(伯尔尼大学哲学系)
专题命中
知识编辑与模型理解
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI
Hallucination Filtering in Radiology Vision-Language Models Using Discrete Semantic Entropy
在放射学视觉-语言模型中使用离散语义熵过滤幻觉
Patrick Wienholt, Sophie Caselitz, Robert Siepmann, Philipp Bruners, Keno Bressem, Christiane Kuhl, Jakob Nikolas Kather, Sven Nebelung, Daniel Truhn
机构
*
Lab for Artificial Intelligence in Medicine, Department of Diagnostic and Interventional Radiology, University Hospital RWTH Aachen(医学人工智能实验室,诊断与介入放射学部,RWTH亚琛大学医院)
;
Department of Diagnostic and Interventional Radiology, University Hospital RWTH Aachen(诊断与介入放射学部,RWTH亚琛大学医院)
;
Department of Diagnostic and Interventional Radiology, Technical University of Munich, School of Medicine and Health, Klinikum rechts der Isar, TUM University Hospital(诊断与介入放射学部,慕尼黑技术大学,医学院与健康学院,Klinikum rechts der Isar,TUM大学医院)
;
Department of Cardiovascular Radiology and Nuclear Medicine, Technical University of Munich, School of Medicine and Health, German Heart Center, TUM University Hospital(心血管放射学与核医学部,慕尼黑技术大学,医学院与健康学院,德国心脏中心,TUM大学医院)
;
Else Kroener Fresenius Center for Digital Health, Faculty of Medicine and University Hospital Carl Gustav Carus, TUD Dresden University of Technology(数字健康中心,医学院与卡尔·古斯塔夫·卡尔斯大学医院,德累斯顿技术大学)
;
Department of Medicine I, Faculty of Medicine and University Hospital Carl Gustav Carus, TUD Dresden University of Technology(第一医学部,医学院与卡尔·古斯塔夫·卡尔斯大学医院,德累斯顿技术大学)
;
Pathology & Data Analytics, Leeds Institute of Medical Research at St James’s, University of Leeds(病理学与数据分析,圣詹姆斯医院医学研究所,利兹大学)
;
Medical Oncology, National Center for Tumor Diseases (NCT), University Hospital Heidelberg(医学肿瘤学,国家肿瘤疾病中心(NCT),海德堡大学医院)
MINAR: Mechanistic Interpretability for Neural Algorithmic Reasoning
MINAR: 图神经网络中神经算法推理的机制可解释性
Jesse He, Helen Jenne, Max Vargas, Davis Brown, Gal Mishne, Yusu Wang, Henry Kvinge
机构
*
Pacific Northwest National Laboratory, Richland, WA(太平洋西北国家实验室)
;
Halıcıoğlu Data Science Institute, University of California, San Diego, San Diego, CA(哈利奇奥格鲁数据科学研究所,加州大学圣地亚哥分校)
;
Department of Computer and Information Science, University of Pennsylvania, Pennsylvaina, PA(计算机与信息科学系,宾夕法尼亚大学)
;
Department of Mathematics, University of Washington, Seattle, WA(数学系,华盛顿大学)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
Probabilistic distances-based hallucination detection in LLMs with RAG
基于概率距离的LLM中幻觉检测方法(RAG)
Rodion Oblovatny, Alexandra Kuleshova, Konstantin Polev, Alexey Zaytsev
机构
*
Markov Lab, Department of Mathematics(马尔可夫实验室,数学系)
;
Computer Science, Saint-Petersburg University(计算机科学,圣彼得堡大学)
;
AI Center, Skoltech(人工智能中心,斯克里普丘克技术学院)
;
SB AI Lab(SB人工智能实验室)
;
AI Center, Skoltech, Risk department, Sber(人工智能中心,斯克里普丘克技术学院,风险部门)
专题命中
知识编辑与模型理解
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI
CommentsUpdated approach to constructing a hallucination detection score. Added results from experiments with the NLI task. The approach with trainable deep kernels has been removed, with a focus on the unsupervised approach
Learning What Matters: Prioritized Concept Learning via Relative Error-driven Sample Selection
学习关键要素:通过相对误差驱动的样本选择进行优先概念学习
Shivam Chandhok, Qian Yang, Oscar Manas, Kanishk Jain, Leonid Sigal, Aishwarya Agrawal
机构
*
Mila - Québec AI Institute(魁北克人工智能研究所)
;
University of British Columbia(不列颠哥伦比亚大学)
;
Université de Montréal(蒙特利尔大学)
;
Vector Institute for AI(人工智能矢量研究所)
;
Canada CIFAR AI Chair(加拿大CIFAR人工智能主席)
DynamicGTR: Leveraging Graph Topology Representation Preferences to Boost VLM Capabilities on Graph QAs
DynamicGTR: 利用图拓扑表示偏好提升视觉语言模型在图问答中的能力
Yanbin Wei, Jiangyue Yan, Chun Kang, Yang Chen, Hua Liu, James Kwok, Yu Zhang
机构
*
Southern University of Science and Technology(南方科技大学)
;
Hong Kong University of Science and Technology(香港理工大学)
;
Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))
;
Beihang University(北京航空航天大学)
机构
*
Vision and AI Lab, Indian Institute of Science, Bangalore, India(印度科学院视觉与人工智能实验室)
;
Chair for Machine Learning, University of Mannheim(曼海姆大学机器学习主任)
;
Max Planck Institute for Informatics, Saarland Informatics Campus, Saarbrücken, Germany(马克斯·普朗克信息研究所,萨尔兰州信息学院,萨尔布吕肯,德国)
SigVLP: Sigmoid Volume-Language Pre-Training for Self-Supervised CT-Volume Adaptive Representation Learning
SigVLP:基于sigmoid体积-语言预训练的自监督CT体积自适应表示学习
Jiayi Wang, Hadrien Reynaud, Ibrahim Ethem Hamamci, Sezgin Er, Suprosanna Shit, Bjoern Menze, Bernhard Kainz
机构
*
Friedrich-Alexander University Erlangen-Nürnberg(弗里德里希-亚历山大大学埃尔兰根-纽伦堡)
;
Department of Quantitative Biomedicine, University of Zurich(苏黎世大学定量生物医学系)
;
ETH AI Center, ETH Zurich(苏黎世联邦理工学院AI中心)
;
International School of Medicine, Istanbul Medipol University(伊斯坦布尔梅迪波尔大学国际医学院)
;
Department of Computing, Imperial College London(伦敦帝国学院计算机系)