Do Foundation Models See Biology? Evaluating Attention Coherence with Spatial Transcriptomics in Glioblastoma
基础模型是否理解生物学?利用空间转录组学评估胶质母细胞瘤中的注意力一致性
Dilakshan Srikanthan, Amoon Jamzad, Paul Wilson, Nooshin Maghsoodi, Robert Policelli, Gabor Fichtinger, John F. Rudan, Parvin Mousavi
机构
*
Translational Medicine, School of Medicine, Queen’s University, Kingston, ON, Canada(转化医学、医学院、皇后大学、金斯顿,ON,加拿大)
;
School of Computing, Queen’s University, Kingston, ON, Canada(计算学院、皇后大学、金斯顿,ON,加拿大)
;
Department of Surgery, Queen’s University, Kingston, ON, Canada(外科部门、皇后大学、金斯顿,ON,加拿大)
CommentsProceedings of the 6th Workshop on Trustworthy NLP (TrustNLP 2026), ACL 2026, San Diego, California, USA. Available at https://openreview.net/forum?id=WJCalficPT
AutoSpecNER: A Fine-Grained Named Entity Recognition Dataset for Vehicle Specification Extraction
AutoSpecNER: 用于车辆规格提取的细粒度命名实体识别数据集
Jordan Lee, Filippos Ventirozos, Abdirahman Abdullahm, Ioanna Nteka, Peter Appleby, Matthew Shardlow
机构
*
Department of Computing and Mathematics, Manchester Metropolitan University(曼彻斯特城市大学计算与数学系)
;
Autotrader Research Group, Autotrader UK(英国Autotrader研究组)
专题命中
评测与基准
:large language model(abstract);language model(abstract);分类 cs.CL
BenchX: Benchmarking AI Models for Cancer Detection and Localization with Demographic and Protocol Biases
BenchX: 基于人口统计和协议偏差的癌症检测与定位AI模型基准测试
Qi Chen, Wenxuan Li, Pedro R. A. S. Bassi, Xinze Zhou, Jakob Wasserthal, Ibrahim Ethem Hamamci, Sezgin Er, Ashwin Kumar, Yiwen Ye, Yuhan Wang, Yuyin Zhou, Akshay S. Chaudhari, Curtis Langlotz, Kang Wang, Yang Yang, Alan L. Yuille, Zongwei Zhou
机构
*
Johns Hopkins University(约翰霍普金斯大学)
;
German Cancer Research Center(德国癌症研究中心)
;
University Hospital Basel(巴塞尔大学医院)
;
University of Zurich(苏黎世大学)
;
ETH AI Center(苏黎世联邦理工学院AI中心)
;
Istanbul Medipol University(伊斯坦布尔梅迪波尔大学)
;
Stanford University(斯坦福大学)
;
École Polytechnique Fédérale de Lausanne(洛桑联邦理工学院)
;
Nanyang Technological University(南洋理工大学)
;
University of California, Santa Cruz(加州大学圣克鲁兹分校)
;
University of California, San Francisco(加州大学旧金山分校)
;
Johns Hopkins Medicine(约翰霍普金斯医学)
专题命中
评测与基准
:large language model(abstract);language model(abstract)
Resonant Minds: Closed-Loop Social Avatars with Theory of Mind
共鸣心智:具备心智理论的闭环社交虚拟人
Jianxu Shangguan, Jing Xu, Hang Ye, Xiaoxuan Ma, Yizhou Wang, Jenq-Neng Hwang, Wentao Zhu
机构
*
University of Washington(华盛顿大学)
;
Peking University(北京大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Eastern Institute of Technology, Ningbo(宁波工程技术学院)
专题命中
评测与基准
:large language model(abstract);language model(abstract)
FALCON: Transforming Cyber Threat Intelligence into Deployable IDS Rules with Self-Reflection
FALCON:通过自我反思将网络威胁情报转化为可部署的入侵检测系统规则
Shaswata Mitra, Subash Neupane, Martin Duclos, Sudip Mittal, Aritran Piplai, Md Rayhanur Rahman, Edward Zieglar, Shahram Rahimi
机构
*
Meharry Medical College(梅哈里医学院)
;
The University of Texas at El Paso(德克萨斯理工大学)
;
The University of Alabama(阿拉巴马大学)
;
National Security Agency(国家安全局)
机构
*
College of Computer Science and Electronic Engineering, Hunan University(湖南大学计算机科学与电子工程学院)
;
Department of Bioengineering and Imperial-X, Imperial College London(帝国理工学院伦敦校区生物工程系)
;
Department of Pathology, Xiangtan Maternal and Child Health Hospital(湘潭 maternal and child health hospital pathology department)
;
Department of Pathology, The First People’s Hospital of Xiangtan City(湘潭市第一人民医院病理科)
A Benchmark for Hallucination Detection in VLMs for Gastrointestinal Endoscopy
胃肠内窥镜中视觉语言模型幻觉检测的基准测试
Aminu Lawal, Niyoj Oli, Sachin Acharya, Prashnna Gyawali, Maria Carmen Romano, Binod Bhattarai
机构
*
University of Aberdeen(阿伯丁大学)
;
Nepal Applied Mathematics and Informatics Institute for Research(尼泊尔应用数学与信息学研究所)
;
West Virginia University(西弗吉尼亚大学)