Navigating Gigapixel Pathology Images with Large Multimodal Models
利用大型多模态模型导航千兆像素病理图像
Thomas A. Buckley, Kian R. Weihrauch, Katherine Latham, Andrew Z. Zhou, Padmini A. Manrai, Arjun K. Manrai
机构
*
Department of Biomedical Informatics, Harvard Medical School(哈佛医学院生物医学信息学系)
;
Department of Pathology, Massachusetts General Hospital(麻省总医院病理学系)
;
Department of Pathology and Laboratory Medicine, Brown University(布朗大学病理学与实验室医学系)
M4FC: a Multimodal, Multilingual, Multicultural, Multitask Real-World Fact-Checking Dataset
M4FC:一个多模态、多语言、多文化、多任务的真实世界事实验证数据集
Jiahui Geng, Jonathan Tonglet, Iryna Gurevych
机构
*
Mohamed bin Zayed University of Artificial Intelligence(Mohamed bin Zayed人工智能大学)
;
Ubiquitous Knowledge Processing Lab(ubiquitous知识处理实验室)
;
Department of Computer Science, TU Darmstadt(TU Darmstadt计算机科学系)
;
National Research Center for Applied Cybersecurity ATHENE(应用网络安全国家研究中心ATHENE)
;
Department of Electrical Engineering, KU Leuven(KU Leuven电气工程系)
;
Department of Computer Science, KU Leuven(KU Leuven计算机科学系)
Sustainability assessment using multimodal AI agents
使用多模态AI代理进行可持续性评估
Zhihan Zhang, Alexander Metzger, Yuxuan Mei, Felix Hähnlein, Zachary Englhardt, Tingyu Cheng, Gregory D. Abowd, Shwetak Patel, Adriana Schulz, Vikram Iyer
机构
*
Paul G. Allen School of Computer Science & Engineering, University of Washington(保罗·G·艾伦计算机科学与工程学院,华盛顿大学)
;
Computer Science and Engineering, University of Notre Dame(计算机科学与工程,诺丁汉大学)
;
Electrical and Computer Engineering, Northeastern University(电气与计算机工程,东北大学)
OGA-AID: Clinician-in-the-loop AI Report Drafting Assistant for Multimodal Observational Gait Analysis in Post-Stroke Rehabilitation
OGA-AID:用于中风后康复多模态观察性步态分析的临床医生在环AI报告起草助手
Khoi T. N. Nguyen, Nghia D. Nguyen, Hui Yu Koh, Patrick W. H. Kwong, Karen Sui Geok Chua, Ananda Sidarta, Baosheng Yu
机构
*
Rehabilitation Research Institute of Singapore, Nanyang Technological University, Singapore(新加坡康复研究中心,南洋理工大学,新加坡)
;
Lee Kong Chian School of Medicine, Nanyang Technological University, Singapore(李光前医学院,南洋理工大学,新加坡)
;
The Grainger College of Engineering, University of Illinois Urbana-Champaign, United States(伊利诺伊大学厄巴纳-香槟分校格雷格学院,美国)
;
Department of Rehabilitation Sciences, The Hong Kong Polytechnic University, Hong Kong(香港理工大学康复科学系,香港)
;
VinUni-Illinois Smart Health Center, VinUniversity, Vietnam(Vin大学Vin-伊利诺伊智能健康中心,越南)
;
Institute of Rehabilitation Excellence, Tan Tock Seng Hospital, NHG Health, Singapore(卓越康复研究所,坦托克桑格医院,NHG健康,新加坡)
RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension
RefBench-PRO:面向感知与推理的指称表达理解基准
Tianyi Gao, Hao Li, Han Fang, Xin Wei, Xiaodong Dong, Hongbo Sun, Ye Yuan, Zhongjiang He, Jinglin Xu, Jingmin Xin, Hao Sun
机构
*
National Key Laboratory of Human-Machine Hybrid Augmented Intelligence(国家人机混合增强智能重点实验室)
;
National Engineering Research Center for Visual Information and Applications(国家视觉信息与应用工程技术研究中心)
;
Institute of Artificial Intelligence and Robotics(人工智能与机器人研究院)
;
Xi’an Jiaotong University(西安交通大学)
;
Institute of Artificial Intelligence (TeleAI)(人工智能研究院(TeleAI))
;
China Telecom(中国电信)
;
Shanghai Jiao Tong University(上海交通大学)
;
University of Science and Technology Beijing(北京科技大学)
Neurai-VN Benchmark: Standardized Machine Learning Models for Multimodal Digital Phenotyping in Mental Health Classification
神经-VN基准:用于心理健康分类中多模态数字表型分析的标准化机器学习模型
Quoc-Cuong Pham, Hoang-Thuy-Duong Vu, Thi-Thanh-Huong Ha, Huy-Hieu Pham
机构
*
College of Engineering and Computer Science, VinUni-Illinois Smart Health Center, VinUniversity(工程与计算机科学学院,VinUni - 伊利诺伊智能健康中心,Vin大学)
;
School of Biomedical Engineering, International University, Vietnam National University HCMC(生物医学工程学院,胡志明市越南国立大学国际大学)
机构
*
National Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室)
;
School of Intelligence Science and Technology, Nanjing University(南京大学智能科学与技术学院)
;
AMAP, Alibaba Group(阿里集团AMAP)
;
School of Computing and Artificial Intelligence, Southwest Jiaotong University(西南交通大学计算机与人工智能学院)
Challenges and proposed solutions in modeling multimodal medical data: A systematic review
多模态医疗数据建模的挑战与解决方案:一项系统综述
Maryam Farhadizadeh, Maria Weymann, Michael Blaß, Johann Kraus, Christopher Gundler, Sebastian Walter, Noah Hempen, Hannah Bast, Harald Binder, Nadine Binder
机构
*
Institute of General Practice/Family Medicine(一般医学/家庭医学研究所)
;
Freiburg Center for Data Analysis, Modeling and AI(弗赖堡数据分析、建模与人工智能中心)
;
Institute of Medical Biometry and Statistics(医学生物统计研究所)
;
Institute for Applied Medical Informatics(应用医学信息研究所)
;
Institute of Medical Systems Biology(医学系统生物学研究所)
;
Department of Computer Science(计算机科学系)
Device-Cloud Collaborative LLM Inference with Multi-Modal, Multi-Task, Multi-Turn Conversations
具有多模态、多任务、多轮对话的设备-云协作大语言模型推理
Liangqi Yuan, Dong-Jun Han, Shiqiang Wang, Christopher G. Brinton
机构
*
School of Electrical and Computer Engineering, Purdue University(普渡大学电气与计算机工程学院)
;
Department of Computer Science and Engineering, Yonsei University(延世大学计算机科学与工程系)
;
Department of Computer Science, University of Exeter(埃克塞特大学计算机科学系)
Quantifying Perovskite Solar Cell Degradation via Machine Learning from Spatially Resolved Multimodal Luminescence Time Series
通过空间分辨多模态发光时间序列的机器学习量化钙钛矿太阳能电池退化
Giulio Barletta, Simon Ternes, Saif Ali, Zohair Abbas, Chiara Ostendi, Marialucia D'Addio, Erica Magliano, Pietro Asinari, Eliodoro Chiavazzo, Aldo Di Carlo
CommentsA subsequent extension of this analysis to a larger corpus found that the heterogeneity predictor is collinear with the industrial-versus-general domain split, so the corresponding effect is not identifiable on the available data
机构
*
Fudan University(复旦大学)
;
Institute of Artificial Intelligence (TeleAI), China Telecom(人工智能研究所(TeleAI),中国电信)
;
Tianjin University(天津大学)
;
Northwestern Polytechnical University(西北工业大学)
;
Tsinghua University(清华大学)
;
City University of Hong Kong(香港城市大学)
机构
*
MoE Key Lab of Artificial Intelligence, AI Institute, School of Computer Science, Shanghai Jiao Tong University(人工智能大模型重点实验室,人工智能学院,计算机科学学院,上海交通大学)
;
Qwen Team(通义实验室)
;
Beijing Institute of Technology(北京理工大学)
;
Tsinghua University(清华大学)
;
Zhejiang University(浙江大学)