Stochastic Siamese MAE Pretraining for Longitudinal Medical Images
随机时序Siamese MAE预训练用于纵向医学图像
Taha Emre, Arunava Chakravarty, Thomas Pinetz, Dmitrii Lachinov, Martin J. Menten, Hendrik Scholl, Sobha Sivaprasad, Daniel Rueckert, Andrew Lotery, Stefan Sacu, Ursula Schmidt-Erfurth, Hrvoje Bogunović
机构
*
Institute of Artificial Intelligence, Center for Medical Data Science, Medical University of Vienna(人工智能研究所,医学数据科学中心,维也纳医科大学)
;
Department of Ophthalmology and Optometry, Medical University of Vienna(眼科学与视光学系,维也纳医科大学)
;
Ophthalmic Image Analysis Group (OPTIMA), Medical University of Vienna(眼科影像分析组(OPTIMA),维也纳医科大学)
;
BioMedIA, Department of Computing, Imperial College London(BioMedIA,计算系,伦敦帝国理工学院)
;
Chair for AI in Healthcare and Medicine, Technical University of Munich(医学与健康人工智能教授职位,慕尼黑技术大学)
;
Moorfields National Institute for Health and Care Biomedical Research Centre, Moorfields Eye Hospital(莫尔菲尔兹国家健康与护理生物医学研究中心,莫尔菲尔兹眼科医院)
PI-MFM: Physics-informed multimodal foundation model for solving partial differential equations
PI-MFM:基于物理的多模态基础模型用于求解偏微分方程
Min Zhu, Jingmin Sun, Zecheng Zhang, Hayden Schaeffer, Lu Lu
机构
*
Department of Statistics and Data Science, Yale University(统计与数据科学系,耶鲁大学)
;
Department of Applied Mathematics and Statistics, Johns Hopkins University(应用数学与统计学系,约翰霍普金斯大学)
;
Department of Applied Computational Mathematics and Statistics, University of Notre Dame(应用计算数学与统计学系,圣母大学)
;
Department of Mathematics, University of California Los Angeles(数学系,加州大学洛杉矶分校)
;
Department of Chemical and Environmental Engineering, Yale University(化学与环境工程系,耶鲁大学)
Enhanced Spatiotemporal Consistency for Image-to-LiDAR Data Pretraining
增强的时空一致性用于图像到LiDAR数据预训练
Xiang Xu, Lingdong Kong, Hui Shuai, Wenwei Zhang, Liang Pan, Kai Chen, Ziwei Liu, Qingshan Liu
机构
*
College of Computer Science and Technology, Nanjing University of Aeronautics and Astronautics(南京航空航天大学计算机科学与技术学院)
;
School of Computing, Department of Computer Science, National University of Singapore(新加坡国立大学计算机学院)
;
School of Computer Science, Nanjing University of Posts and Telecommunications(南京邮电大学计算机学院)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
S-Lab, Nanyang Technological University(南洋理工大学S实验室)
Multi-Aspect Knowledge-Enhanced Medical Vision-Language Pretraining with Multi-Agent Data Generation
多方面知识增强的医学视觉-语言预训练与多代理数据生成
Xieji Li, Siyuan Yan, Yingsheng Liu, H. Peter Soyer, Monika Janda, Victoria Mar, Zongyuan Ge
机构
*
Department of Data Science and AI, Faculty of Information Technology, Monash University(数据科学与人工智能系,信息科技学院,墨尔本大学)
;
Victorian Melanoma Service, Alfred Health(维多利亚黑色素瘤服务,阿尔弗雷德健康)
;
Frazer Institute, The University of Queensland, Dermatology Research Centre(弗雷泽研究所,昆士兰大学,皮肤科研究中心)
Leveraging AI multimodal geospatial foundation models for improved near-real-time flood mapping at a global scale
利用AI多模态地理空间基础模型实现全球范围内的改进型实时洪水制图
Mirela G. Tulbure, Julio Caineta, Mark Broich, Mollie D. Gaines, Philippe Rufin, Leon-Friedrich Thomas, Hamed Alemohammad, Jan Hemmerling, Patrick Hostert
Discover, Learn, and Reinforce: Scaling Vision-Language-Action Pretraining with Diverse RL-Generated Trajectories
发现、学习与强化:通过多样化强化学习生成轨迹扩展视觉-语言-动作预训练
Rushuai Yang, Zhiyuan Feng, Tianxiang Zhang, Kaixin Wang, Chuheng Zhang, Li Zhao, Xiu Su, Yi Chen, Jiang Bian
机构
*
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
Tsinghua University(清华大学)
;
Wuhan University(武汉大学)
;
Central South University(中南大学)
;
Microsoft Research(微软研究院)
Synergizing Multigrid Algorithms with Vision Transformer: A Novel Approach to Enhance the Seismic Foundation Model
Huiwen Wu, Shuo Zhang, Yi Liu, Hongbin Ye
机构
*
Research Center for Scientific Data Hub Zhejiang Laboratory(科学数据枢纽研究中心 浙江实验室)
;
State Key Laboratory of Mathematical Sciences (SKLMS)(数学科学国家重点实验室)
;
State Key Laboratory of Scientific and Engineering Computing (LSEC)(科学与工程计算国家重点实验室)
;
Institute of Computational Mathematics and Scientific/Engineering Computing Academy of Mathematics and Systems Science, Chinese Academy of Sciences, Beijing, China(数学系统科学学院,中国科学院,北京,中国)
;
School of Mathematical Sciences, University of Chinese Academy of Sciences, Beijing, China(中国科学院大学数学科学学院,北京,中国)
机构
*
University of Southern California(南加州大学)
;
Florida State University(佛罗里达州立大学)
;
The Ohio State University(俄亥俄州立大学)
;
Nanyang Technological University(南洋理工大学)
LLM Teacher-Student Framework for Text Classification With No Manually Annotated Data: A Case Study in IPTC News Topic Classification
Taja Kuzman, Nikola Ljubešić
机构
*
Jožef Stefan International Postgraduate School(乔塞夫·斯塔芬国际研究生学校)
;
University of Ljubljana(卢布尔雅那大学)
专题命中
预训练与数据
:LLM(title);large language model(abstract);language model(abstract);分类 cs.CL
CommentsThis work has been accepted and published in the IEEE Access journal. This arXiv version is retained for archival purposes. Readers should use and cite the IEEE Access Version available at https://ieeexplore.ieee.org/document/10900365
机构
*
College of Computer Science and Technology, National University of Defense Technology, China(国防科技大学计算机科学与技术学院)
;
Tsinghua University, China(清华大学)
;
Beijing University of Posts and Telecommunications, China(北京邮电大学)
;
School of Computer Science, Wuhan University, China(武汉大学计算机学院)
;
Zhongguancun Academy, China(中关村学院)
Seeing, Signing, and Saying: A Vision-Language Model-Assisted Pipeline for Sign Language Data Acquisition and Curation from Social Media
Shakib Yazdani, Yasser Hamidullah, Cristina España-Bonet, Josef van Genabith
机构
*
German Research Center for Artificial Intelligence (DFKI GmbH)(德国人工智能研究中心(DFKI GmbH))
;
Saarland Informatics Campus(萨尔兰州信息技术校区)
;
Barcelona Supercomputing Center (BSC-CNS)(巴塞罗那超级计算中心(BSC-CNS))
Language Model Behavioral Phases are Consistent Across Architecture, Training Data, and Scale
James A. Michaelov, Roger P. Levy, Benjamin K. Bergen
机构
*
Department of Brain and Cognitive Sciences, MIT(麻省理工学院脑科学与认知科学系)
;
MIT Libraries CREOS(麻省理工学院图书馆 CREOS)
;
Deparmtent of Cognitive Science, UCSD(加州大学圣地亚哥分校认知科学系)
From TOWER to SPIRE: Adding the Speech Modality to a Translation-Specialist LLM
Kshitij Ambilduke, Ben Peters, Sonal Sannigrahi, Anil Keshwani, Tsz Kin Lam, Bruno Martins, André F. T. Martins, Marcely Zanon Boito
机构
*
ENS Paris-Saclay(巴黎-萨克雷大学)
;
INESC-ID
;
Instituto de Telecomunicações(电信研究所)
;
Instituto Superior Técnico, Universidade de Lisboa(里斯本大学技术学院)
;
Sapienza University of Rome(罗马萨皮恩扎大学)
;
University of Edinburgh(爱丁堡大学)
;
TransPerfect
;
NAVER LABS Europe(NAVER欧洲实验室)
机构
*
Xiamen University Malaysia(马来西亚厦门大学)
;
Columbia University(哥伦比亚大学)
;
The Hong Kong Polytechnic University(香港理工大学)
;
Xiamen University(厦门大学)
;
Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))