Ijaz Ul Haq, Byung Suk Lee, Julia N. Perdrial, David Baude
机构
*
Department of Computer Science, University of Vermont(维珍尼亚大学计算机科学系)
;
Water Resources Institute, University of Vermont(维珍尼亚大学水文资源研究所)
;
Department of Geography and Geosciences, University of Vermont(维珍尼亚大学地理与地质学系)
CommentsSupplementary materials, datasets, and implementation code will be made publicly available upon acceptance for publication in a peer-reviewed journal
CommentsAccepted at ICAART 2026 (18th International Conference on Agents and Artificial Intelligence). The final published version is available in the conference proceedings (SCITEPRESS)
Journal refIn Proceedings of the 18th International Conference on Agents and Artificial Intelligence (ICAART 2026), Vol. 1, pp. 339-346. SCITEPRESS, 2026
3D Modality-Aware Pre-training for Vision-Language Model in MRI Multi-organ Abnormality Detection
面向MRI多器官异常检测的3D模态感知预训练
Haowen Zhu, Ning Yin, Xiaogen Zhou
机构
*
School of Electronic, Electrical Engineering and Physics, Fujian University of Technology(福建工程学院电子电气工程学院)
;
School of Computer Science and Engineering, Southeast University, China(东南大学计算机科学与工程学院)
;
Department of Medical Imaging, Suzhou Traditional Chinese Medicine Hospital, China(苏州中医医院影像科)
Should We Still Pretrain Encoders with Masked Language Modeling?
我们还应该用掩码语言模型预训练编码器吗?
Hippolyte Gisserot-Boukhlef, Nicolas Boizard, Manuel Faysse, Duarte M. Alves, Emmanuel Malherbe, André F. T. Martins, Céline Hudelot, Pierre Colombo
机构
*
Artefact Research Center(Artefact 研究中心)
;
Diabolocom
;
TransPerfect
;
Cohere
;
MICS, CentraleSupélec, Université Paris-Saclay(MICS,CentraleSupélec,巴黎萨克雷大学)
;
Instituto de Telecomunicações(电信研究所)
;
Instituto Superior Técnico & Universidade de Lisboa (Lisbon ELLIS Unit)(里斯本大学(里斯本 ELLIS 单位))
Federated EndoViT: Pretraining Vision Transformers via Federated Learning on Endoscopic Image Collections
联邦端oscopeViT:通过联邦学习在内窥镜图像集上预训练视觉Transformer
Max Kirchner, Alexander C. Jenke, Sebastian Bodenstedt, Fiona R. Kolbinger, Oliver L. Saldanha, Jakob N. Kather, Martin Wagner, Stefanie Speidel
机构
*
National Center for Tumor Diseases (NCT)(国家肿瘤中心)
;
Translational Surgical Oncology(转化外科肿瘤学)
;
Faculty of Medicine and University Hospital Carl Gustav Carus, TUD Dresden University of Technology(医学系和图林根大学技术大学卡尔·古斯塔夫·卡尔医院)
;
DKFZ, Faculty of Medicine and University Hospital Carl Gustav Carus, TUD Dresden University of Technology(德累斯顿大学技术大学医学系和德累斯顿癌症研究中心)
;
Helmholtz-Zentrum Dresden-Rossendorf (HZDR)(德累斯顿-罗斯多夫亥姆霍兹中心)
;
Centre for Tactile Internet with Human-in-the-Loop (CeTI), TU Dresden(具有人类在环的触觉互联网中心,德累斯顿技术大学)
;
Visceral, Thoracic and Vascular Surgery, Faculty of Medicine and University Hospital Carl Gustav Carus, TU Dresden(visceral、胸外科和血管外科,医学系和图林根大学技术大学卡尔·古斯塔夫·卡尔医院)
;
Weldon School of Biomedical Engineering, Purdue University(生物医学工程韦尔登学校,普渡大学)
;
Medical Oncology, NCT, University Hospital Heidelberg, Germany(医学肿瘤学,国家肿瘤中心,海德堡大学医院,德国)
;
Medicine I, Faculty of Medicine and University Hospital Carl Gustav Carus, TU Dresden(医学I,医学系和图林根大学技术大学卡尔·古斯塔夫·卡尔医院)
Modalities, a PyTorch-native Framework For Large-scale LLM Training and Research
模态,一种用于大规模大语言模型训练和研究的PyTorch原生框架
Max Lübbering, Timm Ruland, Richard Rutmann, Felix Stollenwerk, David Fitzek, Michael Fromm, Alexander Weber, Rafet Sifa, Nicolas Flores-Herr, Joachim Köhler, Mehdi Ali
机构
*
Fraunhofer IAIS(弗劳恩霍夫智能系统研究所)
;
AI Sweden(人工智能瑞典)
;
University of Bonn(波恩大学)
;
Lamarr Institute(拉马尔研究所)
Unlocking Noisy Real-World Corpora for Foundation Model Pre-Training via Quality-Aware Tokenization
通过质量感知的分词解锁噪声真实世界语料用于基础模型预训练
Arvid E. Gollwitzer, Paridhi Latawa, David de Gruijl, Deepak A. Subramanian, Adrián Noriega de la Colina
机构
*
Broad Institute of MIT and Harvard, Cambridge, MA, USA(麻省理工学院与哈佛大学Broad研究所)
;
Massachusetts Institute of Technology, Cambridge, MA, USA(麻省理工学院)
;
Koch Institute for Integrative Cancer Research, MIT, Cambridge, MA, USA(麻省理工学院Koch整合癌症研究 institute)
;
Department of Neurology and Neurosurgery, McGill University, Montreal, Canada(麦吉尔大学神经学与神经外科系)
;
The Montreal Neurological Hospital-Institute, Montreal, Canada(蒙特利尔神经科学医院-研究所)
RobustDebias: Debiasing Language Models using Distributionally Robust Optimization
RobustDebias: 使用分布鲁棒优化去偏语言模型
Deep Gandhi, Katyani Singh, Nidhi Hegde
机构
*
Deep Gandhi University of Alberta(Deep Gandhi 阿尔伯塔大学)
;
Katyani Singh University of Alberta(Katyani Singh 阿尔伯塔大学)
;
Nidhi Hegde University of Alberta(Nidhi Hegde 阿尔伯塔大学)
;
Alberta Machine Intelligence Institute(阿尔伯塔人工智能研究所)
机构
*
Department of Electrical Engineering(电气工程系)
;
National Taiwan University(国立台湾大学)
;
Department of Computer Science and Information Engineering(计算机科学与信息工程系)
;
Institute of Information Science(信息科学研究所)
专题命中
预训练与数据
:LLM(title);large language model(abstract);language model(abstract);分类 cs.AI
Understanding the Transfer Limits of Vision Foundation Models
理解视觉基础模型的迁移限制
Shiqi Huang, Yipei Wang, Natasha Thorley, Alexander Ng, Shaheer Saeed, Mark Emberton, Shonit Punwani, Veeru Kasivisvanathan, Dean Barratt, Daniel Alexander, Yipeng Hu
机构
*
Harbin Institute of Technology, Shenzhen, China(哈尔滨工业大学(深圳))
;
Hong Kong Baptist University, China(香港 Baptist大学)
;
City University of Hong Kong, China(香港城市大学)
Training Language Models with homotokens Leads to Delayed Overfitting
通过homotokens训练语言模型导致过拟合延迟
Adrian Cosma, Stefan Ruseti, Emilian Radoi, Mihai Dascalu
机构
*
Dalle Molle Institute for Artificial Intelligence (IDSIA)(达勒莫莱人工智能研究所)
;
National University of Science and Technology POLITEHNICA Bucharest(科学与技术国家大学)