机构
*
National Observatory of Athens(国家天文台)
;
National Technical University of Athens(雅典国家技术大学)
;
National and Kapodistrian University of Athens(雅典国家与卡波迪斯特里亚大学)
;
Athena Research Center(雅典研究所以及研究中心)
Pretraining Induces a Reusable Spectral Basis for Downstream Task Adaptation
预训练诱导了可重用的谱基底以用于下游任务适应
Junjie Yu, Yue Wang, Zihan Deng, Yan Zhu, Wenxiao Ma, Quanying Liu
机构
*
Department of Biomedical Engineering, Southern University of Science and Technology(生物医学工程系,南方科技大学)
;
Department of Psychology, The University of Hong Kong(心理学系,香港大学)
Location-Aware Pretraining for Medical Difference Visual Question Answering
面向位置的预训练方法用于医学差异视觉问答
Denis Musinguzi, Caren Han, Prasenjit Mitra
机构
*
Department of Electrical and Computer Engineering, Carnegie Mellon University, Kigali, Rwanda(电气与计算机工程系,卡内基梅隆大学,刚果(金)基利齐)
;
University of Melbourne, Melbourne, Australia(墨尔本大学,墨尔本,澳大利亚)
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学人工智能学院)
;
Beijing Key Laboratory of Research on Large Models and Intelligent Governance(北京大型模型与智能治理研究重点实验室)
;
Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)
;
Institute for AI Industry Research, Tsinghua University(清华大学人工智能产业研究院)
;
Advanced Computing and Storage Lab, Huawei Technologies(华为技术有限公司先进计算与存储实验室)
;
School of Intelligence Science and Technology, Nanjing University(南京大学智能科学与技术学院)
;
Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(教育部下一代智能搜索与推荐工程研究中心)
GeoMeld: Toward Semantically Grounded Foundation Models for Remote Sensing
GeoMeld:迈向遥感领域语义引导的基模模型
Maram Hasan, Md Aminur Hossain, Savitra Roy, Souparna Bhowmik, Ayush V. Patel, Mainak Singha, Subhasis Chaudhuri, Muhammad Haris Khan, Biplab Banerjee
机构
*
Indian Institute of Technology Bombay(印度理工学院孟买分校)
;
Space Applications Centre, ISRO(印度空间研究组织空间应用中心)
;
University of Trento(特伦托大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
CliPPER: Contextual Video-Language Pretraining on Long-form Intraoperative Surgical Procedures for Event Recognition
CliPPER:基于长形式术中手术程序的上下文视频-语言预训练用于事件识别
Florian Stilz, Vinkle Srivastav, Nassir Navab, Nicolas Padoy
机构
*
University of Strasbourg, CNRS, INSERM, ICube, UMR7357, France(斯特拉斯堡大学,法国国家科学研究中心,法国国家医学研究院,ICube,UMR7357,法国)
;
IHU Strasbourg, France(斯特拉斯堡IHU,法国)
;
Technical University of Munich, Germany(慕尼黑技术大学,德国)
From Panel to Pixel: Zoom-In Vision-Language Pretraining from Biomedical Scientific Literature
从面板到像素:从生物医学科学文献中进行缩放视觉-语言预训练
Kun Yuan, Min Woo Sun, Zhen Chen, Alejandro Lozano, Xiangteng He, Shi Li, Nassir Navab, Xiaoxiao Sun, Nicolas Padoy, Serena Yeung-Levy
机构
*
University of Strasbourg, CNRS, INSERM, ICube, UMR7357, Strasbourg, France(斯特拉斯堡大学、法国国家科学研究中心、法国国家医学研究院、ICube、UMR7357、斯特拉斯堡,法国)
;
Technical University of Munich(慕尼黑技术大学)
;
Stanford University(斯坦福大学)
;
DSAI, The Hong Kong Polytechnic University(香港理工大学DSAI实验室)
;
University of British Columbia(不列颠哥伦比亚大学)
;
Munich Center for Machine Learning(慕尼黑机器学习中心)
;
IHU Strasbourg, Strasbourg(斯特拉斯堡IHU医院,斯特拉斯堡)
;
Vector Institute for AI(人工智能向量研究所)
;
Yale University(耶鲁大学)
DeeperBrain: A Neuro-Grounded EEG Foundation Model Towards Universal BCI
DeeperBrain: 一种面向通用脑机接口的神经 grounded EEG 基础模型
Jiquan Wang, Sha Zhao, Yangxuan Zhou, Yiming Kang, Shijian Li, Gang Pan
机构
*
State Key Laboratory of Brain-machine Intelligence, Zhejiang University(浙江大学脑机智能国家重点实验室)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
MOE Frontier Science Center for Brain Science and Brain-machine Integration(教育部脑科学与脑机集成前沿科学中心)