机构
*
School of Computer Science and Technology, Beijing Jiaotong University(北京交通大学计算机科学与技术学院)
;
School of Computer Science, Peking University(北京大学计算机科学学院)
;
Huazhong University of Science and Technology(华中科技大学)
;
Department of Computer Science and Information Technology, La Trobe University(拉筹伯大学计算机科学与信息技术系)
;
UCAS-Terminus AI Lab, University of Chinese Academy of Sciences(中国科学院大学UCAS-Terminus人工智能实验室)
DynFly: Dynamic-Aware Continuous Trajectory Generation for UAV Vision-Language Navigation in Urban Environments
DynFly: 面向城市环境无人机视觉语言导航的动态感知连续轨迹生成
Wen Jiang, Hanfang Liang, Li Wang, Kangyao Huang, Wang Xu, Wei Fan, Jinyuan Liu, Shaoyu Liu, Hongwei Duan, Bin Xu, Xiangyang Ji, Huaping Liu
机构
*
Beijing Institute of Technology(北京理工大学)
;
Tsinghua University(清华大学)
;
Dalian University of Technology(大连理工大学)
;
Xidian University(西安电子科技大学)
;
Huazhong University of Science and Technology(华中科技大学)
机构
*
MoE Key Lab of Artificial Intelligence(人工智能MOE实验室)
;
AI Institute(人工智能研究院)
;
School of Computer Science(计算机科学学院)
;
Shanghai Jiao Tong University(上海交通大学)
;
Department of Radiology(放射科)
;
The First Affiliated Hospital(第一附属医院)
;
School of Medicine(医学院)
;
Zhejiang University(浙江大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
State Key Laboratory of Infrared Physics(红外物理国家重点实验室)
;
Shanghai Institute of Technical Physics(上海技术物理研究所)
;
Chinese Academy of Science(中国科学院)
AMALIA-VL: A Native European Portuguese Open-Source Vision and Language Model
AMALIA-VL: 一个原生欧洲葡萄牙语开源视觉与语言模型
Diogo Glória-Silva, João Cardeira, Manuel Letras da Luz, Afonso Simplício, Gonçalo Vinagre, Diogo Tavares, Rafael Ferreira, Inês Calvo, Inês Vieira, David Semedo, João Magalhães
机构
*
NOVA School of Science and Technology(NOVA科学与技术学校)
;
NOVA LINCS
机构
*
Seoul National University(首尔大学)
;
Robotics Lab, Hyundai Motor Company(现代汽车公司机器人实验室)
;
Pohang University of Science and Technology (POSTECH)(浦项科技大学)
Rethinking the Role of Feature Engineering and Learning Strategies in Few-Shot Hidden Emotion Recognition
重新思考特征工程与学习策略在少样本隐藏情感识别中的作用
Xiaochuan Guo, Jihao Gu, Haixu Liu, Yuxin Liu, Qi Wang, Yufei Wang, Fei Wang, Kun Li, Dan Guo
机构
*
Hefei University of Technology(合肥工业大学)
;
University College London(伦敦大学学院)
;
The University of Sydney(悉尼大学)
;
Beijing QBoson Quantum Technology Co., Ltd.(北京量子芯光科技有限公司)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合性国家科学中心人工智能研究院)
;
Beihang University(北京航空航天大学)
;
The University of New South Wales(新南威尔士大学)
;
United Arab Emirates University(阿拉伯联合酋长国大学)
HSD: Training-Free Acceleration for Document Parsing Vision-Language Models with Hierarchical Speculative Decoding
HSD:基于分层推测解码的文档解析视觉语言模型训练免费加速
Wenhui Liao, Hongliang Li, Pengyu Xie, Xinyu Cai, Yufan Shen, Yi Xin, Qi Qin, Shenglong Ye, Tianbin Li, Ming Hu, Junjun He, Yihao Liu, Wenhai Wang, Min Dou, Bin Fu, Botian Shi, Yu Qiao, Lianwen Jin
机构
*
South China University of Technology(华南理工大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Shenzhen Institute of Advanced Technology, CAS(中国科学院深圳先进技术研究院)
;
Nanjing University(南京大学)
CheXanatomy: Anatomy-Aware Vision-Language Modeling for Chest Radiographs
CheXanatomy: 面向胸部X光片的解剖感知视觉-语言建模
Sergios Gatidis, Curtis Langlotz, Christian Bluethgen
机构
*
Stanford Center for Artificial Intelligence in Medicine and Imaging, Stanford University(斯坦福大学医学与影像人工智能中心)
;
Department of Radiology, Stanford University(斯坦福大学放射学系)
机构
*
School of Biomedical Engineering, Division of Life Sciences and Medicine, University of Science and Technology of China (USTC)(生物医学工程学院,生命科学与医学系,中国科学技术大学)
;
Center for Medical Imaging, Robotics, Analytic Computing & Learning (MIRACLE)(医学影像、机器人、分析计算与学习中心)
;
Suzhou Institute for Advanced Research, USTC(苏州先进研究院,中国科学技术大学)
;
Department of Radiology, The First Affiliated Hospital of USTC, Division of Life Sciences and Medicine, USTC(放射科,中国科学技术大学第一附属医院,生命科学与医学系,中国科学技术大学)
;
T Magnetic Resonance Translational Medicine Research Center, Department of Radiology, The First Affiliated Hospital (Southwest Hospital) of Army Medical University(7T磁共振转化医学研究中心,放射科,中国医学大学第一附属医院(西南医院))
6 Fingers, 1 Kidney: Natural Adversarial Medical Images Reveal Critical Weaknesses of Vision-Language Models
6根手指,1个肾脏:自然对抗性医学图像揭示视觉语言模型的关键弱点
Leon Mayer, Piotr Kalinowski, Caroline Ebersbach, Marcel Knopp, Tim Rädsch, Evangelia Christodoulou, Annika Reinke, Fiona R. Kolbinger, Lena Maier-Hein
机构
*
German Cancer Research Center (DKFZ) Heidelberg, Division of Intelligent Medical Systems(德国癌症研究中心(DKFZ)海德堡,智能医学系统部门)
;
Medical Faculty, Heidelberg University(海德堡大学医学院)
;
Faculty of Mathematics and Computer Science, Heidelberg University(海德堡大学数学与计算机科学学院)
;
HIDSS4Health - Helmholtz Information and Data Science School for Health, Karlsruhe/Heidelberg(HIDSS4Health - 哈勃-马克斯信息与数据科学健康学院,卡尔斯鲁厄/海德堡)
;
Helmholtz Imaging, German Cancer Research Center (DKFZ)(哈勃-马克斯成像,德国癌症研究中心(DKFZ))
;
Engineering Faculty, Heidelberg University(海德堡大学工程学院)
;
School of Computation, Information and Technology, TUM(技术大学(TUM)计算、信息与技术学院)
;
Weldon School of Biomedical Engineering, Purdue University(普渡大学韦尔登生物医学工程学院)
;
Department of Visceral, Thoracic and Vascular Surgery, University Hospital and Faculty of Medicine Carl Gustav Carus, TUD Dresden University of Technology(visceral、胸腔和血管外科部门,技术大学(TUD)德累斯顿大学医院和医学院)
;
National Center for Tumor Diseases (NCT), NCT Heidelberg, a partnership between DKFZ and University Hospital Heidelberg(肿瘤疾病国家中心(NCT),海德堡NCT,DKFZ与海德堡大学医院之间的合作)
;
Heidelberg University Hospital, Surgical Clinic, Surgical AI Research Group(海德堡大学医院,外科诊所,外科人工智能研究组)
;
Mohamed Bin Zayed University of Artificial Intelligence (MBZUAI), Abu Dhabi, UAE(Mohamed Bin Zayed人工智能大学(MBZUAI),阿布扎赫,阿拉伯联合酋长国)
Jolia: Concept-Level Vision-Language Alignment for 3D CT Contrastive Learning
Jolia: 用于3D CT对比学习的概念级视觉-语言对齐
Julien Khlaut, Charles Corbière, Baptiste Callard, Amaury Prat, Leo Butsanets, Antoine Saporta, Théo Danielou, Leo Machado, Korentin Le Floch, Tom Boeken, Pierre Manceron, Corentin Dancette
机构
*
Raidium
;
Department of Vascular and Oncological Interventional Radiology, Hôpital Européen Georges Pompidou, AP-HP(欧洲乔治·蓬皮杜医院血管与肿瘤介入放射科,AP-HP)
;
Faculté de Santé, Université Paris-Cité(巴黎西岱大学健康学院)
;
HEKA, INRIA(HEKA,法国国家信息与自动化研究所)
;
Imaging Department, Fondation Ophtalmologique Adolphe de Rothschild(阿道夫·罗斯柴尔德眼科基金会影像科)
机构
*
University of Chinese Academy of Sciences(中国科学院大学)
;
State Key Lab of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences, Beijing, China(中国科学院人工智能安全国家重点实验室,计算技术研究所,北京,中国)
;
Harbin Institute of Technology (Weihai)(哈尔滨工业大学(威海))
HANCLIP: A Family of Hyperbolic Angular Negation Vision Language Models
HANCLIP:双曲角否定视觉语言模型系列
Hoang-Bao Le, Aiden Durrant, Thai Son Mai, Binh T. Nguyen, Liting Zhou, Cathal Gurrin
机构
*
ADAPT Centre Dublin City University, Ireland(爱尔兰都柏林城市大学ADAPT中心)
;
University of East Anglia Norwich, UK(英国东英吉利大学)
;
Queen’s University Belfast Belfast, UK(英国贝尔法斯特女王大学)
;
University of Science Vietnam National University Ho Chi Minh City, Vietnam(越南胡志明市国家大学理科大学)