机构
*
Taizhou Institute of Science and Technology, Nanjing University of Science and Technology(泰州科技学院、南京理工大学)
;
Department of Intelligence Science, Xi’an Jiaotong-Liverpool University(智能科学系,西安交通大学利物浦大学)
;
School of Computer Science and Technology, Soochow University(计算机科学与技术学院,苏州大学)
;
Department of Statistical Sciences, University of Toronto(统计科学系,多伦多大学)
Semantics-Guided Multimodal Masked Autoencoder Pretraining for 3D BEV Object Detection
语义引导的多模态掩码自编码器预训练用于3D BEV目标检测
Prabuddhi Wariyapperuma, Rajitha de Silva, Marc Hanheide, Thomas Bohné, Leonardo Guevara
机构
*
University of Lincoln, Lincoln Centre for Autonomous Systems(林肯大学,林肯自主系统中心)
;
University of Cambridge, Institute for Manufacturing, Department of Engineering(剑桥大学,制造研究所,工程系)
机构
*
Key Laboratory of Child Development and Learning Science (Ministry of Education), School of Biological Sciences and Medical Engineering, Southeast University(儿童发展与学习科学重点实验室(教育部)、生物科学与医学工程学院、东南大学)
;
Department of Artificial Intelligence, Westlake University(人工智能学院、西湖大学)
;
Department of Artificial Intelligence, Vrije Universiteit Amsterdam(人工智能学院、阿姆斯特丹自由大学)
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
;
Baidu Inc(百度公司)
A Multimodal Framework for Dementia Detection via Linguistic and Acoustic Representation Learning
基于语言和声学表征学习的多模态痴呆检测框架
Loukas Ilias, Dimitris Askounis
机构
*
Decision Support Systems Laboratory, School of Electrical and Computer Engineering, National Technical University of Athens(决策支持系统实验室,电气与计算机工程学院,国家技术大学雅典)
机构
*
Department of Data Science \& AI, Faculty of Information Technology, Monash University, Melbourne, VIC 3800, Australia Alfred Health Radiology, Alfred Health, Melbourne, VIC 3004, Australia School of Translational Medicine, Faculty of Medicine, Nursing
;
Health Sciences, Monash University, Melbourne, VIC 3800, Australia Hong Kong Polytechnic University, Hong Kong SAR, China
CommentsEarly accepted by MICCAI 2026. This version of the contribution has been accepted for publication, after peer review (when applicable) but is not the Version of Record and does not reflect post-acceptance improvements, or any corrections
Image-Conditioned Instance Prompt Network for Referring Remote Sensing Image Segmentation
图像条件实例提示网络用于遥感图像指代分割
Biaoyu Ren, Qingsheng Wang, Cun Xu, Dingkang Yang, Wenxuan Wang
机构
*
School of Computer Science, Northwestern Polytechnical University, Xi'an, China(西北工业大学计算机科学学院,西安,中国)
;
College of Intelligent Robotics and Advanced Manufacturing, Fudan University, Shanghai, China(复旦大学智能机器人与先进制造学院,上海,中国)
;
Shenzhen Research Institute of Northwestern Polytechnical University, Shenzhen, China(西北工业大学深圳研究院,深圳,中国)
IVR-R1: Refining Trajectories through Iterative Visual-Grounded Reasoning in Reinforcement Learning
IVR-R1:通过强化学习中的迭代视觉基础推理优化轨迹
Chenghao Li, Fusheng Hao, Xikai Zhang, Likang Xiao, Yanwei Ren, Fuxiang Wu, Quan Chen, Liu Liu
机构
*
Hangzhou International Innovation Institute, Beihang University(北京航空航天大学杭州国际创新研究院)
;
School of Artificial Intelligence, Beihang University(北京航空航天大学人工智能学院)
;
Kuaishou Technology(快手科技)
;
Shenzhen Institute of Advanced Integration Technology, Shenzhen(深圳先进集成技术研究院)
机构
*
University of Macau(澳门大学)
;
Guangdong Institute of Intelligence Science and Technology(广东智能科学与技术研究院)
;
Peking University(北京大学)
;
Independent Researcher(独立研究员)
;
Institute of Science Tokyo(东京科学研究院)
;
Morgan Stanley(摩根大通)
;
Halmstad University(哈马碧大学)