机构
*
Indian Institute of Technology, Bombay(印度理工学院班加罗尔分校)
;
Indian Institute of Technology, Roorkee(印度理工学院罗尔基分校)
;
Microsoft Research(微软研究院)
;
Stanford University(斯坦福大学)
;
Adobe MDSR(Adobe MDSR实验室)
;
Carnegie Mellon University(卡内基梅隆大学)
Multimodal Coherent Explanation Generation of Robot Failures
Pradip Pramanick, Silvia Rossi
机构
*
Interdepartmental Center for Advances in Robotic Surgery - ICAROS, University of Naples Federico II(跨部门先进机器人手术中心 - ICAROS,那不勒斯费德里科二世大学)
;
Department of Electrical Engineering and Information Technologies - DIETI, University of Naples Federico II(电气工程与信息科技系 - DIETI,那不勒斯费德里科二世大学)
专题命中
多模态生成
:multimodal(title,abstract);分类 cs.AI
Journal ref2024 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
MIFNet: Learning Modality-Invariant Features for Generalizable Multimodal Image Matching
Yepeng Liu, Zhichao Sun, Baosheng Yu, Yitian Zhao, Bo Du, Yongchao Xu, Jun Cheng
机构
*
National Engineering Research Center for Multimedia Software(多媒体软件国家工程研究中心)
;
Institute of Artificial Intelligence(人工智能研究院)
;
School of Computer Science(计算机科学学院)
;
Hubei Key Laboratory of Multimedia and Network Communication Engineering(湖北省多媒体与网络通信工程重点实验室)
;
Lee Kong Chian School of Medicine(李光耀医学院)
;
Nanyang Technological University(南洋理工大学)
;
Ningbo Institute of Materials Technology and Engineering(宁波材料技术与工程研究所)
;
Chinese Academy of Sciences(中国科学院)
;
Institute for Infocomm Research (I 2 R)(信息与通信研究所以(I 2 R))
;
Agency for Science, Technology and Research (A*STAR)(科技研究局(A*STAR))
COSMMIC: Comment-Sensitive Multimodal Multilingual Indian Corpus for Summarization and Headline Generation
Raghvendra Kumar, S. A. Mohammed Salman, Aryan Sahu, Tridib Nandi, Pragathi Y. P., Sriparna Saha, Jose G. Moreno
机构
*
Department of Computer Science and Engineering, Indian Institute of Technology Patna, India(印度理工学院帕纳分校计算机科学与工程系)
;
Department of Metallurgical and Materials Engineering, National Institute of Technology Tiruchirappalli, India(印度理工学院 Tiruchirappalli 金属与材料工程系)
;
Department of Computer Science and Information Systems, BITS Pilani – Goa Campus, India(比斯·帕尼学院 Goa 分校计算机科学与信息系统系)
;
Department of Computer Science and Engineering, Indian Institute of Information Technology Vadodara, India(印度信息科技学院瓦达拉分校计算机科学与工程系)
;
Department of Computer Science and Engineering, B.M.S. College of Engineering, Bangalore, India(班加罗尔 B.M.S. 工程学院计算机科学与工程系)
;
Université de Toulouse, IRIT UMR 5505 CNRS, France(图卢兹大学 IRIT UMR 5505 CNRS 实验室)
BrainMAP: Multimodal Graph Learning For Efficient Brain Disease Localization
Nguyen Linh Dan Le, Jing Ren, Ciyuan Peng, Chengyao Xie, Bowen Li, Feng Xia
机构
*
School of Computing Technologies, RMIT University(计算技术学院,拉筹纳斯大学)
;
Institute of Innovation, Science and Sustainability, Federation University Australia(创新、科学与可持续性研究所,联邦大学澳大利亚)
seg2med: a bridge from artificial anatomy to multimodal medical images
Zeyu Yang, Zhilin Chen, Yipeng Sun, Anika Strittmatter, Anish Raj, Ahmad Allababidi, Johann S. Rink, Frank G. Zöllner
机构
*
Computer Assisted Clinical Medicine, Medical Faculty Mannheim, Heidelberg University(计算机辅助临床医学,曼海姆医学院,海德堡大学)
;
Pattern Recognition Lab, Friedrich-Alexander-University Erlangen-Nuremberg(模式识别实验室,埃尔兰根-纽伦堡弗里德里希-亚历山大大学)
;
Department of Radiology and Nuclear Medicine, University Medical Center Mannheim(放射学与核医学系,曼海姆大学医学中心)
;
Mannheim Institute for Intelligent Systems in Medicine, Medical Faculty Mannheim, Heidelberg University(曼海姆智能医学研究所,曼海姆医学院,海德堡大学)
;
Optical Bioimaging Laboratory, Department of Biomedical Engineering, College of Design and Engineering, National University of Singapore(光学生物成像实验室,生物医学工程系,设计与工程学院,新加坡国立大学)
机构
*
Faculty of Computer Science and Technology, Qilu University of Technology (Shandong Academy of Sciences)(计算机科学与技术学院,齐鲁工业大学(山东科学院))
;
School of Computing, National University of Singapore(computing 学院,新加坡国立大学)
;
School of Computing and Information Technology, Great Bay University(computing 与信息学院,大湾大学)
;
National Engineering Laboratory for Big Data System Computing Technology, Shenzhen University(大数据系统计算技术国家工程实验室,深圳大学)
专题命中
多模态生成
:multi-modal(title,abstract);分类 cs.CV
CommentsAccepted by IEEE Transactions on Information Forensics and Security 2025
机构
*
Department of Civil and Environmental Engineering, University of California, Berkeley(加州大学伯克利分校土木与环境工程系)
;
The Singapore-MIT Alliance for Research and Technology(新加坡-麻省理工联盟研究技术中心)
;
Department of Urban and Regional Planning, University of Florida(佛罗里达大学城市与区域规划系)
;
Department of Civil and Environmental Engineering, Massachusetts Institute of Technology(麻省理工学院土木与环境工程系)
;
Department of Urban Planning, Tsinghua University(清华大学城市规划系)
;
Department of Urban Studies and Planning, Massachusetts Institute of Technology(麻省理工学院城市研究与规划系)
Foundation Molecular Grammar: Multi-Modal Foundation Models Induce Interpretable Molecular Graph Languages
Michael Sun, Weize Yuan, Gang Liu, Wojciech Matusik, Jie Chen
机构
*
MIT CSAIL(麻省理工学院计算机科学与人工智能实验室)
;
MIT Chemistry(麻省理工学院化学系)
;
MIT-IBM Watson AI Lab, IBM Research(麻省理工-IBM Watson人工智能实验室,IBM研究院)
;
University of Notre Dame(诺埃伯大学)
Cross-Modal Causal Intervention for Medical Report Generation
Weixing Chen, Yang Liu, Ce Wang, Jiarui Zhu, Guanbin Li, Cheng-Lin Liu, Liang Lin
机构
*
School of Computer Science and Engineering, Sun Yat-sen University(中山大学计算机科学与工程学院)
;
Guangdong Key Laboratory of Big Data Analysis and Processing(广东大数据分析与处理重点实验室)
;
School of Science, Sun Yat-sen University(中山大学理学院)
;
Hong Kong Polytechnic University(香港理工大学)
;
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
专题命中
多模态生成
:cross-modal(title,abstract);分类 cs.CV
CommentsAccepted by IEEE TIP 2025, 16 pages, 11 figures, 7 tables
Journal refIEEE Transactions on Image Processing 34 (2025) 2970-2985