RANGER: A Monocular Zero-Shot Semantic Navigation Framework through Visual Contextual Adaptation
RANGER:通过视觉上下文适应的单目零样本语义导航框架
Ming-Ming Yu, Yi Chen, Börje F. Karlsson, Wenjun Wu
机构
*
Beihang University(北京航空航天大学)
;
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Hangzhou International Innovation Institute, Beihang University(北京航空航天大学杭州国际创新研究院)
专题命中
视频多模态
:multimodal(abstract);multimodal foundation model(abstract);分类 cs.CV
Seeing Beyond the Image: ECG and Anatomical Knowledge-Guided Myocardial Scar Segmentation from Late Gadolinium-Enhanced Images
超越图像:结合心电图与解剖知识的晚期钆增强图像心肌瘢痕分割
Farheen Ramzan, Yusuf Kiberu, Nikesh Jathanna, Meryem Jabrane, Vicente Grau, Shahnaz Jamil-Copley, Richard H. Clayton, Chen, Chen
机构
*
School of Computer Science, University of Sheffield(谢菲尔德大学计算机科学学院)
;
Department of Cardiology, Trent Cardiac Centre, Nottingham City Hospital, Nottingham University Hospitals NHS Trust(诺丁汉大学医院NHS信托基金诺丁汉市医院特伦特心脏中心心脏病科)
;
School of Medicine, University of Nottingham(诺丁汉大学医学院)
;
Department of Engineering Science, University of Oxford(牛津大学工程科学系)
TTA-Vid: Generalized Test-Time Adaptation for Video Reasoning
TTA-Vid: 视频推理的通用测试时间适应
Soumya Shamarao Jahagirdar, Edson Araujo, Anna Kukleva, M. Jehanzeb Mirza, Saurabhchand Bhati, Samuel Thomas, Brian Kingsbury, Rogerio Feris, James R. Glass, Hilde Kuehne
机构
*
University of Tübingen(蒂宾根大学)
;
MIT(麻省理工学院)
;
IBM Research(IBM研究院)
;
MIT-IBM Watson AI Lab(MIT-IBM沃森人工智能实验室)
;
Tuebingen AI Center(蒂宾根人工智能中心)
;
Max Planck Institute for Informatics(马克斯·普朗克信息学研究所)