MIMO: A medical vision language model with visual referring multimodal input and pixel grounding multimodal output
Yanyuan Chen, Dexuan Xu, Yu Huang, Songkun Zhan, Hanpin Wang, Dongxue Chen, Xueping Wang, Meikang Qiu, Hang Li
机构
*
School of Software & Microelectronics, Peking University(软件与微电子学院,北京大学)
;
School of Computer Science, Peking University(计算机学院,北京大学)
;
National Engineering Research Center for Software Engineering, Peking University(软件工程国家工程研究中心,北京大学)
;
Peking University Sixth Hospital(北京大学第六医院)
;
Augusta University(奥古斯塔大学)
;
Peking University First Hospital(北京大学第一医院)
Continual Adapter Tuning with Semantic Shift Compensation for Class-Incremental Learning
Qinhao Zhou, Yuwen Tan, Boqing Gong, Xiang Xiang
机构
*
HUST AI & Visual Learning Lab (HAIV Lab)(华中科技大学人工智能与视觉学习实验室)
;
Huazhong University of Science and Technology (HUST)(华中科技大学)
;
Department of Computer Science(计算机科学系)
;
Boston University(波士顿大学)
;
ByteDance, Inc.(字节跳动公司)
机构
*
School of Electronic Information and Communications, Huazhong University of Science and Technology(电子信息与通讯学院,华中科技大学)
;
Meta
;
AI Chip Center, Hong Kong University of Science and Technology(香港科技大学人工智能芯片中心)
;
Nanjing Forestry University(南京林业大学)
;
Autel Robotics
CommentsMonSter++: Unified Stereo Matching, Multi-view Stereo, and Real-time Stereo with Monodepth Priors, is the extended journal version of our earlier conference paper (arXiv:2501.08643) accepted to CVPR 2025
Weakly and Self-Supervised Class-Agnostic Motion Prediction for Autonomous Driving
Ruibo Li, Hanyu Shi, Zhe Wang, Guosheng Lin
机构
*
College of Computing and Data Science, Nanyang Technological University, Singapore(计算与数据科学学院,南洋理工大学,新加坡)
;
SenseTime Research, Hong Kong, China(时光科技研究院,香港,中国)
CommentsAn extension of our CVPR 2023 paper, "Weakly Supervised Class-Agnostic Motion Prediction for Autonomous Driving," accepted for publication in TPAMI