AnyUser: Translating Sketched User Intent into Domestic Robots
AnyUser: 将草图用户意图翻译成家用机器人
Songyuan Yang, Huibin Tan, Kailun Yang, Wenjing Yang, Shaowu Yang
机构
*
College of Computer Science and Technology, National University of Defense Technology(国防科技大学计算机科学与技术学院)
;
National Engineering Research Center of Robot Visual Perception and Control Technology, Hunan University(湖南大学国家机器人视觉感知与控制技术工程研究中心)
UniSurgSAM: A Unified Promptable Model for Reliable Surgical Video Segmentation
UniSurgSAM: 一种用于可靠外科视频分割的统一提示模型
Haofeng Liu, Ziyue Wang, Alex Y. W. Kong, Guanyi Qin, Yunqiu Xu, Chang Han Low, Mingqi Gao, Lap Yan Lennon Chan, Yueming Jin
机构
*
Department of Biomedical Engineering, National University of Singapore(新加坡国立大学生物医学工程系)
;
Department of Electrical and Computer Engineering, National University of Singapore(新加坡国立大学电气与计算机工程系)
;
School of Computer Science, The University of Sheffield(谢菲尔德大学计算机科学学院)
;
Department of Computer Science and Engineering, The Chinese University of Hong Kong(香港中文大学计算机科学与工程系)
机构
*
Beihang University(北京航空航天大学)
;
Centre for Artificial Intelligence and Robotics, HKISI-CAS(香港智能科学与工业研究院人工智能与机器人中心)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Hangzhou International Innovation Institute, Beihang University(北京航空航天大学杭州国际创新研究院)
;
Beijing Jiaotong University(北京交通大学)
GenSmoke-GS: A Multi-Stage Method for Novel View Synthesis from Smoke-Degraded Images Using a Generative Model
GenSmoke-GS:一种基于生成模型的多阶段方法,用于从烟雾退化图像中生成新视角
Qida Cao, Xinyuan Hu, Changyue Shi, Jiajun Ding, Zhou Yu, Jun Yu
机构
*
School of Computer Science and Technology, Hangzhou Dianzi University(杭州电子科技大学计算机科学与技术学院)
;
School of Computer Science and Technology, Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)计算机科学与技术学院)
A Generative Foundation Model for Multimodal Histopathology
多模态病理生成模型基础框架
Jinxi Xiang, Mingjie Li, Siyu Hou, Yijiang Chen, Xiangde Luo, Yuanfeng Ji, Xiang Zhou, Ehsan Adeli, Akshay Chaudhari, Curtis P. Langlotz, Kilian M. Pohl, Ruijiang Li
机构
*
Department of Radiation Oncology, Stanford University School of Medicine(斯坦福大学医学院放射肿瘤学系)
;
Department of Psychiatry and Behavioral Sciences, Stanford University School of Medicine(斯坦福大学医学院精神病学与行为科学系)
;
Department of Statistics and Data Science, Yale University(耶鲁大学统计与数据科学系)
;
Department of Computer Science, Stanford University(斯坦福大学计算机科学系)
;
Department of Electrical Engineering, Stanford University(斯坦福大学电气工程系)
;
Department of Biomedical Data Science, Stanford University(斯坦福大学生物医学数据科学系)
;
Department of Radiology, Stanford University(斯坦福大学放射学系)
;
Center for Artificial Intelligence in Medicine and Imaging, Stanford University(斯坦福大学医学与影像人工智能中心)
CommentsThis is a submission to the "Pattern Analysis and Applications". The manuscript includes 14 pages and 6 figures. All authors have approved the submission, and there is no conflict of interest to declare
机构
*
National Center of Technology Innovation for Intelligent Design and Numerical Control, Huazhong University of Science and Technology(华中科技大学国家智能设计与数控技术创新中心)
;
School of Artificial Intelligence and Robotics, Hunan University(湖南大学人工智能与机器人学院)
;
Department of Intelligent Manufacturing, Contemporary Amperex Technology Ltd.(宁德时代新能源科技股份有限公司智能制造部)
;
Institute for Infocomm Research, A*STAR(新加坡科技研究局资讯通信研究院)
;
School of Electrical and Electronic Engineering, Nanyang Technological University(南洋理工大学电气与电子工程学院)
Enhancing Foundation VLM Robustness to Missing Modality: Scalable Diffusion for Bi-directional Feature Restoration
增强基础视觉语言模型对缺失模态的鲁棒性:可扩展的扩散用于双向特征恢复
Wei Dai, Haoyu Wang, Honghao Chang, Lijun He, Fan Li, Jian Sun, Haixia Bi
机构
*
Department of Electronics
;
Information \ 'an Jiaotong University Xi'an China
;
School of Information
;
Communications Engineering \ 'an Jiaotong University Xi'an China
;
School of Mathematics
;
Statistics \ 'an Jiaotong University Xi'an China
;
Information \ 'an Jiaotong University
;
Communications Engineering \ 'an Jiaotong University
;
Statistics \ 'an Jiaotong University
机构
*
School of Computer Science and Engineering, Northeastern University(东北大学计算机科学与工程学院)
;
Xingning Power Supply Bureau, Guangdong Power Grid Co., Ltd.(广东电网有限责任公司兴宁供电局)
MuDD: A Multimodal Deception Detection Dataset and GSR-Guided Progressive Distillation for Non-Contact Deception Detection
MuDD:一种多模态欺骗检测数据集和GSR引导的渐进性知识蒸馏用于非接触欺骗检测
Peiyuan Jiang, Yao Liu, Yanglei Gan, Jiaye Yang, Lu Liu, Daibing Yao, Qiao Liu
机构
*
School of Computer Science and Engineering, University of Electronic Science and Technology of China(电子科技大学计算机科学与工程学院)
;
School of Information and Software Engineering, University of Electronic Science and Technology of China(电子科技大学信息与软件工程学院)
Multimodal Urban Tree Detection from Satellite and Street-Level Imagery via Annotation-Efficient Deep Learning Strategies
通过高效深度学习策略实现多模态城市树木检测:结合卫星和街景图像
In Seon Kim, Ali Moghimi
机构
*
Department of Computer Science, University of California, Davis(加州大学戴维斯分校计算机科学系)
;
Department of Biological and Agricultural Engineering, University of California, Davis(加州大学戴维斯分校生物与农业工程系)