MPFlow: Multi-modal Posterior-Guided Flow Matching for Zero-Shot MRI Reconstruction
MPFlow: 多模态后验引导的流匹配用于零样本MRI重建
Seunghoi Kim, Chen Jin, Henry F. J. Tregidgo, Matteo Figini, Daniel C. Alexander
机构
*
1 Hawkes Institute, UCL \, 2 Dept. of Medical Physics
;
Biomedical Engineering, UCL 3 Dept. of Computer Science, UCL 4 Centre for AI, DS\&AI, AstraZeneca, UK
机构
*
National Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室)
;
Data Science & Artificial Intelligence Research Institute, China Unicom(中国unicom数据科学与人工智能研究院)
;
Unicom Data Intelligence, China Unicom(中国unicom数据智能)
Align-KD: Distilling Cross-Modal Alignment Knowledge for Mobile Vision-Language Model Enhancement
Align-KD:为移动视觉语言模型增强提取跨模态对齐知识
Qianhan Feng, Wenshuo Li, Tong Lin, Xinghao Chen
机构
*
State Key Laboratory of General Artificial Intelligence, School of Intelligence Science and Technology, Peking University, China(通用人工智能国家重点实验室,智能科学与技术学院,北京大学,中国)
;
Huawei Noah’s Ark Lab, China(华为诺亚方舟实验室,中国)
UF-AMA: A unified framework for cross-domain emotion recognition via adaptive multimodal alignment
UF-AMA: 通过自适应多模态对齐的跨域情感识别统一框架
Zheng Wang, Shuo Wang, Junhong Wang
机构
*
Institute of Advanced Technology, University of Science and Technology of China(中国科学技术大学先进技术研究院)
;
Department of Electronic Engineering and Information Science, University of Science and Technology of China(中国科学技术大学电子工程与信息科学系)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合国家科学中心人工智能研究院)
Variational Adapter for Cross-modal Similarity Representation
变分适配器用于跨模态相似性表示
WenZhang Wei, Zhipeng Gui, Dehua Peng, Tiandi Ye, Huayi Wu
机构
*
School of Remote Sensing and Information Engineering(遥感与信息工程学院)
;
Wuhan University(武汉大学)
;
School of Data Science and Engineering(数据科学与工程学院)
;
East China Normal University(华东师范大学)
;
State Key Laboratory of Information Engineering in Surveying, Mapping and Remote Sensing(测绘遥感信息工程国家重点实验室)
机构
*
National University of Singapore(国立新加坡大学)
;
University of Science and Technology of China(中国科学技术大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))
;
Central South University(中南大学)
SEATrack: Simple, Efficient, and Adaptive Multimodal Tracker
SEATrack: 简单、高效且自适应的多模态跟踪器
Junbin Su, Ziteng Xue, Shihui Zhang, Kun Chen, Weiming Hu, Zhipeng Zhang
机构
*
School of Artificial Intelligence (School of Software), Yanshan University(燕山大学人工智能学院(软件学院))
;
School of Software, Beihang University(北航软件学院)
;
Hebei Key Laboratory of Computer Virtual Technology and System Integration(河北省计算机虚拟技术与系统集成重点实验室)
;
State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), CASIA(多模态人工智能系统国家重点实验室(MAIS),CASIA)
;
School of Artificial Intelligence, UCAS(中国科学院大学人工智能学院)
;
AutoLab, School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院自动化实验室)
From Reasoning to Pixels: Benchmarking the Alignment Gap in Unified Multimodal Models
从推理到像素:统一多模态模型中对齐差距的基准测试
Cheng Yang, Chufan Shi, Bo Shui, Yaokang Wu, Muzi Tao, Huijuan Wang, Ivan Yee Lee, Yong Liu, Xuezhe Ma, Taylor Berg-Kirkpatrick
机构
*
University of California San Diego(加利福尼亚大学圣迭戈分校)
;
University of Southern California(南加利福尼亚大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Carnegie Mellon University(卡内基梅隆大学)
机构
*
State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(西安交通大学人工智能与机器人研究所人机混合增强智能全国重点实验室)
Comparative analysis of dual-form networks for live land monitoring using multi-modal satellite image time series
多模态卫星图像时间序列用于实时土地监测的双形式网络比较分析
Iris Dumeur, Jérémy Anger, Gabriele Facciolo
机构
*
1 Kayrros SAS 2 Universit\'e Paris-Saclay, ENS Paris-Saclay, CNRS, Centre Borelli, 91190, Gif-sur-Yvette, France 3 Institut Universitaire de France
Multi-Modal Image Fusion via Intervention-Stable Feature Learning
多模态图像融合 via 干预稳定的特征学习
Xue Wang, Zheng Guan, Wenhua Qian, Chengchao Wang, Runzhuo Ma
机构
*
School of Information Science and Engineering, Yunnan University(云南大学信息科学与工程学院)
;
School of Artificial Intelligence, Nanyang Normal University(南阳师范学院人工智能学院)
;
Department of Electrical and Electronic Engineering, Hong Kong Polytechnic University(香港理工大学电子与电气工程系)
Learning Progressive Adaptation for Multi-Modal Tracking
多模态跟踪的渐进适应学习
He Wang, Tianyang Xu, Zhangyong Tang, Xiao-Jun Wu, Josef Kittler
机构
*
School of Artificial Intelligence and Computer Science, Jiangnan University(江南大学人工智能与计算机科学学院)
;
Centre for Vision, Speech and Signal Processing, University of Surrey(Surrey大学视觉、语音和信号处理中心)
机构
*
School of Computer Science and Engineering, Northeastern University, Shenyang, China(东北大学计算机科学与工程学院)
;
Key Laboratory of Intelligent Computing in Medical Image of Ministry of Education, Northeastern University, Shenyang, China(教育部医学图像智能计算重点实验室)
;
National Frontiers Science Center for Industrial Intelligence and Systems Optimization, Shenyang, China(工业智能与系统优化国家级前沿科学中心)
;
Alberta Machine Intelligence Institute, University of Alberta, Edmonton, Canada(阿尔伯塔机器智能研究所,阿尔伯塔大学,加拿大爱德蒙顿)