Multimodal Priors-Augmented Text-Driven 3D Human-Object Interaction Generation
多模态先验增强的文本驱动3D人-物交互生成
Yin Wang, Ziyao Zhang, Zhiying Leng, Haitian Liu, Frederick W. B. Li, Mu Li, Xiaohui Liang
机构
*
State Key Laboratory of Virtual Reality Technology and Systems(虚拟现实技术与系统国家重点实验室)
;
Beihang University(北京航空航天大学)
;
Department of Computer Science, University of Durham(杜伦大学计算机科学系)
;
Zhongguancun Laboratory(中关村实验室)
A Unified Framework for Multimodal Image Reconstruction and Synthesis using Denoising Diffusion Models
基于去噪扩散模型的多模态图像重建与合成统一框架
Weijie Gan, Xucheng Wang, Tongyao Wang, Wenshang Wang, Chunwei Ying, Yuyang Hu, Yasheng Chen, Hongyu An, Ulugbek S. Kamilov
机构
*
Department of Computer Science and Engineering, Washington University in St. Louis(华盛顿大学圣路易斯分校计算机科学与工程系)
;
Mallinckrodt Institute of Radiology, Washington University in St. Louis(华盛顿大学圣路易斯分校马林克罗德特放射医学研究所)
;
Department of Electrical and Systems Engineering, Washington University in St. Louis(华盛顿大学圣路易斯分校电气与系统工程系)
;
Department of Neurology, Washington University in St. Louis(华盛顿大学圣路易斯分校神经病学系)
;
Department of Biomedical Engineering, Washington University in St. Louis(华盛顿大学圣路易斯分校生物医学工程系)
;
Division of Biology and Biomedical Sciences, Washington University in St. Louis(华盛顿大学圣路易斯分校生物学与生物医学科学 division)
;
Department of Electrical and Computer Engineering, University of Wisconsin–Madison(威斯康星大学麦迪逊分校电气与计算机工程系)
机构
*
Institute for Artificial Intelligence, Peking University(北京大学人工智能研究院)
;
School of Psychological and Cognitive Sciences, Peking University(北京大学心理与认知科学学院)
;
Beijing Key Lab of Behavior and Mental Health, Peking University(北京大学行为与心理健康北京市重点实验室)
;
Beijing Institute for General Artificial Intelligence(北京一般人工智能研究院)
;
State Key Lab for General Artificial Intelligence(一般人工智能国家重点实验室)
;
Embodied Intelligence Lab, PKU-Wuhan Institute for Artificial Intelligence(具身智能实验室,北京大学武汉人工智能研究院)
;
Department of Computer Science and Technology, University of Cambridge(剑桥大学计算机科学与技术系)
机构
*
Center for Data Science, Peking University(北京大学数据科学中心)
;
Alibaba Group(阿里巴巴集团)
;
CASIA
;
Center for Machine Learning Research, Peking University(北京大学机器学习研究中心)
;
State Key Laboratory of General Artificial Intelligence, Peking University(北京大学通用人工智能国家重点实验室)
Diff4MMLiTS: Advanced Multimodal Liver Tumor Segmentation via Diffusion-Based Image Synthesis and Alignment
Diff4MMLiTS: 通过基于扩散的图像合成与对齐的先进多模态肝肿瘤分割
Shiyun Chen, Li Lin, Pujin Cheng, ZhiCheng Jin, JianJian Chen, HaiDong Zhu, Kenneth K. Y. Wong, Xiaoying Tang
机构
*
Department of Electronic and Electrical Engineering, Southern University of Science and Technology, Shenzhen, China(电子与电气工程系,南方科技大学,深圳,中国)
;
Department of Electrical and Electronic Engineering, The University of Hong Kong, Hong Kong SAR, China(电气与电子工程系,香港大学,香港特别行政区,中国)
;
Department of Radiology, Zhongda Hospital, Medical School, Southeast University, Nanjing, China(放射科,中大医院,医学院,东南大学,南京,中国)
;
Jiaxing Research Institute, Southern University of Science and Technology, Jiaxing, China(嘉兴研究所,南方科技大学,嘉兴,中国)
Multimodal Scientific Learning Beyond Diffusions and Flows
多模态科学学习超越扩散与流
Leonardo Ferreira Guilhoto, Akshat Kaushal, Paris Perdikaris
机构
*
Graduate Group in Applied Mathematics and Computational Science(应用数学与计算科学联合研究生组)
;
Department of Computer and Information Science(计算机与信息科学系)
;
Department of Mechanical Engineering and Applied Mechanics(机械工程与应用力学系)
RMFlow: Refined Mean Flow by a Noise-Injection Step for Multimodal Generation
RMFlow:通过噪声注入步骤细化均流以实现多模态生成
Yuhao Huang, Shih-Hsin Wang, Andrea L. Bertozzi, Bao Wang
机构
*
Department of Mathematics and Scientific Computing and Imaging (SCI) Institute University of Utah(数学与科学计算及成像学院(SCI)院,犹他大学)
;
Department of Mathematics, UCLA(数学系,加州大学洛杉矶分校)
WordCraft: Scaffolding the Keyword Method for L2 Vocabulary Learning with Multimodal LLMs
WordCraft: 通过多模态大语言模型 scaffolding 关键词方法用于L2词汇学习
Yuheng Shao, Junjie Xiong, Chaoran Wu, Xiyuan Wang, Ziyu Zhou, Yang Ouyang, Qinyi Tao, Quan Li
机构
*
School of Information Science and Technology, ShanghaiTech University(信息科学与技术学院,上海科技大学)
;
School of Creativity and Art, ShanghaiTech University(创意与艺术学院,上海科技大学)
;
Shanghai Fengxian Dai Wen Middle School(上海奉贤戴文中学)
Multi-modal Imputation for Alzheimer's Disease Classification
多模态缺失数据填补用于阿尔茨海默病分类
Abhijith Shaji, Tamoghna Chattopadhyay, Sophia I. Thomopoulos, Greg Ver Steeg, Paul M. Thompson, Jose-Luis Ambite
机构
*
Information Sciences Institute(信息科学研究所)
;
University of Southern California(美国南加州大学)
;
University of California(加州大学)
;
Stevens Neuroimaging and Informatics Institute(史蒂文斯神经影像与信息学研究所)
机构
*
Institut Curie, Université Paris-Saclay(巴黎-萨克勒大学Curie研究所)
;
Institute of Intelligent Systems and Robotics (ISIR), Sorbonne University(索邦大学智能系统与机器人研究所)
;
Assistance Publique – Hôpitaux de Paris (AP-HP), Université Paris Cité(巴黎公共医院(AP-HP)与巴黎城市大学)