MLLMRec: A Preference Reasoning Paradigm with Graph Refinement for Multimodal Recommendation
MLLMRec: 基于图细化的多模态推荐偏好推理范式
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract)
AI总结 MLLMRec通过图细化和多模态大语言模型提升多模态推荐的用户偏好推理与物品表示学习准确性。
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
MLLMRec: 基于图细化的多模态推荐偏好推理范式
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract)
AI总结 MLLMRec通过图细化和多模态大语言模型提升多模态推荐的用户偏好推理与物品表示学习准确性。
Wave2Word: 一种多模态Transformer框架,用于神经重症监护中的联合EEG-文本对齐和多任务表示学习
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
AI总结 Wave2Word提出一种多模态Transformer框架,通过整合信号域建模与结构化临床语言监督,实现EEG-文本对齐和多任务表示学习,提升神经重症监护中的EEG分析效果。
Doctor Sun: 一种双语多模态大语言模型用于生物医学AI
机构 * Key Laboratory of Smart Manufacturing in Energy Chemical Process, Ministry of Education East China University of Science and Technology(能源化工过程智能制造重点实验室,东华大学) ; Research Institute of Intelligent Control and Systems Harbin Institute of Technology(智能控制与系统研究室,哈尔滨工业大学) ; Department of Emergency Medicine, Sir Run Run Shaw Hospital Zhejiang University School of Medicine(浙江大学医学院急诊医学科) ; Provincial Key Laboratory of Precise Diagnosis Treatment of Abdominal Infection, Sir Run Run Shaw Hospital Zhejiang University School of Medicine(腹部感染精准诊断治疗省级重点实验室,浙江大学医学院) ; School of Medicine Shaoxing University(绍兴大学医学院)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL、cs.AI、cs.MM
AI总结 Doctor Sun是一种双语多模态大语言模型,通过整合预训练视觉编码器和医学LLM,提升生物医学多模态任务的性能,并提供SunMed-VL数据集支持研究进展。
具有语义空间对齐的层次多模态大语言模型用于增强的时间序列分类
机构 * State Key Laboratory of Cognitive Intelligence, University of Science and Technology of China(认知智能国家重点实验室,中国科学技术大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
AI总结 HiTime通过层次多模态大语言模型和语义空间对齐,提升时间序列分类的性能。
多模态融合与可解释性在人体活动识别中的应用:一种可复现的基于传感器建模框架
专题命中 多模态训练与对齐 :multimodal(title,abstract);multi-modal(abstract)
AI总结 本文提出了一种可复现的多模态融合框架,通过统一预处理和融合策略提升人体活动识别的准确性和可解释性。
Comments 33 pages, 12 figures, 4 tables
Senti-iFusion: 一种以完整性为中心的多模态情感分析多模态融合框架,用于在不确定模态缺失情况下
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
AI总结 Senti-iFusion提出了一种以完整性为中心的多模态融合框架,通过分层结构处理模态缺失问题,提升多模态情感分析的鲁棒性和准确性。
机构 * The University of Sydney(悉尼大学) ; School of Computer Science, The University of Sydney(悉尼大学计算机科学学院) ; National Engineering Laboratory for Integrated Aero-Space-Ground-Ocean Big Data Application Technology(集成空天地海大数据应用技术国家工程实验室) ; School of Computer Science and Engineering, Northwestern Polytechnical University(西北工业大学计算机科学与工程学院) ; University of Maryland(马里兰大学) ; Ningbo Institute of Northwestern Polytechnical University(西北工业大学宁波学院)
专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV、cs.AI、cs.MM
Comments Accepted by IEEE Transactions on Medical Imaging (TMI). Code available at https://github.com/TianyiFranklinWang/MIRROR. Project page: https://tianyifranklinwang.github.io/MIRROR
Journal ref IEEE Trans. Med. Imaging (2025)
机构 * Department of Computing, The Hong Kong Polytechnic University(计算系,香港理工大学) ; Research Institute of Multiple Agents and Embodied Intelligence, Pengcheng Laboratory(多智能体与具身智能研究院,鹏城实验室)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments Accepted to IEEE TMM
专题命中 多模态训练与对齐 :multi-modal(title,abstract);cross-modal(abstract)
Comments Accepted by AAAI 2026
专题命中 多模态训练与对齐 :cross-modal(title,abstract);multimodal(abstract)
Comments 7 pages, 2 figures, 1 table
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
机构 * Scuola Internazionale Superiore di Studi Avanzati (SISSA)(国际先进研究学院(SISSA))
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
Comments 24 pages, 11 figures
机构 * University of Pennsylvania(宾夕法尼亚大学) ; University of Hong Kong(香港大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
Comments Accepted by NeruIPS 2025
机构 * Center for Wireless Communications, University of Oulu(无线通信中心,奥卢大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
Comments 5 pages, 3 figures, 1 table
机构 * College of Information Science and Electronic Engineering, Zhejiang University(信息科学与电子工程学院,浙江大学) ; College of Computer Science and Technology, Zhejiang University of Technology(计算机科学与技术学院,浙江工业大学) ; Zhejiang University Jinhua Research Institute(浙江大学金华研究院)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.MM
机构 * MBZUAI(穆扎芬人工智能研究所) ; University of Zhengzhou(郑州大学) ; HKUST(香港科技大学) ; University of Aberdeen(爱丁堡大学) ; MBZUAI, Weizmann Institute of Science(穆扎芬人工智能研究所、威斯曼科学研究所)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
专题命中 多模态训练与对齐 :multi-modal(title,abstract);cross-modal(abstract)
机构 * Dept. of Electronic Engineering at Sogang University(ソガン大学电子工程系) ; Dept. of Immersive Media and Engineering at Sungkyunkwan University(顺天大学沉浸媒体与工程系) ; College of Medicine, Yonsei University(延世大学医学院)
专题命中 多模态训练与对齐 :cross-modal(title,abstract);multi-modal(abstract)
机构 * University of Maryland, College Park(马里兰大学学院 park)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
机构 * Radha Gulhane(独立研究者) ; Sathish Reddy Indurthi(独立研究者)
专题命中 多模态训练与对齐 :MLLM(title);multimodal(abstract);分类 cs.CV、cs.CL、cs.AI
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
Comments ACM MM 2025
机构 * Northwestern University(西北大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
Comments Accepted to Neurips 2025 (Spotlight)
机构 * School of Computer Science, Wuhan University(武汉大学计算机学院) ; School of Artificial Intelligence, Wuhan University(武汉大学人工智能学院)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
Comments Accepted by CIKM 2025
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
机构 * University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments Accepted by COLM 2025
专题命中 多模态训练与对齐 :cross-modal(title,abstract);multimodal(abstract)
Comments This paper has been accepted by ACM MM 2025
机构 * New York UniversityUSA(纽约大学) ; Max Planck SocietyGermany(马克斯·普朗克研究所)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CL、cs.MM、eess.AS
Comments Interspeech 2025