Hierarchical MoE: Continuous Multimodal Emotion Recognition with Incomplete and Asynchronous Inputs
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
机构 * Arizona State University(亚利桑那州立大学) ; Clemson University(克莱姆森大学) ; LinkedIn Corporation(领英公司) ; Ludwig Maximilian University of Munich(慕尼黑路德维希-马克西米利安大学)
专题命中 多模态训练与对齐 :multi-modal(title);multimodal(abstract);分类 cs.CV、cs.CL、cs.AI
专题命中 多模态训练与对齐 :multi-modal(title,abstract);multimodal(abstract)
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract)
机构 * Feng Xu 1,2(作者1单位) ; Hui Wang 1(作者1单位) ; Yuting Huang 3(作者3单位) ; Danwei Zhang 4(作者4单位) ; Zizhu Fan 5(作者5单位)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
Comments 9 pages, 5 figures
机构 * The Chinese University of Hong Kong(香港中文大学) ; University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments Accepted to ICCV 2025
机构 * Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(中国教育部多媒体可信感知与高效计算重点实验室,厦门大学) ; Institute of Artificial Intelligence, Xiamen University(厦门大学人工智能研究院)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.MM
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
机构 * Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学人工智能学院) ; Stanford University(斯坦福大学) ; Microsoft Corporation(微软公司)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments Homepage: https://haon-chen.github.io/MoCa/
机构 * Baidu Inc.(百度公司)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
机构 * Georgia Institute of Technology(佐治亚理工学院)
专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments Accepted to CVPR 2025
机构 * Dept. of Electrical and Computer Engineering(电气与计算机工程系) ; New York University(纽约大学) ; Tandon School of Engineering(坦顿工程学院)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
机构 * College of Computer Science and Technology, Zhejiang University of Technology(浙江工业大学计算机科学与技术学院)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.MM
Comments Submitted to TAC. The code is available at https://github.com/gw-zhong/CIDer
机构 * State Key Laboratory of Brain Cognition and Brain-inspired Intelligence Technology(脑认知与脑启发智能技术重点实验室) ; Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) ; School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) ; Institute of Neuroscience, State Key Laboratory of Brain Cognition and Brain-Inspired Intelligence Technology(神经科学研究所) ; CAS Center for Excellence in Brain Science and Intelligence Technology(中国科学院脑科学与智能技术卓越创新中心) ; Chinese Academy of Sciences(中国科学院)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments Published on Nature Machine Intelligence
Journal ref Nature Machine Intelligence, 2025
机构 * Singapore University of Technology and Design (SUTD)(新加坡科技设计大学) ; Institute for Infocomm Research (I2R), A*STAR, Singapore(信息与通信研究院(I2R))
专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV、cs.AI、cs.MM
Comments Accepted at the 2025 International Conference on Machine Learning (ICML)
专题命中 多模态训练与对齐 :multimodal(title);multi-modal(abstract);cross-modal(abstract)
机构 * Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院,清华大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
机构 * Department of Artificial Intelligence, Korea University(人工智能系,韩国大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.AI、cs.MM
Comments Accepted to ACL 2025 (Findings)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments Project Page: https://stylemotif.github.io
专题命中 多模态训练与对齐 :multi-modal(title,abstract);cross-modal(abstract)
Comments 6 pages,3 figures
Journal ref ICME 2025
专题命中 多模态训练与对齐 :multi-modal(title,abstract);image-text(abstract)
Comments This paper is accepted by CVPR 2025
专题命中 多模态训练与对齐 :multi-modal(title,abstract);cross-modal(abstract)
专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments update comparison with sota and analysis
专题命中 多模态训练与对齐 :cross-modal(title,abstract);multimodal(abstract)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
Comments 8 pages,6 figures,2 tables,submitted to the 33rd ACM International Conference on Multimedia(ACM MM 2025)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments 11 pages, 9 figures
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments Accepted to ICLR 2025. Code: https://github.com/ictnlp/LLaVA-Mini Model: https://huggingface.co/ICTNLP/llava-mini-llama-3.1-8b
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract)
Comments Accepted by KDD 2025 (CR)