Token Activation Map to Visually Explain Multimodal LLMs
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.AI
Comments ICCV2025 Accepted
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.AI
Comments ICCV2025 Accepted
机构 * Chinese Academy of Military Science(中国军事科学院) ; Changchun University of Science and Technology(长春理工大学) ; Hunan University(湖南大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments To appear in ICCV2025
机构 * Institute for AI, Peking University, Beijing, China(人工智能研究院,北京大学,北京,中国) ; State Key Laboratory of General Artificial Intelligence, Institute for AI, Peking University, Beijing, China(通用人工智能国家重点实验室,人工智能研究院,北京大学,北京,中国)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CL、cs.AI
Comments 17 pages, 13 figures
机构 * Skywork AI(Skywork人工智能) ; Kunlun Inc.(昆仑公司)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.CL
机构 * University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
机构 * School of Computer Science and Technology, Beijing Jiaotong University(计算机科学与技术学院,北京交通大学) ; Qifu Technology(启福科技)
专题命中 多模态训练与对齐 :MLLM(title,abstract);multimodal(abstract);分类 cs.CV、cs.AI
Comments Accepted by IJCAI 2025
机构 * Mohamed bin Zayed University of Artificial Intelligence(Mohamed bin Zayed大学人工智能学院) ; Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) ; Kyoto University(京都大学) ; Tencent AI Lab(腾讯AI实验室) ; Tomorrow Advancing Life(明天生命科技)
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CL、cs.AI
Comments Accepted by ACL2025 Main Conference
机构 * South China University of Technology, China(华南理工大学) ; Beihang University, China(北航) ; Pazhou Laboratory, China(琶洲实验室)
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CL、cs.MM
Comments Accepted in ACL 2025 Main Track
专题命中 多模态训练与对齐 :multi-modal(title,abstract);image-text(abstract);分类 cs.CV、cs.CL
Comments 6 pages, 1 figure, accepted by 2024 IEEE Conference on Artificial Intelligence (CAI)
Journal ref 2024 IEEE Conference on Artificial Intelligence (CAI), 2024, 480-485
机构 * Zhejiang University(浙江大学) ; Huawei Noah’s Ark Lab(华为诺亚实验室)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
机构 * Huawei Inc.(华为公司) ; University of Chinese Academy of Sciences(中国科学院大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.AI
机构 * Shanghai Jiao Tong University(上海交通大学) ; EPIC Lab, SJTU(EPIC实验室,SJTu) ; Shanghai AI Laboratory(上海人工智能实验室) ; Sun Yat-sen University(中山大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.CL
Comments ACL 2025 Findings
机构 * Queen Mary University of London(伦敦玛丽女王大学) ; Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) ; Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) ; The Hong Kong University of Science and Technology (GZ)(香港科学与技术大学)
专题命中 多模态训练与对齐 :MLLM(title,abstract);multimodal(abstract);分类 cs.CL、cs.MM
专题命中 多模态训练与对齐 :image-text(title,abstract);cross-modal(abstract);分类 cs.CV、cs.MM
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CL、cs.AI
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CL、cs.AI
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments Accepted by AAAI 2025
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.AI
Comments CVPR2025 Camera-ready
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.CL
Comments Accepted by CVPR 2025
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.AI
Comments 19 pages, 4 figures, submitted to Engineering Applications of Artificial Intelligence
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.CL
Comments 14 pages
专题命中 多模态训练与对齐 :cross-modal(title,abstract);multimodal(abstract);分类 cs.CV、cs.AI
Comments 11 pages, 2 figures, 3 tables
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.AI
Comments Github: https://github.com/NVlabs/Eagle, HuggingFace: https://huggingface.co/NVEagle
专题命中 多模态训练与对齐 :multi-modal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments In this version, we corrected some typos
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.CL
Comments Project Page: https://mm-rlhf.github.io/
专题命中 多模态训练与对齐 :MLLM(title,abstract);multimodal(abstract);分类 cs.CV、cs.AI
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
专题命中 多模态训练与对齐 :multi-modal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.AI
专题命中 多模态训练与对齐 :MLLM(title,abstract);multimodal(abstract);分类 cs.CV、cs.AI
Comments Technical Report on Slow Thinking with LLMs: Visual Reasoning