Multimodal HD Mapping for Intersections by Intelligent Roadside Units
专题命中 多模态生成 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV
Comments Accepted by ITSC'25
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 多模态生成 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV
Comments Accepted by ITSC'25
机构 * School of Electronic and Information Engineering, Beijing Jiaotong University(电子信息工程学院,北京交通大学) ; School of Electrical and Electronic Engineering, Imperial College London(电子电气工程学院,帝国理工学院伦敦分校)
专题命中 多模态生成 :multimodal(title,abstract);MLLM(abstract);分类 cs.AI
Comments This work has been submitted to the IEEE for possible publication
机构 * AI Geeks, Australia ; Australian Artificial Intelligence Institute, Australia(澳大利亚人工智能研究所) ; University of Liverpool, United Kingdom(利物浦大学) ; La Trobe University, Australia(拉特罗布大学)
专题命中 多模态生成 :multimodal(title,abstract);audio-visual(abstract);分类 cs.CV
机构 * School of Electronic Engineering and Computer Science(电子工程与计算机科学学院) ; Queen Mary University of London(伦敦女王学院) ; School of Engineering, College of Engineering and Physical Sciences(工程学院,工程与物理科学学院) ; University of Birmingham(伯明翰大学) ; Guangdong University of Technology(广东工业大学) ; Meta Inc. US(Meta美国公司) ; Nuffield Department of Clinical Neurosciences(临床神经科学系) ; University of Oxford(牛津大学) ; William Harvey Research Institute, NIHR Barts Biomedical Research Centre, Queen Mary University London(威廉·哈里弗研究所在NIHR巴茨生物医学研究中心,伦敦女王学院)
专题命中 多模态生成 :multi-modal(title,abstract);multimodal(abstract);分类 cs.CL
Comments Accepted by MICCAI 2025
机构 * Tsinghua University, Beijing, China(清华大学) ; Harbin Institute of Technology(哈尔滨工业大学)
专题命中 多模态生成 :multimodal(title,abstract);MLLM(abstract);分类 cs.AI
Comments Accepted by ACL 2025 Main, Camera Ready
机构 * Sun Yat-sen University(中山大学) ; Beihang University(北航) ; University of Chinese Academy of Sciences(中国科学院大学) ; Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) ; Beijing National Research Center for Information Science and Technology, Tsinghua University(北京信息科学与技术国家研究中心,清华大学) ; Department of Automation, Tsinghua University(清华大学自动化系)
专题命中 多模态生成 :multimodal(title,abstract);MLLM(abstract);分类 cs.CL
机构 * Zhejiang University(浙江大学) ; Alibaba Group(阿里巴巴集团)
专题命中 多模态生成 :MLLM(title,abstract);multimodal(abstract);分类 cs.CV
机构 * MedAI Technology (Wuxi) Co. Ltd.(MedAI技术(无锡)有限公司) ; Technical University of Munich(慕尼黑技术大学)
专题命中 多模态生成 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV
机构 * School of Computer Science and Technology, Chongqing University of Posts and Telecommunications(重庆邮电大学计算机科学与技术学院) ; Chongqing Key Laboratory of Image Recognition, Chongqing University of Posts and Telecommunications(重庆邮电大学图像识别重点实验室) ; Key Laboratory of Cyberspace Big Data Intelligent Security (Chongqing University of Posts and Telecommunications), Ministry of Education(教育部重庆邮电大学网络大数据智能安全重点实验室) ; College of Computer and Information Science, Chongqing Normal University(重庆师范大学计算机与信息科学学院)
专题命中 多模态生成 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV
Comments This paper has been accepted by IEEE Transactions on Multimedia (TMM) in March 2025
机构 * Inclusion AI Ant Group(Inclusion AI蚂蚁集团)
专题命中 多模态生成 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV
Comments https://github.com/inclusionAI/Ming/tree/Ming-Lite-Omni-Preview/Ming-unify
机构 * Inclusion AI
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments 18 pages,8 figures
机构 * Dept. of Pathology, University Medical Center Utrecht(病理学系,乌得勒支大学医学中心) ; Dept. of Biomedical Engineering, Eindhoven University of Technology(生物医学工程系,埃因霍温理工大学) ; Dept. of Mathematics and Computer Science, Eindhoven University of Technology(数学与计算机科学系,埃因霍温理工大学)
专题命中 多模态生成 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV
Comments 11 pages, 1 figure
机构 * S-Lab, Nanyang Technological University(南洋理工大学S实验室) ; SenseTime Research(商汤科技研究院) ; SenseTime Research and Tetras.AI(商汤研究与Tetras.AI)
专题命中 多模态生成 :multimodal(title,abstract);image-text(abstract);分类 cs.CV
机构 * Yonsei University(延世大学)
专题命中 多模态生成 :any-to-any(title,abstract);cross-modal(abstract);分类 cs.CL
Journal ref ACL 2025
机构 * University of Rochester(罗切斯特大学) ; Purdue University(普渡大学) ; NVIDIA(英伟达)
专题命中 多模态生成 :multi-modal(title,abstract);multimodal(abstract);分类 cs.CV
机构 * School of Electronic and Computer Engineering, Peking University(北京大学电子与计算机工程学院) ; ByteDance Inc.(字节跳动公司)
专题命中 多模态生成 :MLLM(title,abstract);multimodal(abstract);分类 cs.CV
机构 * Apple(苹果公司) ; Fudan University(复旦大学)
专题命中 多模态生成 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV
Comments Technical report
机构 * ByteDance Seed(字节跳动种子)
专题命中 多模态生成 :multi-modal(title,abstract);omni-modal(abstract);分类 cs.CV
Comments Mogao Technical Report
机构 * Tencent Hunyuan(腾讯文汇)
专题命中 多模态生成 :multimodal(title);multi-modal(abstract);image-text(abstract);分类 cs.CV
专题命中 多模态生成 :multimodal(title,abstract);multi-modal(abstract);分类 cs.CV
Comments Code available at: https://github.com/Alpha-VLLM/Lumina-mGPT
专题命中 多模态生成 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV
Comments Project page: https://mme-unify.github.io/
专题命中 多模态生成 :multimodal(title,abstract);any-to-any(abstract);分类 cs.CV
专题命中 多模态生成 :multi-modal(title,abstract);cross-modal(abstract);分类 cs.CV
Comments Our project page: https://jiepengwang.github.io/MMGen/
专题命中 多模态生成 :cross-modal(title,abstract);multi-modal(abstract);分类 cs.CV
Comments 11 pages, 9 figures, Accepted by IEEE Transactions on Medical Imaging
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments CVPR 2025 Camera-ready. Project page: https://silmm.github.io/
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments ICLR 2025 Camera-ready
专题命中 多模态生成 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV
Comments 31 pages, 6 figures
专题命中 多模态生成 :cross-modal(title);multimodal(abstract);分类 cs.CL、cs.AI、cs.MM
Comments Accepted by RepL4NLP 2025 @ NAACL 2025
专题命中 多模态生成 :multi-modal(title);multimodal(abstract);MLLM(abstract);分类 cs.CV
Comments [CVPR 2025] The project page is https://jianzongwu.github.io/projects/diffsensei/
专题命中 多模态生成 :multimodal(title,abstract);image-text(abstract);分类 cs.CV