Multimodal Conditioned Diffusive Time Series Forecasting
机构 * University of Science and Technology of China(中国科学技术大学) ; University of Washington(华盛顿大学)
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.CL
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * University of Science and Technology of China(中国科学技术大学) ; University of Washington(华盛顿大学)
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.CL
机构 * Department of Computing, and Mathematics, Manchester Metropolitan University(计算机与数学系,曼彻斯特 Metropolitan 大学) ; Blackpool Teaching Hospitals NHS Foundation Trust(布莱克浦教学医院 NHS 基础信托) ; Northern Care Alliance, NHS Foundation Trust(北方护理联盟,NHS 基础信托) ; Department of Engineering, Manchester Metropolitan University(工程系,曼彻斯特 Metropolitan 大学) ; Faculty of Business and Law, Manchester Metropolitan University(商业与法律学院,曼彻斯特 Metropolitan 大学) ; Faculty of Medical Sciences, Newcastle University(医学科学学院,纽卡斯尔大学)
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.AI
机构 * Sun Yat-sen University, China(中山大学) ; Tongyi Lab, Alibaba Group(通义实验室,阿里巴巴集团) ; Peng Cheng Laboratory, Shenzhen, China(鹏城实验室,深圳,中国) ; Key Laboratory of Machine Intelligence and Advanced Computing, Ministry of Education, China(教育部人工智能与先进计算重点实验室)
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.CV
机构 * University of Bristol(布里斯托大学) ; X-Intelligence Labs(X-智能实验室) ; Meta
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.CV
机构 * Department of Mechanical Science and Engineering, University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校机械科学与工程系) ; Department of Mechanical Engineering, University of Michigan(密歇根大学机械工程系)
专题命中 视频多模态 :multi-modal(title,abstract);分类 cs.CV
Comments 17 pages, 12 figures
专题命中 视频多模态 :cross-modal(title,abstract);分类 cs.CV
Comments 23 pages, 10 figures
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.CV
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.CV
Comments Accepted by CVPR25
专题命中 视频多模态 :multimodal(title);分类 cs.CV、cs.CL、cs.AI
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.MM
专题命中 视频多模态 :multi-modal(title,abstract);分类 cs.CV
Comments Accepted by TPAMI
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.AI
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.AI
Comments Presented at AI Agent for Information Retrieval: Generating and Ranking (Agent4IR) @ AAAI 2025 [https://sites.google.com/view/ai4ir/aaai-2025]
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.AI
Comments Published in the Proceedings of the 16th International Conference on Social Robotics (ICSR) 2024,15 pages,5 figures,2 tables; work was co-funded by Horizon Europe project TERAIS under Grant agreement number 101079338
Journal ref In: Palinko, O., et al. Social Robotics. ICSR + AI 2024. vol 15563. Springer (2025)
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.CV
专题命中 视频多模态 :multi-modal(title,abstract);分类 cs.CV
Comments This is a part of article arXiv:2504.02287
专题命中 视频多模态 :multi-modal(title,abstract);分类 cs.CV
Comments Technical report
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.AI
Comments 12 pages, 5 figures, 1 table
专题命中 视频多模态 :multi-modal(title,abstract);分类 cs.CV
Comments To appear at CVPR 2025
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.CV
Comments CVPR 2025
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.CV
Comments Technical Report of VideoGLaMM
专题命中 视频多模态 :multi-modal(title,abstract);分类 cs.CV
专题命中 视频多模态 :multi-modal(title);multimodal(abstract);分类 cs.CV
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.AI
Comments Accepted at Human-Centered Explainable AI workshop at CHI 2024
专题命中 视频多模态 :multi-modal(title,abstract);分类 cs.CV
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.CV
Comments 13 pages, 8 figures, Github: https://github.com/mbzuai-oryx/TrackingMeetsLMM
专题命中 视频多模态 :multi-modal(title,abstract);分类 cs.CV
Comments Accepted to CVPR 2025. Dataset and code are available at https://github.com/google-research-datasets/egotempo.git
专题命中 视频多模态 :MLLM(title,abstract);分类 cs.CV
Comments CVPR 2025
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.AI
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.CV
Comments CVPR 2025