Generate the browsing process for short-video recommendation
机构 * Kuaishou Technology(快手科技)
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * Kuaishou Technology(快手科技)
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
机构 * Microsoft Redmond, WA 98052(微软红mond分校)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.AI
Comments In The Thirty-Ninth Annual Conference on Neural Information Processing Systems (NeurIPS 2025)
机构 * Robotics Institute, Carnegie Mellon University(卡内基梅隆大学机器人研究所) ; Department of Engineering Design, Indian Institute of Technology, Madras(印度理工学院Madras分校工程设计系)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.AI
Comments CoRL 2025
机构 * School of Automation, Northwestern Polytechnical University(自动化学院,西北工业大学) ; The University of Hong Kong(香港大学) ; School of Software, Northwestern Polytechnical University(软件学院,西北工业大学) ; Unmanned System Research Institute, Northwestern Polytechnical University(无人系统研究院,西北工业大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments ACM MM 2025
专题命中 视频多模态 :multimodal(abstract);分类 cs.CL
Comments Paper submitted to Language Sciences Journal
机构 * NUS(新加坡国立大学) ; HKUST(GZ)(香港科技大学(广州)) ; NTU(南洋理工大学) ; HKUST(香港科技大学) ; I 2 R, A*STAR(新加坡科技研究局) ; IPAL, CNRS(法国国家科学研究中心IPAL) ; CerCo, CNRS(法国国家科学研究中心CerCo)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Abstract Paper (Non-Archival) @ ICCV 2025 NeVi Workshop
机构 * Beihang University(北京航空航天大学) ; Dcar, ByteDance(字节跳动Dcar部门) ; Qfin Holdings,Inc(Qfin控股公司) ; MAIS, Institute of Automation, Chinese Academy of Sciences(自动化研究所,中国科学院)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
机构 * College of Intelligence Science and Technology, National University of Defense Technology(智能科学与技术学院,国防科技大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.AI
Comments 20 pages, 19 figures, accepted by IEEE Transactions on Robotics
机构 * University of Science and Technology of China(中国科学技术大学) ; The University of Adelaide(阿德莱德大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Accepted to ICCV 2025
机构 * University of Twente(特文特大学) ; University of Bath(巴斯大学) ; Shanghai Jiao Tong University(上海交通大学) ; PhiGent Robotics(PhiGent机器人)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
Comments Accepted by ICRA 2025
机构 * Keye Team, Kuaishou Group(快手集团Keye团队)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Github page: https://github.com/Kwai-Keye/Keye
机构 * Zhejiang University and Alibaba Cloud(浙江大学和阿里云) ; Alibaba Cloud(阿里云) ; Zhejiang University(浙江大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
Comments Accepted by T-ASE and CoRL25 GenPriors Workshop
机构 * Department of Machine Learning Mohamed bin Zayed University of Artificial Intelligence (MBZUAI) Abu Dhabi, UAE(机器学习系,Mohamed bin Zayed人工智能大学(MBZUAI),阿布扎比,阿联酋) ; Department of Computer Vision Mohamed bin Zayed University of Artificial Intelligence (MBZUAI) Abu Dhabi, UAE(计算机视觉系,Mohamed bin Zayed人工智能大学(MBZUAI),阿布扎比,阿联酋)
专题命中 视频多模态 :image-text(abstract);分类 cs.CV
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
Comments 8 pages, 6 figures, 10th International Conference on Affective Computing and Intelligent Interaction (ACII 2022)
机构 * University of Colorado Anschutz Medical Campus(科罗拉多大学安舒茨医学校园) ; University of Colorado Boulder(科罗拉多大学波德分校) ; Northeastern University(东北大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CL
机构 * Faculty of Computer Science, Electronics and Telecommunications(计算机科学与电子技术学院) ; AGH University of Science and Technology(AGH科技大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments 12 pages, 8 figures, 3 tables; dataset descriptor paper introducing DVS-PedX (synthetic-and-real event-based pedestrian dataset with baselines) External URL: https://doi.org/10.5281/zenodo.17030898
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
Comments arXiv admin note: text overlap with arXiv:2403.07815 by other authors
Journal ref Ocean Engineering, Volume 341, Part 2, 1 December 2025, Article 122502
机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) ; SpreeAI
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
Comments Project Page: https://immortalco.github.io/DressAndDance/
专题命中 视频多模态 :multimodal(abstract);分类 cs.MM
Comments EMNLP2025 Findings
专题命中 视频多模态 :multi-modal(abstract);分类 cs.AI
Comments 10 pages, 6 figures
机构 * University of Science(科学大学) ; University of Dayton(代顿大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments ACM Multimedia 2025
机构 * Kunming University of Science and Technology(昆明理工大学)
专题命中 视频多模态 :cross-modal(abstract);分类 cs.CV
机构 * School of Information Science and Electronic Engineering, Shanghai Jiao Tong University(上海交通大学信息科学与电子工程学院) ; MoE Key Lab of Artificial Intelligence, AI Institute, Shanghai Jiao Tong University(上海交通大学人工智能教育部重点实验室) ; The University of Tokyo(东京大学) ; Shanghai AI Laboratory(上海人工智能实验室) ; E-surfing Vision Technology Co., Ltd(亿欧视觉科技有限公司)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Accepted by ICCV2025
机构 * College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) ; Home Robotics Lab, E-surfing Digital Life Technology Co., Ltd., China Telecom(E-surfing数字生活技术有限公司) ; Zhejiang Lab(浙江实验室) ; Trinity College Dublin(都柏林大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
Comments Accepted by EMNLP 2025 (Main Conference)
专题命中 视频多模态 :cross-modal(abstract);分类 cs.AI
Comments The paper is submitted to IAAI26. Total 9 pages with 8 figures
机构 * AMAP, Alibaba Group(阿里集团AMAP)
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
机构 * Xidian University(西电大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
机构 * Universidad de Buenos Aires, Facultad de Ciencias Exactas y Naturales(布宜诺斯艾利斯大学,精确科学与自然学院) ; Institute of Engineering Sciences, Universidad de O’Higgins(工程科学研究所,奥希金斯大学) ; CONICET-UBA, Instituto de Ciencias de la Computacion (ICC)(CONICET-UBA,计算科学研究所) ; L3S Research Center, Leibniz Universität Hannover(L3S研究中心,汉诺威莱布尼茨大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
Journal ref 2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCVW); 2nd Workshop on Neuromorphic Vision (NeVi)
机构 * Louisiana State University(路易斯安那州立大学) ; Northeastern University(东北大学) ; Yale University(耶鲁大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
Comments This paper will appear in CCS 2025
机构 * MoE Key Lab of Artificial Intelligence, AI Institute Shanghai Jiao Tong University Shanghai China(人工智能联合实验室,人工智能研究院,上海交通大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments ACM MM 2025; Code is released at https://github.com/VISION-SJTU/AnchorSync