ClearSight: Human Vision-Inspired Solutions for Event-Based Motion Deblurring
机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
专题命中 视频多模态 :cross-modal(abstract);分类 cs.CV
Comments Accepted by ICCV 2025
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
专题命中 视频多模态 :cross-modal(abstract);分类 cs.CV
Comments Accepted by ICCV 2025
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Published in 2025 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP). Project page: https://physid.github.io/
Journal ref 2025 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Hyderabad, India, 2025, pp. 1-5
机构 * University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) ; Amazon.com LLC(亚马逊公司)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
机构 * National Key Laboratory of Human-Machine Hybrid Augmented Intelligence(国家人类-机器混合增强智能重点实验室) ; National Engineering Research Center for Visual Information and Applications(国家视觉信息与应用工程研究中心) ; Institute of Artificial Intelligence and Robotics(人工智能与机器人研究所) ; Xi’an Jiaotong University(西安交通大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
机构 * University College London(伦敦大学学院) ; University of Washington(华盛顿大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CL
机构 * Tsinghua University(清华大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
机构 * Tsinghua University(清华大学) ; Shanghai Research Institute for Intelligent Autonomous Systems,Tongji University(上海智能自主系统研究所,同济大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
机构 * Center for Visual Information Technology (CVIT)(视觉信息科技中心)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
机构 * School of Computer and Communication Engineering, University of Science and Technology Beijing(计算机与通信工程学院,北京科技大学)
专题命中 视频多模态 :cross-modal(abstract);分类 cs.CL
机构 * Beijing Normal University(北京师范大学) ; Tsinghua University(清华大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
机构 * SKLCCSE, Institute of Artificial Intelligence, Beihang University, Beijing, China(信息与通信工程学院,人工智能研究院,北京航空航天大学,北京,中国) ; Beijing Advanced Innovation Center for Future Blockchain and Privacy Computing, Beihang University(未来区块链与隐私计算先进创新中心,北京航空航天大学) ; Hangzhou International Innovation Institute, Beihang University, Hangzhou, China(杭州国际创新研究院,北京航空航天大学,杭州,中国)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments code and training recipes are available at https://github.com/ZhangXJ199/TinyLLaVA-Video
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
Comments 26 pages, 10 tables, 6 figures, accepted at Image and Vision Computing (IMAVIS)
Journal ref In: Image and Vision Computing 155 (2025), p. 105437. issn: 0262-8856
机构 * Sun Yat-sen University(中山大学) ; Lanzhou University(兰州大学) ; University of Hong Kong(香港大学) ; Hefei University of Technology(合肥工业大学) ; National University of Singapore(新加坡国立大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
机构 * Lenovo Research(联想研究院) ; Tsinghua University(清华大学) ; School of Artificial Intelligence, University of Chinese Academy of Sciences (UCAS)(中国科学院大学人工智能学院) ; Institute of Automation, Chinese Academy of Sciences(CAS)(中国科学院自动化研究所) ; Zhongguancun Academy(中关村学院) ; University of Chinese Academy of Sciences(中国科学院大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
机构 * Robotics and Perception Group, University of Zurich, Switzerland(苏黎世大学机器人与感知组)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments CVPR 2025 Workshop on Event-based Vision
专题命中 视频多模态 :multi-modal(abstract);分类 cs.AI
机构 * Monash University(蒙纳士大学) ; MBZUAI ; XJTLU(西安交通大学) ; Shanghai Jiaotong University(上海交通大学) ; Fudan University(复旦大学) ; University of Minnesota(明尼苏达大学) ; Cornell University(康奈尔大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Clarification note for the CVPR 2025 paper (FarSight). Prepared by a subset of the original authors; remaining co-authors are acknowledged in the text
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
Comments 14 pages
机构 * aiMotive Budapest, Hungary(aiMotive布达佩斯,匈牙利)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
机构 * CUHK(香港中文大学) ; HKU(香港大学) ; PolyU ; Peking University(北京大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
Comments 19 pages, 13 figures, Website: https://Perceive-Anything.github.io
机构 * University of Florence(佛罗伦萨大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
机构 * Hangzhou City University(杭州城市大学) ; University of Wisconsin–Madison(威斯康星大学麦迪逊分校) ; Hangzhou Normal University(杭州师范大学) ; Zhejiang Provincial Engineering Research Center for Real-Time SmartTech in Urban Security Governance(浙江省实时智能安防技术工程研究中心)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
机构 * Carnegie Mellon University(卡内基梅隆大学) ; Massachusetts Institute of Technology(麻省理工学院)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments IEEE FG 2025, \c{opyright} 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work
机构 * School of Computer Science and Technology, Chongqing University of Posts and Telecommunications(重庆邮电大学计算机科学与技术学院) ; Chongqing Institute for Brain and Intelligence(重庆脑科学与智能技术研究院) ; Guangyang Bay Laboratory(广阳湾实验室)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Accepted by IEEE TIFS
机构 * The University of Tokyo(东京大学) ; Adobe Research(Adobe研究)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
Comments 14 pages,10 figures
机构 * BUAA(北京航空航天大学) ; NUS(国立大学新加坡) ; Meituan(美团)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV