AccidentBench: Benchmarking Multimodal Understanding and Reasoning in Vehicle Accidents and Beyond
机构 * UC Berkeley(伯克利大学) ; Stanford(斯坦福大学) ; UCL(伦敦大学学院) ; Virginia Tech(弗吉尼亚理工学院) ; Nvidia(英伟达公司)
专题命中 视频多模态 :multimodal(title,abstract)
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * UC Berkeley(伯克利大学) ; Stanford(斯坦福大学) ; UCL(伦敦大学学院) ; Virginia Tech(弗吉尼亚理工学院) ; Nvidia(英伟达公司)
专题命中 视频多模态 :multimodal(title,abstract)
机构 * Institute of Neuroinformatics, University of Zurich(神经信息学研究所,苏黎世大学) ; ETH Zurich(苏黎世联邦理工学院) ; Digital Society Initiative, University of Zurich(数字社会倡议,苏黎世大学) ; Department of Geography(地理系)
专题命中 视频多模态 :multimodal(title,abstract)
Comments 8 pages
机构 * Arizona State University(亚利桑那州立大学) ; University of Miami(迈阿密大学) ; The University of Texas at Austin(德克萨斯大学奥斯汀分校)
专题命中 视频多模态 :multimodal(title,abstract)
专题命中 视频多模态 :multimodal(title,abstract)
专题命中 视频多模态 :multi-modal(title,abstract)
机构 * Max Planck ETH CLS(马克斯·普朗克-ETH CLS) ; German Research Foundation(德国研究基金会) ; Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所) ; ETH Zürich(苏黎世联邦理工学院) ; Institute for Data Science in Mechanical Engineering (DSME)(机械工程数据科学研究所) ; RWTH Aachen University(亚琛工业大学)
专题命中 视频多模态 :multi-modal(title,abstract)
机构 * Respiratory Disease AI Laboratory in Epidemic Intelligence and Applications of Medical Big Data Instruments, Macau University of Science and Technology(呼吸疾病人工智能实验室(流行病智能与医学大数据应用)) ; Faculty of Innovation Engineering, Macau University of Science and Technology(创新工程学院) ; Institute of Systems Engineering, Macau University of Science and Technology(系统工程研究所) ; School of Business, Macau University of Science and Technology(商学院) ; State Key Laboratory of Respiratory Disease, National Clinical Research Center for Respiratory Disease, Guangzhou Institute of Respiratory Health, The First Affiliated Hospital of Guangzhou Medical University(呼吸疾病国家重点实验室、呼吸疾病临床研究中心、广州呼吸健康研究院、广州医学院第一附属医院) ; Guangzhou National Laboratory(广州国家实验室)
专题命中 视频多模态 :multi-modal(title,abstract)
机构 * Zhejiang University(浙江大学)
专题命中 视频多模态 :multimodal(title);分类 cs.CV、cs.CL、cs.MM
Comments Accepted at EMNLP2025 Main
机构 * Laboratoire Informatique d'Avignon, Avignon University, France(阿维尼翁信息实验室,阿维尼翁大学,法国)
专题命中 视频多模态 :multimodal(title,abstract)
Comments Paper accepted at ICPRAM 2025
专题命中 视频多模态 :multi-modal(title,abstract)
专题命中 视频多模态 :multi-modal(title,abstract)
专题命中 视频多模态 :multimodal(title,abstract)
机构 * Department of Electrical and Electronic Engineering, Imperial College London, London SW7 2AZ, UK(帝国理工学院电子与电气工程系) ; Institute of Biomedical Engineering, Department of Engineering Science, University of Oxford, Oxford OX3 7DQ, UK(牛津大学生物医学工程研究所)
专题命中 视频多模态 :multimodal(title,abstract)
Comments 19 pages, 4 figures. *Correspondence: m.shi16@imperial.ac.uk. Accepted by the IUPESM World Congress on Medical Physics and Biomedical Engineering 2025
机构 * Zalando SE Berlin Germany(泽尔安多德国分公司) ; Zalando Switzerland AG Zürich Switzerland(泽尔安多瑞士分公司)
专题命中 视频多模态 :multimodal(title,abstract)
专题命中 视频多模态 :multimodal(title,abstract)
机构 * Shanghai Jiao Tong University(上海交通大学) ; Zhuoyu Technology, Co., Ltd.(珠海宇科技有限公司) ; Department of Electronic and Computer Engineering, Hong Kong University of Science and Technology(香港理工大学电子与计算机工程系)
专题命中 视频多模态 :multimodal(title,abstract)
专题命中 视频多模态 :multimodal(title,abstract)
Comments 17 pages, 9 figures
机构 * College of Cyber Security(网络安全学院) ; Jinan University(济南大学) ; Department of Electrical and Electronic Engineering(电子与电气工程系) ; Xi’an Jiaotong-Liverpool University(西安交通大学利物浦大学)
专题命中 视频多模态 :cross-modal(title,abstract)
Comments Submitted to the 2025 IEEE International Conference on Data Mining (ICDM)
专题命中 视频多模态 :multimodal(title,abstract)
机构 * Laboratory of Innovations in Transportation (LiTrans), Toronto Metropolitan University, Toronto, Canada(创新交通实验室(LiTrans),多伦多 Metropolitan 大学,多伦多,加拿大)
专题命中 视频多模态 :multi-modal(title,abstract)
机构 * Department of Computer Science, University of Copenhagen(计算机科学系,哥本哈根大学)
专题命中 视频多模态 :multimodal(title,abstract)
Comments Accepted to be presented at the 35th IEEE International Workshop on Machine Learning for Signal Processing (IEEE MLSP 2025). Source code available at https://github.com/HughYau/UniPhyNet
专题命中 视频多模态 :multimodal(title,abstract)
Comments 11 pages, 7 figures, 3 tables. This work has been submitted to the IEEE for possible publication
专题命中 视频多模态 :multi-modal(title,abstract)
Comments Accepted by ACMMM2025
机构 * Information Processing and Telecommunications Center, ETSI Telecomunicación, Universidad Politécnica de Madrid, Spain(信息处理与电信中心,电信工程学院,马德里理工大学,西班牙)
专题命中 视频多模态 :multimodal(title,abstract)
Comments 29 pages, 9 Figures
专题命中 视频多模态 :multimodal(title,abstract)
Comments 15 pages, 4 figures
Journal ref Annals of Physics 480 (2025) 170104
机构 * Nara Institute of Science and Technology(奈良科学技術大學) ; Kyushu University(九州大學)
专题命中 视频多模态 :multimodal(title,abstract)
专题命中 视频多模态 :multimodal(title,abstract)
Comments Accepted to WACV 2025
机构 * Westlake University(西湖大学) ; Zhejiang University(浙江大学)
专题命中 视频多模态 :multimodal(title,abstract)
专题命中 视频多模态 :multimodal(title,abstract)
专题命中 视频多模态 :multimodal(title,abstract)
Comments 29 pages, 9 tables, 2 figures, and