V-SenseDrive: A Privacy-Preserving Road Video and In-Vehicle Sensor Fusion Framework for Road Safety & Driver Behaviour Modelling
Muhammad Naveed, Nazia Perwaiz, Sidra Sultana, Mohaira Ahmad, Muhammad Moazam Fraz
机构
*
School of Electrical Engineering and Computer Science (SEECS), National University of Sciences and Technology (NUST)(电气工程与计算机科学学院(SEECS),国立科学与技术大学(NUST))
Mitigating Query Selection Bias in Referring Video Object Segmentation
Dingwei Zhang, Dong Zhang, Jinhui Tang
机构
*
Nanjing University of Science and Technology(南京理工大学)
;
The Hong Kong University of Science and Technology(香港科学大学)
;
Nanjing Forestry University(南京林业大学)
Large Language Models for Crash Detection in Video: A Survey of Methods, Datasets, and Challenges
Sanjeda Akter, Ibne Farabi Shihab, Anuj Sharma
机构
*
Department of Computer Science, Iowa State University(计算机科学系,爱荷华州立大学)
;
Department of Civil, Construction and Environmental Engineering, Iowa State University(土木、建设与环境工程系,爱荷华州立大学)
Jingwei Liu, Ling Yang, Hao Luo, Fan Wang, Hongyan Li, Mengdi Wang
机构
*
School of Intelligence Science and Technology, Peking University(北京理工大学智能科学与技术学院)
;
DAMO Academy, Alibaba group(阿里巴巴集团大模型研究院)
;
Hupan Lab(虎扑实验室)
;
National Key Laboratory of General Artificial Intelligence, Peking University(北京人工智能 general artificial intelligence 国家重点实验室)
;
Department of Electrical and Computer Engineering, Princeton University(普林斯顿大学电气与计算机工程系)
Livia: An Emotion-Aware AR Companion Powered by Modular AI Agents and Progressive Memory Compression
Rui Xi, Xianghan Wang
专题命中
视频多模态
:multimodal(abstract);分类 cs.AI、cs.MM
CommentsAccepted to the Proceedings of the 2025 International Conference on Artificial Intelligence and Virtual Reality (AIVR 2025). \c{opyright} 2025 Springer. This is the author-accepted manuscript. Rui Xi and Xianghan Wang contributed equally to this work. The final version will be available via SpringerLink
Enhancing Partially Relevant Video Retrieval with Robust Alignment Learning
Long Zhang, Peipei Song, Jianfeng Dong, Kun Li, Xun Yang
机构
*
University of Science and Technology of China(中国科学技术大学)
;
MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition, University of Science and Technology of China(中国科学技术大学脑启发式智能感知与认知实验室)
;
Zhejiang Gongshang University(浙江工商大学)
;
ReLER, CCAI, Zhejiang University(ReLER,中国计算机学会,浙江大学)
Boosting Temporal Sentence Grounding via Causal Inference
Kefan Tang, Lihuo He, Jisheng Dang, Xinbo Gao
机构
*
School of Electronic Engineering, Xidian University Xi'an China
;
School of Information Science \& Engineering, Lanzhou University Lanzhou China
;
Xidian University
;
Lanzhou University
MetaOcc: Spatio-Temporal Fusion of Surround-View 4D Radar and Camera for 3D Occupancy Prediction with Dual Training Strategies
Long Yang, Lianqing Zheng, Wenjin Ai, Minghao Liu, Sen Li, Qunshu Lin, Shengyu Yan, Jie Bai, Zhixiong Ma, Tao Huang, Xichan Zhu
机构
*
School of Automotive Studies, Tongji University(同济大学汽车学院)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
School of Automobile, Chang'an University(长安大学汽车学院)
;
School of Information and Electrical Engineering, Hangzhou City University(杭州城市学院信息与电气工程学院)
;
College of Science and Engineering, James Cook University(詹姆斯库克大学科学与工程学院)
Punching Bag vs. Punching Person: Motion Transferability in Videos
Raiyaan Abdullah, Jared Claypoole, Michael Cogswell, Ajay Divakaran, Yogesh Rawat
机构
*
Center for Research in Computer Vision, University of Central Florida(计算机视觉研究中心,中央佛罗里达大学)
;
Center for Vision Technology, SRI International(视觉技术中心,SRI国际)
Hierarchical Sub-action Tree for Continuous Sign Language Recognition
Dejie Yang, Zhu Xu, Xinjie Gao, Yang Liu
机构
*
Wangxuan Institute of Computer Technology, Peking University, Beijing, China(王轩计算机技术研究所,北京大学,北京,中国)
;
State Key Laboratory of General Artificial Intelligence, Peking Universitys, Beijing, China(通用人工智能国家重点实验室,北京大学,北京,中国)
HLFormer: Enhancing Partially Relevant Video Retrieval with Hyperbolic Learning
Jun Li, Jinpeng Wang, Chaolei Tan, Niu Lian, Long Chen, Yaowei Wang, Min Zhang, Shu-Tao Xia, Bin Chen
机构
*
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))
;
Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院)
;
Research Center of Artificial Intelligence, Peng Cheng Laboratory(鹏城实验室人工智能研究中心)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
专题命中
视频多模态
:cross-modal(abstract);分类 cs.CV、cs.MM
CommentsAccepted by ICCV'25. 13 pages, 6 figures, 4 tables