Multi-Modal Decouple and Recouple Network for Robust 3D Object Detection
多模态解耦与耦合网络用于抗干扰的3D目标检测
Rui Ding, Zhaonian Kuang, Yuzhe Ji, Meng Yang, Xinhu Zheng, Gang Hua
机构
*
State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(人机混合增强智能国家重点实验室,人工智能与机器人研究院,西安交通大学)
;
Intelligent Transportation Thrust of the Systems Hub, The Hong Kong University of Science and Technology (Guangzhou)(系统枢纽智能交通方向,香港科技大学(广州))
;
Multimodal Experiences Research Lab, Dolby Laboratories(多模态体验研究实验室,Dolby实验室)
Bandwidth-adaptive Cloud-Assisted 360-Degree 3D Perception for Autonomous Vehicles
带宽自适应的云辅助360度3D感知用于自动驾驶车辆
Faisal Hawladera, Rui Meireles, Gamal Elghazaly, Ana Aguiar, Raphaël Frank
机构
*
Interdisciplinary Centre for Security, Reliability, and Trust (SnT), University of Luxembourg, L-1855, Luxembourg(安全、可靠性与信任跨学科研究中心(SnT),卢森堡大学)
;
Computer Science Department, Vassar College, Poughkeepsie, NY 12604, USA(计算机科学系,瓦萨学院)
机构
*
College of Information Science and Electronic Engineering, Zhejiang University(浙江大学信息科学与电子工程学院)
;
School of Automotive Studies, Tongji University(同济大学汽车学院)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
Jinhua Institute of Zhejiang University(浙江大学金华研究院)
;
School of Information and Electrical Engineering, Hangzhou City University(杭州城市学院信息与电气工程学院)
;
Hangzhou City University Binjiang Innovation Center(杭州城市学院滨江创新中心)
CP-uniGuard: A Unified, Probability-Agnostic, and Adaptive Framework for Malicious Agent Detection and Defense in Multi-Agent Embodied Perception Systems
CP-uniGuard: 多智能体具身体验系统中恶意代理检测与防御的统一、概率无关和自适应框架
Senkang Hu, Yihang Tao, Guowen Xu, Xinyuan Qian, Yiqin Deng, Xianhao Chen, Sam Tak Wu Kwong, Yuguang Fang
机构
*
Hong Kong JC STEM Lab of Smart City and Department of Computer Science, City University of Hong Kong(香港JC STEM实验室及城市大学计算机科学系)
;
School of Computer Science and Engineering, University of Electronic Science and Technology of China(电子科技大学计算机科学与工程学院)
;
Department of Electrical and Electronic Engineering, The University of Hong Kong(香港大学电子与电气工程系)
;
School of Data Science, Lingnan University(岭南大学数据科学学院)
Unlocking Past Information: Temporal Embeddings in Cooperative Bird's Eye View Prediction
解锁过去信息:合作鸟瞰图预测中的时间嵌入
Dominik Rößle, Jeremias Gerner, Klaus Bogenberger, Daniel Cremers, Stefanie Schmidtner, Torsten Schön
机构
*
Department of Computer Science and AImotion Bavaria, Technische Hochschule Ingolstadt(计算机科学系和AImotion巴伐利亚,因戈尔施塔特技术大学)
;
Department of Electrical Engineering and AImotion Bavaria, Technische Hochschule Ingolstadt(电气工程系和AImotion巴伐利亚,因戈尔施塔特技术大学)
;
School of Engineering and Design, Technical University of Munich(工程与设计学院,慕尼黑技术大学)
;
School of Computation, Information and Technology, Technical University of Munich(计算、信息与技术学院,慕尼黑技术大学)
CommentsCopyright 2024 IEEE. This is the accepted version of the paper. In 2024 IEEE Intelligent Vehicles Symposium (IV), pp. 2220-2225. Official paper available at https://doi.org/10.1109/IV55156.2024.10588608
Journal refIEEE Intelligent Vehicles Symposium (IV), pp. 2220-2225, 2024
CommentsAccepted at WACV 2026 Proceedings (Oral), 5th Workshop on Image, Video, and Audio Quality Assessment in Computer Vision, with a focus on VLM and Diffusion Models
V2X-Radar: A Multi-modal Dataset with 4D Radar for Cooperative Perception
V2X-Radar: 一种包含4D雷达的多模态数据集用于协作感知
Lei Yang, Xinyu Zhang, Jun Li, Chen Wang, Jiaqi Ma, Zhiying Song, Tong Zhao, Ziying Song, Li Wang, Mo Zhou, Yang Shen, Kai Wu, Chen Lv
机构
*
School of Vehicle and Mobility, Tsinghua University(车辆与移动性学院,清华大学)
;
Nanyang Technological University(南洋理工大学)
;
CUMTB(中国交通车辆技术研究所)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
Beijing Jiaotong University(北京交通大学)
;
ByteDance(字节跳动)
SparseLaneSTP: Leveraging Spatio-Temporal Priors with Sparse Transformers for 3D Lane Detection
SparseLaneSTP: 利用稀疏变换器结合时空先验进行3D车道检测
Maximilian Pittner, Joel Janai, Mario Faigle, Alexandru Paul Condurache
机构
*
Bosch Mobility Solutions, Robert Bosch GmbH(博世移动解决方案,罗伯特·博世有限公司)
;
Institute of Neuro- and Bioinformatics, University of Lübeck(神经与生物医学研究所,吕贝克大学)
;
Institute for Signal Processing and System Theory, University of Stuttgart(信号处理与系统理论研究所,斯图加特大学)
SCAFusion: A Multimodal 3D Detection Framework for Small Object Detection in Lunar Surface Exploration
SCAFusion: 一种针对月球表面探索的小目标多模态3D检测框架
Xin Chen, Kang Luo, Yangyi Xiao, Hesheng Wang
机构
*
Department of Automation, Key Laboratory of System Control and Information Processing of Ministry of Education, Key Laboratory of Marine Intelligent Equipment and System of Ministry of Education, Shanghai Engineering Research Center of Intelligent Control and Management, Shanghai Jiao Tong University(自动化系、教育部系统控制与信息处理重点实验室、教育部海洋智能装备与系统重点实验室、上海智能控制与管理工程研究中心、上海交通大学)
StereoMV2D: A Sparse Temporal Stereo-Enhanced Framework for Robust Multi-View 3D Object Detection
StereoMV2D: 一种稀疏时间立体增强框架,用于鲁棒的多视角3D目标检测
Di Wu, Feng Yang, Wenhui Zhao, Jinwen Yu, Pan Liao, Benlian Xu, Dingwen Zhang
机构
*
the school of automation, Northwestern Polytechnical University(自动化学院,西北工业大学)
;
the school of electronic and information engineering, Suzhou University of Science and Technology(电子信息工程学院,苏州科技大学)
机构
*
School of Computer Science, Beijing Institute of Technology(北京理工大学计算机科学学院)
;
Robotics and Autonomous Driving Lab (RAL), Baidu Research(百度研究机器人与自动驾驶实验室)
;
State Key Laboratory of Internet of Things for Smart City, Department of Computer and Information Science, University of Macau(澳门大学智能城市物联网国家重点实验室,计算机与信息科学系)
Rethinking the Encoding and Annotating of 3D Bounding Box: Corner-Aware 3D Object Detection from Point Clouds
重新思考3D边界框的编码与标注:从点云中基于角点的3D目标检测
Qinghao Meng, Junbo Yin, Jianbing Shen, Yunde Jia
机构
*
School of Computer Science, Beijing Institute of Technology(计算机科学学院,北京理工大学)
;
Computer Science Program, Computer, Electrical and Mathematical Sciences and Engineering (CEMSE) Division, Center of Excellence for Smart Health, and Center of Excellence for Generative AI, King Abdullah University of Science and Technology (KAUST)(计算机科学项目,计算机、电气和数学科学与工程(CEMSE)部门,智能健康卓越中心,生成人工智能卓越中心,国王阿卜杜勒·阿齐兹大学科学与技术(KAUST))
;
State Key Laboratory of Internet of Things for Smart City, Department of Computer and Information Science, University of Macau(智慧城市物联网国家重点实验室,澳门大学计算机与信息科学系)
;
Guangdong Provincial Key Laboratory of Machine Perception and Intelligent Computing, Shenzhen MSU-BIT University(广东省机器感知与智能计算重点实验室,深圳MSU-BIT大学)
;
Beijing Key Laboratory of Intelligent Information Technology, School of Computer Science, Beijing Institute of Technology(北京智能信息技术重点实验室,计算机科学学院,北京理工大学)
CommentsI made an operational error. I intended to update the paper with Identifier arXiv:2502.15488, not submit a new paper with a different identifier. Therefore, I would like to withdraw the current submission and resubmit an updated version for Identifier arXiv:2502.15488
DGFusion: Dual-guided Fusion for Robust Multi-Modal 3D Object Detection
Feiyang Jia, Caiyan Jia, Ailin Liu, Shaoqing Xu, Qiming Xia, Lin Liu, Lei Yang, Yan Gong, Ziying Song
机构
*
School of Computer Science and Technology, Beijing Key Laboratory of Traffic Data Mining and Embodied Intelligence, Beijing Jiaotong University(计算机科学与技术学院、交通数据挖掘与具身智能北京市重点实验室、北京交通大学)
;
State Key Laboratory of Internet of Things for Smart City and Department of Electrome chanical Engineering, University of Macau(智能城市物联网国家重点实验室、澳门大学机电工程系)
;
Fujian Key Laboratory of Sensing and Computing for Smart Cities, Xiamen University(智能城市感知与计算福建省重点实验室、厦门大学)
;
School of Mechanical and Aerospace Engineering, Nanyang Technological University(机械与航空航天工程学院、南洋理工大学)
;
State Key Laboratory of Robotics and System, Harbin Institute of Technology(机器人系统国家重点实验室、哈尔滨工业大学)