arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

多模态信息融合

面向图像、视频、多传感器和跨模态感知的信息融合,包括 Image Fusion、红外可见光、遥感、医学影像、LiDAR/雷达/相机和音视频融合。

共收录 259 信号源:cs.CV, eess.IV, eess.SP, cs.RO, cs.MM

1. 自动驾驶多传感器融合 259 篇

2507.10376 2025-07-15 cs.RO 57%

Raci-Net: Ego-vehicle Odometry Estimation in Adverse Weather Conditions

Mohammadhossein Talebi, Pragyan Dahal, Davide Possenti, Stefano Arrigoni, Francesco Braghin

机构 * Politecnico di Milano(米兰理工大学) Michigan State University(密歇根州立大学)

专题命中 自动驾驶多传感器融合 :sensor fusion(abstract);分类 cs.RO

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10715 2025-07-14 cs.CV 57%

EVT: Efficient View Transformation for Multi-Modal 3D Object Detection

Yongjin Lee, Hyeon-Mun Jeong, Yurim Jeon, Sanghyun Kim

机构 * ThorDrive Co., Ltd(ThorDrive公司) Seoul National University(首尔国立大学)

专题命中 自动驾驶多传感器融合 :sensor fusion(abstract);分类 cs.CV

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.07938 2025-07-11 cs.MM 57%

Multimodal Framework for Explainable Autonomous Driving: Integrating Video, Sensor, and Textual Data for Enhanced Decision-Making and Transparency

Abolfazl Zarghani, Amirhossein Ebrahimi, Amir Malekesfandiari

专题命中 自动驾驶多传感器融合 :sensor fusion(abstract);分类 cs.MM

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01484 2025-07-03 cs.CV 57%

What Really Matters for Robust Multi-Sensor HD Map Construction?

Xiaoshuai Hao, Yuting Zhao, Yuheng Ji, Luanyuan Dai, Peng Hao, Dingzhe Li, Shuai Cheng, Rong Yin

机构 * Beijing Academy of Artificial Intelligence(北京人工智能研究院) Institute of Automation, Chinese Academy of Science(中国科学院自动化研究所) Nanjing University of Science and Technology(南京理工大学) Samsung R&D Institute China–Beijing(三星中国北京研发中心) China North Artificial Intelligent & Innovation Research Institute(中国北方人工智能与创新研究院) Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)

专题命中 自动驾驶多传感器融合 :multi-modal fusion(abstract);分类 cs.CV

Comments Accepted by IROS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.12696 2025-06-23 cs.CV 57%

Collaborative Perception Datasets for Autonomous Driving: A Review

Naibang Wang, Deyong Shang, Yan Gong, Xiaoxi Hu, Ziying Song, Lei Yang, Yuhan Huang, Xiaoyu Wang, Jianli Lu

机构 * School of Mechanical and Electrical Engineering, China University of Mining and Technology (Beijing)(中国矿业大学(北京)机械与电子工程学院) State Key Laboratory of Robotics and System, Harbin Institute of Technology(哈尔滨工业大学机器人系统国家重点实验室) State Key Laboratory of Intelligent Green Vehicle and Mobility, Tsinghua University(清华大学智能绿色车辆与移动系统国家重点实验室) School of Mechanical and Aerospace Engineering, Nanyang Technological University(南洋理工大学机械与航空航天工程学院) School of Mechatronics Engineering, Harbin Institute of Technology(哈尔滨工业大学机械电子工程学院) Department of Electronic & Electrical Engineering, University of Bath(巴斯大学电子与电气工程系)

专题命中 自动驾驶多传感器融合 :information fusion(abstract);分类 cs.CV

Comments 18pages, 7figures, journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21079 2025-05-28 cs.CV 57%

Uni3D-MoE: Scalable Multimodal 3D Scene Understanding via Mixture of Experts

Yue Zhang, Yingzhao Jian, Hehe Fan, Yi Yang, Roger Zimmermann

机构 * Zhejiang University(浙江大学) National University of Singapore(新加坡国立大学)

专题命中 自动驾驶多传感器融合 :multimodal fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15005 2025-05-22 cs.RO cs.SE 57%

UniSTPA: A Safety Analysis Framework for End-to-End Autonomous Driving

Hongrui Kou, Zhouhang Lyu, Ziyu Wang, Cheng Wang, Yuxin Zhang

机构 * National Key Laboratory of Automotive Chassis Integration and Bionics, Jilin University, Changchun, China(吉林大学汽车底盘集成与仿生国家级重点实验室) School of Engineering and Physical Sciences, Heriot‐Watt University, Edinburgh, United Kingdom(赫瑞-沃森大学工程与物理科学学院)

专题命中 自动驾驶多传感器融合 :sensor fusion(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.12693 2025-05-20 cs.CV 57%

TACOcc:Target-Adaptive Cross-Modal Fusion with Volume Rendering for 3D Semantic Occupancy

Luyao Lei, Shuo Xu, Yifan Bai, Xing Wei

机构 * School of Software Engineering(软件工程学院) Xi’an Jiaotong University(西安交通大学)

专题命中 自动驾驶多传感器融合 :multi-modal fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.09422 2025-05-15 cs.CV 57%

MoRAL: Motion-aware Multi-Frame 4D Radar and LiDAR Fusion for Robust 3D Object Detection

Xiangyuan Peng, Yu Wang, Miao Tang, Bierzynski Kay, Lorenzo Servadei, Robert Wille

机构 * Technical University of Munich(慕尼黑技术大学) Infineon Technologies AG(英飞凌科技AG) China University of Geosciences(中国地质大学)

专题命中 自动驾驶多传感器融合 :multi-modal fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.16368 2025-04-24 cs.CV 57%

Revisiting Radar Camera Alignment by Contrastive Learning for 3D Object Detection

Linhua Kong, Dongxia Chang, Lian Liu, Zisen Kong, Pengyuan Li, Yao Zhao

机构 * IEEE Publication Technology Department(IEEE出版技术部)

专题命中 自动驾驶多传感器融合 :radar camera fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08172 2025-04-14 cs.RO 57%

Enhanced Cooperative Perception Through Asynchronous Vehicle to Infrastructure Framework with Delay Mitigation for Connected and Automated Vehicles

Nithish Kumar Saravanan, Varun Jammula, Yezhou Yang, Jeffrey Wishart, Junfeng Zhao

专题命中 自动驾驶多传感器融合 :sensor fusion(abstract);分类 cs.RO

Comments 9 pages, 9 figures, This paper is under review of SAE Journal of Connected and Automated Vehicles

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03563 2025-04-07 cs.CV 57%

PF3Det: A Prompted Foundation Feature Assisted Visual LiDAR 3D Detector

Kaidong Li, Tianxiao Zhang, Kuan-Chuan Peng, Guanghui Wang

专题命中 自动驾驶多传感器融合 :multi-modal fusion(abstract);分类 cs.CV

Comments This paper is accepted to the CVPR 2025 Workshop on Distillation of Foundation Models for Autonomous Driving (WDFM-AD)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20614 2025-03-27 eess.IV 57%

SaViD: Spectravista Aesthetic Vision Integration for Robust and Discerning 3D Object Detection in Challenging Environments

Tanmoy Dam, Sanjay Bhargav Dharavath, Sameer Alam, Nimrod Lilith, Aniruddha Maiti, Supriyo Chakraborty, Mir Feroskhan

专题命中 自动驾驶多传感器融合 :sensor fusion(abstract);分类 eess.IV

Comments This paper has been accepted for ICRA 2025, and copyright will automatically transfer to IEEE upon its availability on the IEEE portal

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08092 2025-03-12 cs.CV 57%

SparseVoxFormer: Sparse Voxel-based Transformer for Multi-modal 3D Object Detection

Hyeongseok Son, Jia He, Seung-In Park, Ying Min, Yunhao Zhang, ByungIn Yoo

专题命中 自动驾驶多传感器融合 :multi-modal fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.00299 2025-03-07 cs.CV 57%

GSPR: Multimodal Place Recognition Using 3D Gaussian Splatting for Autonomous Driving

Zhangshuo Qi, Junyi Ma, Jingyi Xu, Zijie Zhou, Luqi Cheng, Guangming Xiong

专题命中 自动驾驶多传感器融合 :multimodal fusion(abstract);分类 cs.CV

Comments 8 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.03689 2025-03-06 cs.CV 57%

DualDiff+: Dual-Branch Diffusion for High-Fidelity Video Generation with Reward Guidance

Zhao Yang, Zezhong Qian, Xiaofan Li, Weixiang Xu, Gongpeng Zhao, Ruohong Yu, Lingsi Zhu, Longjun Liu

专题命中 自动驾驶多传感器融合 :multimodal fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20111 2025-02-28 cs.CV cs.AI 57%

MITracker: Multi-View Integration for Visual Object Tracking

Mengjie Xu, Yitao Zhu, Haotian Jiang, Jiaming Li, Zhenrong Shen, Sheng Wang, Haolin Huang, Xinyu Wang, Qing Yang, Han Zhang, Qian Wang

专题命中 自动驾驶多传感器融合 :information fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.05075 2025-02-24 cs.CV 57%

DeepInteraction++: Multi-Modality Interaction for Autonomous Driving

Zeyu Yang, Nan Song, Wei Li, Xiatian Zhu, Li Zhang, Philip H. S. Torr

专题命中 自动驾驶多传感器融合 :multi-modal fusion(abstract);分类 cs.CV

Comments Journal extension of NeurIPS 2022. arXiv admin note: text overlap with arXiv:2208.11112

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.05847 2025-02-14 cs.RO cs.AI 57%

Federated Data-Driven Kalman Filtering for State Estimation

Nikos Piperigkos, Alexandros Gkillas, Christos Anagnostopoulos, Aris S. Lalos

专题命中 自动驾驶多传感器融合 :sensor fusion(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.10580 2025-02-13 cs.CV 57%

Normal Transformer: Extracting Surface Geometry from LiDAR Points Enhanced by Visual Semantics

Ancheng Lin, Jun Li, Yusheng Xiang, Wei Bian, Mukesh Prasad

专题命中 自动驾驶多传感器融合 :information fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.02858 2025-01-07 cs.CV 57%

A Novel Vision Transformer for Camera-LiDAR Fusion based Traffic Object Segmentation

Toomas Tahves, Junyi Gu, Mauro Bellone, Raivo Sell

专题命中 自动驾驶多传感器融合 :multimodal fusion(abstract);分类 cs.CV

Comments International Conference on Agents and Artificial Intelligence 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13454 2024-12-19 cs.CV cs.AI 57%

Pre-training a Density-Aware Pose Transformer for Robust LiDAR-based 3D Human Pose Estimation

Xiaoqi An, Lin Zhao, Chen Gong, Jun Li, Jian Yang

专题命中 自动驾驶多传感器融合 :multi-modal fusion(abstract);分类 cs.CV

Comments Accepted to AAAI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10438 2024-12-17 cs.CV cs.AI cs.LG 57%

Automatic Image Annotation for Mapped Features Detection

Maxime Noizet, Philippe Xu, Philippe Bonnifait

专题命中 自动驾驶多传感器融合 :multi-modal fusion(abstract);分类 cs.CV

Journal ref 2024 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2024), Oct 2024, Abu Dhabi, United Arab Emirates

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.06777 2024-12-10 cs.CV cs.AI cs.LG 57%

Driv3R: Learning Dense 4D Reconstruction for Autonomous Driving

Xin Fei, Wenzhao Zheng, Yueqi Duan, Wei Zhan, Masayoshi Tomizuka, Kurt Keutzer, Jiwen Lu

专题命中 自动驾驶多传感器融合 :sensor fusion(abstract);分类 cs.CV

Comments Code is available at: https://github.com/Barrybarry-Smith/Driv3R

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.05881 2024-11-12 cs.RO 57%

MIPD: A Multi-sensory Interactive Perception Dataset for Embodied Intelligent Driving

Zhiwei Li, Tingzhen Zhang, Meihua Zhou, Dandan Tang, Pengwei Zhang, Wenzhuo Liu, Qiaoning Yang, Tianyu Shen, Kunfeng Wang, Huaping Liu

专题命中 自动驾驶多传感器融合 :multi-modal fusion(abstract);分类 cs.RO

Comments Data, development kit and more details will be available at https://github.com/BUCT-IUSRC/Dataset MIPD

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.17734 2024-10-24 cs.CV cs.IR 57%

YOLO-Vehicle-Pro: A Cloud-Edge Collaborative Framework for Object Detection in Autonomous Driving under Adverse Weather Conditions

Xiguang Li, Jiafu Chen, Yunhe Sun, Na Lin, Ammar Hawbani, Liang Zhao

专题命中 自动驾驶多传感器融合 :multimodal fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07268 2024-10-11 cs.CV cs.AI 57%

Learning Content-Aware Multi-Modal Joint Input Pruning via Bird's-Eye-View Representation

Yuxin Li, Yiheng Li, Xulei Yang, Mengying Yu, Zihang Huang, Xiaojun Wu, Chai Kiat Yeo

专题命中 自动驾驶多传感器融合 :sensor fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.14170 2024-09-24 cs.CV 57%

LFP: Efficient and Accurate End-to-End Lane-Level Planning via Camera-LiDAR Fusion

Guoliang You, Xiaomeng Chu, Yifan Duan, Xingchen Li, Sha Zhang, Jianmin Ji, Yanyong Zhang

专题命中 自动驾驶多传感器融合 :sensor fusion(abstract);分类 cs.CV

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.12716 2024-09-20 cs.CV cs.AI 57%

Optical Flow Matters: an Empirical Comparative Study on Fusing Monocular Extracted Modalities for Better Steering

Fouad Makiyeh, Mark Bastourous, Anass Bairouk, Wei Xiao, Mirjana Maras, Tsun-Hsuan Wangb, Marc Blanchon, Ramin Hasani, Patrick Chareyre, Daniela Rus

专题命中 自动驾驶多传感器融合 :hybrid fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.04133 2024-09-09 cs.CV cs.CY 57%

Secure Traffic Sign Recognition: An Attention-Enabled Universal Image Inpainting Mechanism against Light Patch Attacks

Hangcheng Cao, Longzhi Yuan, Guowen Xu, Ziyang He, Zhengru Fang, Yuguang Fang

专题命中 自动驾驶多传感器融合 :image fusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏