arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

共收录 21140 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 6044 篇

2606.08680 2026-06-09 cs.CV cs.RO 新提交 92%

Distortion-Aware PETR for BEV Object Detection with Mixed Pinhole-Fisheye Cameras

畸变感知的PETR用于混合针孔-鱼眼相机的BEV目标检测

Xiangzhong Liu

机构 * fortiss GmbH(fortiss有限公司)

专题命中 感知 :BEV(title,title_cn);autonomous driving(abstract);driving perception(abstract);分类 cs.RO、cs.CV

AI总结 针对鱼眼相机径向畸变破坏BEV检测器均匀采样假设的问题,提出DAPETR,通过畸变感知位置编码和双向特征-几何协同调制模块,在KITTI-360基准上优于基线方法,并揭示了学习适应与显式几何重参数化之间的冲突。

Comments 8 pages, 5 figures, accepted at ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.06647 2026-06-02 cs.RO cs.AI cs.CV 91%

DeepIPCv2: LiDAR-powered Robust Environmental Perception and Navigational Control for Autonomous Vehicle

DeepIPCv2: 基于LiDAR的鲁棒环境感知与自动驾驶导航控制

Oskar Natan, Jun Miura

机构 * Department of Computer Science and Electronics, Universitas Gadjah Mada(计算机科学与电子系,加查马达大学) Department of Computer Science and Engineering, Toyohashi University of Technology(计算机科学与工程系,toyohashi技术大学)

专题命中 感知 :LiDAR(title,title_cn);autonomous driving(abstract);分类 cs.RO、cs.CV、cs.AI

AI总结 提出DeepIPCv2端到端自动驾驶框架,通过融合LiDAR点云分割与多视图投影构建鲁棒场景表示,结合门控循环单元、命令特定多层感知器和PID控制器实现路径点与导航控制命令的联合估计,在光照变化下取得最低总指标误差和最少驾驶干预。

Comments This work has been accepted for publication in IEEE Access. https://ieeexplore.ieee.org/document/11313052

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07626 2026-06-09 cs.CV cs.AI 新提交 91%

Eyes All Around: Design and Analysis of 360-Degree LiDAR Perception Using Equivariant Feature Learning in Unstructured Traffic

全方位视角:非结构化交通中基于等变特征学习的360度LiDAR感知设计与分析

Pranav Darshan, Raghuveer Narayanan Rajesh, M Uttara Kumari

机构 * RV College of Engineering(RV工程学院)

专题命中 感知 :LiDAR(title,title_cn);autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 针对非结构化城市交通中感知难题,提出结合扇形全景处理与旋转等变稀疏卷积的360度LiDAR感知框架,在印度城市交通数据集上验证了多类别检测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.12918 2026-05-27 cs.CV 91%

Radar-Camera BEV Multi-Task Learning with Cross-Task Attention Bridge for Joint 3D Detection and Segmentation

雷达-相机BEV多任务学习:用于联合3D检测与分割的跨任务注意力桥

Ahmet İnanç, Özgür Erkent

机构 * Hacettepe University(哈切特佩大学)

专题命中 感知 :BEV(title,title_cn);autonomous driving(abstract);分类 cs.CV

AI总结 提出CTAB(跨任务注意力桥)模块,通过共享BEV空间中的多尺度可变形注意力在检测和分割分支间交换特征,实现联合3D检测与分割的多任务学习,在nuScenes上提升分割性能且检测几乎不受影响。

Comments 8 pages, 5 figures, 3 Tables, Accepted at Radar in Robotics: New Frontiers workshop, at IEEE International Conference on Robotics & Automation (ICRA), 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.05245 2024-03-05 cs.RO 90%

Influence of Camera-LiDAR Configuration on 3D Object Detection for Autonomous Driving

Ye Li, Hanjiang Hu, Zuxin Liu, Xiaohao Xu, Xiaonan Huang, Ding Zhao

专题命中 感知 :autonomous driving(title,abstract);LiDAR(title,abstract);self-driving(abstract);driving perception(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.09629 2026-07-14 cs.CV cs.AI 版本更新 90%

4DR360: State Reasoning for Joint 3D Detection and Occupancy Prediction in 4D Radar-Camera Full-Scene Perception

4DR360:用于4D雷达-相机全场景感知中联合3D检测和占用预测的状态推理

Xiaokai Bai, Lianqing Zheng, Runwei Guan, Songkai Wang, Siyuan Cao, Hui-liang Shen

专题命中 感知 :BEV(summary_cn,abstract);occupancy(title,abstract);autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 针对4D雷达-相机全场景感知,提出\method框架,遵循跨模态状态推理范式,通过状态引导的BEV增强和多普勒引导的时间融合进行联合3D检测和占用预测,扩展数据集并实验,提升多任务学习效果。

Comments 5 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.24224 2026-07-28 cs.CV 新提交 90%

MATS: A novel multi-modality multi-task learning framework for 3D perception in autonomous driving

MATS:一种用于自动驾驶中3D感知的新型多模态多任务学习框架

Junchen Huo, Wanming Hao, Song Wang, Enqing Chen, Shouyi Yang, Guanghui Wang

机构 * School of Electrical and Information Engineering, Zhengzhou University(郑州大学电气与信息工程学院) Tianping College of Suzhou University of Science and Technology(苏州科技大学天平学院) Toronto Metropolitan University(多伦多都会大学)

专题命中 感知 :BEV(summary_cn,abstract);autonomous driving(title,abstract);LiDAR(abstract);分类 cs.CV

AI总结 针对自动驾驶3D感知,提出MATS多模态多任务学习框架,通过模态自适应BEV融合和特定任务MoE模块,在nuScenes基准测试中,多模态输入下显著优于现有技术,单任务也优于基线。

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.20752 2026-06-23 cs.CV cs.CR 新提交 90%

Mirage: a Clean-Label Backdoor against LiDAR 3D Object Detection

Mirage:针对LiDAR 3D目标检测的干净标签后门攻击

Ziba Parsons, Ang Li

机构 * University of Michigan - Dearborn(密歇根大学迪尔伯恩分校)

专题命中 感知 :LiDAR(title,title_cn);分类 cs.CV

AI总结 提出Mirage,一种黑盒、干净标签的后门攻击方法,通过注入少量标签一致的毒化样本,使LiDAR 3D目标检测模型学习触发器与目标类别的恶意关联,实现73%误分类成功率且仅需0.5%毒化率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.19122 2026-06-18 cs.RO 新提交 90%

Monocular 3D Occupancy Perception for Robots on Sidewalks via Hybrid 2D-3D Learning

基于混合2D-3D学习的人行道机器人单目3D占用感知

Yukai Ma, Joe Lin, Liu Liu, Honglin He, Lulu Ricketts, Brad Squicciarini, Yong Liu, Bolei Zhou

机构 * University of California, Los Angeles(加州大学洛杉矶分校) Zhejiang University(浙江大学) Coco Robotics(Coco机器人) Massachusetts Institute of Technology(麻省理工学院)

专题命中 感知 :LiDAR(summary_cn,abstract);occupancy(title,abstract);autonomous driving(abstract);分类 cs.RO

AI总结 提出WalkOCC框架,通过混合射线行进单目3D占用感知,结合LiDAR-RGB配对数据与大规模无配对单目图像学习,提升人行道机器人导航的预测精度和泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10535 2024-11-19 cs.RO cs.CV 90%

Advancing Autonomous Driving Perception: Analysis of Sensor Fusion and Computer Vision Techniques

Urvishkumar Bharti, Vikram Shahapur

专题命中 感知 :autonomous driving(title,abstract);driving perception(title,abstract);self-driving(abstract);分类 cs.RO、cs.CV

Comments 7 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.00349 2023-11-28 cs.CV cs.LG 89%

CALICO: Self-Supervised Camera-LiDAR Contrastive Pre-training for BEV Perception

Jiachen Sun, Haizhong Zheng, Qingzhao Zhang, Atul Prakash, Z. Morley Mao, Chaowei Xiao

专题命中 感知 :BEV(title,abstract);LiDAR(title,abstract);autonomous driving(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.00623 2022-12-02 cs.CV 89%

BEV-LGKD: A Unified LiDAR-Guided Knowledge Distillation Framework for BEV 3D Object Detection

Jianing Li, Ming Lu, Jiaming Liu, Yandong Guo, Li Du, Shanghang Zhang

专题命中 感知 :BEV(title,abstract);LiDAR(title,abstract);autonomous driving(abstract);分类 cs.CV

Comments 12pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.10780 2021-07-13 cs.CV 89%

BEVDetNet: Bird's Eye View LiDAR Point Cloud based Real-time 3D Object Detection for Autonomous Driving

Sambit Mohapatra, Senthil Yogamani, Heinrich Gotzig, Stefan Milz, Patrick Mader

专题命中 感知 :autonomous driving(title,abstract);LiDAR(title,abstract);BEV(abstract);分类 cs.CV

Comments Accepted for Oral Presentation at IEEE Intelligent Transportation Systems Conference (ITSC) 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.12464 2020-10-12 cs.CV cs.LG stat.ML 89%

End-to-end Autonomous Driving Perception with Sequential Latent Representation Learning

Jianyu Chen, Zhuo Xu, Masayoshi Tomizuka

专题命中 感知 :autonomous driving(title,abstract);driving perception(title,abstract);LiDAR(abstract);分类 cs.CV

Comments 8 pages, 10 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.16974 2020-07-01 cs.CR cs.CV cs.LG 89%

Towards Robust LiDAR-based Perception in Autonomous Driving: General Black-box Adversarial Sensor Attack and Countermeasures

Jiachen Sun, Yulong Cao, Qi Alfred Chen, Z. Morley Mao

专题命中 感知 :autonomous driving(title,abstract);LiDAR(title,abstract);self-driving(abstract);分类 cs.CV

Comments 18 pages, 27 figures, to be published in USENIX Security 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.10117 2026-05-12 cs.CV cs.AI 89%

Think as Needed: Geometry-Driven Adaptive Perception for Autonomous Driving

按需思考:基于几何的自适应感知用于自动驾驶

Donghyun Kim, Jaehyoung Park

机构 * Stony Brook University(史蒂文尼森布鲁克大学)

专题命中 感知 :LiDAR(summary_cn,abstract);autonomous driving(title,abstract);分类 cs.CV、cs.AI

AI总结 本文提出Enhanced HOPE架构,通过几何复杂度估计动态调整LiDAR帧处理路径,减少计算资源浪费,提升复杂场景下的感知能力与遮挡物体追踪性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.18940 2026-04-22 cs.CV cs.RO 89%

Localization-Guided Foreground Augmentation in Autonomous Driving

基于定位的前景增强在自动驾驶中

Jiawei Yong, Deyuan Qu, Qi Chen, Kentaro Oguchi, Shintaro Fukushima

机构 * Toyota Motor Corporation, Japan(日本电产公司) Toyota Motor North America, USA(美国电产北美公司)

专题命中 感知 :BEV(summary_cn,abstract);autonomous driving(title,abstract);分类 cs.RO、cs.CV

AI总结 本文提出LG-FA模块,通过在线增强几何上下文提升自动驾驶中的前景感知,提高BEV表示的几何完整性和时间稳定性,减少定位误差,并实现全局一致的车道和拓扑重建。

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.01471 2023-08-04 cs.CV cs.AI cs.LG cs.RO 89%

Implicit Occupancy Flow Fields for Perception and Prediction in Self-Driving

Ben Agro, Quinlan Sykora, Sergio Casas, Raquel Urtasun

专题命中 感知 :self-driving(title,abstract);occupancy(title,abstract);分类 cs.RO、cs.CV、cs.AI

Comments 19 pages, 13 figures

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2023, pp. 1379-1388

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.00493 2023-01-03 cs.CV cs.AI cs.LG cs.RO 89%

Argoverse 2: Next Generation Datasets for Self-Driving Perception and Forecasting

Benjamin Wilson, William Qi, Tanmay Agarwal, John Lambert, Jagjeet Singh, Siddhesh Khandelwal, Bowen Pan, Ratnesh Kumar, Andrew Hartnett, Jhony Kaesemodel Pontes, Deva Ramanan, Peter Carr, James Hays

专题命中 感知 :self-driving(title,abstract);driving perception(title);LiDAR(abstract);分类 cs.RO、cs.CV、cs.AI

Comments Proceedings of the Neural Information Processing Systems Track on Datasets and Benchmarks

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.00373 2022-05-05 cs.RO cs.AI cs.CV cs.LG 89%

Investigating the Impact of Multi-LiDAR Placement on Object Detection for Autonomous Driving

Hanjiang Hu, Zuxin Liu, Sharad Chitlangia, Akhil Agnihotri, Ding Zhao

专题命中 感知 :LiDAR(title,abstract);autonomous driving(title);self-driving(abstract);分类 cs.RO、cs.CV、cs.AI

Comments CVPR 2022 camera-ready version:15 pages, 14 figures, 9 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.00208 2019-07-18 cs.CV cs.AI cs.LG cs.RO 89%

RGB and LiDAR fusion based 3D Semantic Segmentation for Autonomous Driving

Khaled El Madawy, Hazem Rashed, Ahmad El Sallab, Omar Nasr, Hanan Kamel, Senthil Yogamani

专题命中 感知 :autonomous driving(title,abstract);LiDAR(title,abstract);分类 cs.RO、cs.CV、cs.AI

Comments Accepted for Oral Presentation at IEEE Intelligent Transportation Systems Conference (ITSC) 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.14710 2026-07-17 cs.CV cs.SY eess.SY 新提交 89%

Variational Inference for Bird's Eye View Segmentation in Autonomous Driving

自动驾驶中鸟瞰视角分割的变分推理

Jingyue Shi, Huaicheng Li, Junhui Zhao, Yanxiang Jiang

机构 * School of Electronic and Information Engineering, Beijing Jiaotong University(北京交通大学电子信息工程学院) School of Information Science and Engineering, Southeast University(东南大学信息科学与工程学院)

专题命中 感知 :BEV(summary_cn,abstract);autonomous driving(title,abstract);分类 cs.CV

AI总结 针对自动驾驶中鸟瞰视角分割难题,提出基于变压器的变分流变换网络TVB,通过后验BEV监督学习映射,结合条件变分自编码器、归一化流及注意力融合模块,在多摄像头视图BEV分割等方面性能优越。

Comments 13 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.25652 2026-06-25 cs.CV 新提交 89%

Auto-Labelling-Based Domain Transfer for 3D Object Detection on a Bicycle-Mounted LiDAR Platform

基于自动标注的域迁移用于自行车搭载LiDAR平台上的3D目标检测

Mario Finkbeiner, Max A. Buettner, Kanak Mazumder, Fabian B. Flohr

机构 * Intelligent Vehicles Lab (IVL), Munich University of Applied Sciences(慕尼黑应用科学大学智能车辆实验室)

专题命中 感知 :LiDAR(title,title_cn);autonomous driving(abstract);分类 cs.CV

AI总结 针对自行车视角下3D标注数据稀缺的问题,提出利用自动标注流水线迁移车辆训练检测器,在FUSE-Bike数据集上微调后mAP提升23.4点,证明自动标注可替代人工标注。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.09143 2026-06-09 cs.CV 新提交 89%

CAMF-Det: Closure-Aware Multimodal Fusion for LiDAR-Camera 3D Object Detection on UAV Platforms

CAMF-Det: 面向无人机平台的激光雷达-相机闭合感知多模态融合3D目标检测

Yanze Jiang, Yanfeng Gu, Xian Li

机构 * School of Electronics and Information Engineering, Harbin Institute of Technology(哈尔滨工业大学电子与信息工程学院)

专题命中 感知 :BEV(summary_cn,abstract);LiDAR(title,abstract);分类 cs.CV

AI总结 针对无人机俯视场景中树冠遮挡导致的多模态信息退化问题,提出基于比尔-朗伯定律的闭合感知融合框架CAMF-Det,通过显式建模双模态遮挡强度并注入检测流程,在自建数据集上实现困难级别mAP_BEV提升9.43%和4.88%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.24331 2026-05-26 cs.CV 89%

Spatial-aware Vision Language Model for Autonomous Driving

面向自动驾驶的空间感知视觉语言模型

Weijie Wei, Zhipeng Luo, Ling Feng, Venice Erin Liong

机构 * Motional University of Amsterdam(阿姆斯特丹大学)

专题命中 感知 :LiDAR(summary_cn,abstract);autonomous driving(title,abstract);分类 cs.CV

AI总结 提出LVLDrive框架,通过融合LiDAR点云与视觉语言模型,利用渐进融合Q-Former和空间感知问答数据集,解决3D度量空间推理瓶颈,提升自动驾驶场景理解与决策可靠性。

Comments Accepted to CVPR AutoPilot Workshop 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.25405 2026-04-29 cs.CV cs.RO 88%

Leveraging Previous-Traversal Point Cloud Map Priors for Camera-Based 3D Object Detection and Tracking

利用先前遍历点云地图先验进行基于摄像头的3D物体检测与跟踪

Markus Käppeler, Özgün Çiçek, Yakov Miron, Abhinav Valada

机构 * Department of Computer Science, University of Freiburg(弗赖堡大学计算机科学系) Bosch Research, Robert Bosch GmbH(博世研究)

专题命中 感知 :LiDAR(summary_cn,abstract);BEV(abstract,abstract_cn);autonomous driving(abstract);分类 cs.RO、cs.CV

AI总结 本文提出DualViewMapDet框架,通过在线检索先前遍历生成的点云地图先验,提升无LiDAR情况下基于摄像头的3D物体检测与跟踪性能,通过双空间融合策略增强特征表示。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08302 2025-09-11 cs.RO cs.CV 88%

Foundation Models for Autonomous Driving Perception: A Survey Through Core Capabilities

Rajendramayavan Sathyam, Yueqi Li

机构 * Zoox Inc.(Zoox公司)

专题命中 感知 :autonomous driving(title,abstract);driving perception(title,abstract);分类 cs.RO、cs.CV

Comments 32 pages, 14 figures, accepted at IEEE Open Journal of Vehicular Technology (OJVT)

URL PDF HTML 收藏
2412.17226 2024-12-24 cs.CV cs.RO 88%

OLiDM: Object-aware LiDAR Diffusion Models for Autonomous Driving

Tianyi Yan, Junbo Yin, Xianpeng Lang, Ruigang Yang, Cheng-Zhong Xu, Jianbing Shen

专题命中 感知 :autonomous driving(title,abstract);LiDAR(title,abstract);分类 cs.RO、cs.CV

Comments AAAI 2025, https://yanty123.github.io/OLiDM

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.08850 2024-11-20 cs.CV cs.RO 88%

LiDAR-BEVMTN: Real-Time LiDAR Bird's-Eye View Multi-Task Perception Network for Autonomous Driving

Sambit Mohapatra, Senthil Yogamani, Varun Ravi Kumar, Stefan Milz, Heinrich Gotzig, Patrick Mäder

专题命中 感知 :autonomous driving(title,abstract);LiDAR(title,abstract);分类 cs.RO、cs.CV

Comments Accepted for publication at IEEE Transactions on Intelligent Transportation Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.17265 2024-10-23 cs.CV cs.AI 88%

Image-Guided Outdoor LiDAR Perception Quality Assessment for Autonomous Driving

Ce Zhang, Azim Eskandarian

专题命中 感知 :autonomous driving(title,abstract);LiDAR(title,abstract);分类 cs.CV、cs.AI

Comments This work has been submitted to the IEEE for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏