arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

2026-05-26 至 2026-05-26 共收录 44 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 激光雷达 9 篇

2605.24753 2026-05-26 cs.CV 90%

Ghosts in the Point Clouds: De-glaring LiDAR in the Transient Domain

点云中的鬼影:瞬态域中的LiDAR去眩光

Avery Gump, Connor Henley, Sungjin Cheong, Akarsh Prabhakara, Mohit Gupta

机构 * University of Wisconsin–Madison(威斯康星大学麦迪逊分校)

专题命中 激光雷达 :LiDAR(title,title_cn);分类 cs.CV

AI总结 针对固态LiDAR内部多径眩光导致的伪影问题,提出基于瞬态眩光扩散函数(TGSF)的物理模型和无训练算法,在点云形成前抑制眩光,保留真实场景结构。

Comments CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.25293 2026-05-26 cs.CV cs.AI cs.RO 89%

Neuromorphic LiDAR-based Bird's Eye View Object Detection using Energy-efficient Spiking Neural Networks

基于神经形态激光雷达的鸟瞰图目标检测:使用节能脉冲神经网络

Sambit Mohapatra, Senthil Yogamani, Heinrich Gotzig, Patrick Mader

机构 * Valeo, Germany(德国瓦莱欧公司) Valeo, Ireland(爱尔兰瓦莱欧公司) TU Ilmenau, Germany(德国伊门豪大学)

专题命中 激光雷达 :LiDAR(title,abstract);BEV(abstract,abstract_cn);autonomous driving(abstract);driving perception(abstract)

AI总结 提出一种端到端脉冲编码器-解码器网络,用于激光雷达点云鸟瞰图表示中的目标检测,通过代理梯度反向传播训练,在KITTI基准上达到高精度,并实现3.33倍突触操作能耗降低。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.22809 2026-05-26 cs.CV 89%

Sensor2Sensor: Cross-Embodiment Sensor Conversion for Autonomous Driving

Sensor2Sensor: 自动驾驶的跨本体传感器转换

Jiahao Wang, Bo Sun, Yijing Bai, Vincent Casser, Songyou Peng, Zehao Zhu, Meng-Li Shih, Xander Masotto, Shih-Yang Su, Kanaad V Parvate, Tiancheng Ge, Linn Bieske, Dragomir Anguelov, Mingxing Tan, Chiyu Max Jiang

机构 * Waymo Johns Hopkins University(约翰霍普金斯大学) Google DeepMind(谷歌DeepMind) University of Washington(华盛顿大学)

专题命中 激光雷达 :LiDAR(summary_cn,abstract);autonomous driving(title,abstract);分类 cs.CV

AI总结 提出Sensor2Sensor生成模型,将单目行车记录仪视频转换为多模态传感器数据(多视角相机图像和LiDAR点云),通过4D高斯泼溅重建和扩散架构解决无配对数据问题,为自动驾驶开发解锁外部数据源。

Comments Accepted by CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24098 2026-05-26 cs.CV 87%

D2-V2X: Depth-Driven Cooperative V2X Reasoning for Autonomous Driving

D2-V2X: 面向自动驾驶的深度驱动协同V2X推理

Kevin Richard, Alphin Varghese, Colin Pham, David Oh, Srijan Das

机构 * University of North Carolina at Charlotte(北卡罗来纳州立大学)

专题命中 激光雷达 :LiDAR(summary_cn,abstract);autonomous driving(title);分类 cs.CV

AI总结 针对单车辆视觉语言模型受传感器遮挡限制的问题,提出D2-V2X基准和基线模型,通过融合3D LiDAR特征与VLM潜空间,利用链式思维推理实现遮挡目标识别和空间估计,在识别遮挡危险和降低空间估计误差上取得显著提升。

Comments Accepted to the DriveX Workshop at CVPR 2026 (Non-archival)

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24950 2026-05-26 cs.RO cs.LG 84%

ARCANE-PedSynth: Synthetic Multi-Pedestrian Datasets with Behavioural Crossing Annotations

ARCANE-PedSynth:具有行为穿越注释的合成多行人数据集

Muhammad Naveed Riaz, Maciej Wielgosz, Antonio M. López Peña

机构 * Computer Vision Center (CVC), Universitat Aut\` o noma de Barcelona (UAB), Bellaterra, Barcelona, Spain Institute of Electronics, Faculty of Computer Science, Electronics Telecommunications, AGH University of Krakow, Krak\' o w, Poland

专题命中 激光雷达 :LiDAR(summary_cn,abstract);autonomous driving(abstract);分类 cs.RO

AI总结 提出基于CARLA的开源框架ARCANE-PedSynth,通过混合AI-手动控制架构和12状态行为有限状态机生成高穿越率的多行人合成数据,支持RGB、LiDAR和DVS模态及行为标注,用于自动驾驶中的行人穿越预测。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24074 2026-05-26 cs.CV cs.RO 82%

WideDepth: Millimeter-Accurate Benchmark for Fisheye Depth Estimation

WideDepth: 用于鱼眼深度估计的毫米级精度基准

Ilia Indyk, Ignat Penshin, Ivan Sosin, Maxim Monastyrny, Aleksei Valenkov, Ilya Makarov

机构 * Robotics Center(机器人中心) AXXX Trusted AI Research Center, RAS(可信人工智能研究中心,俄罗斯科学院)

专题命中 激光雷达 :LiDAR(summary_cn,abstract);分类 cs.RO、cs.CV

AI总结 提出首个室内鱼眼深度估计数据集WideDepth,包含101个场景的5K高分辨率立体对和毫米级真值,并引入基于LiDAR的立体鱼眼图像生成方法,评估多种模型,微调后性能提升高达62%。

Comments Accepted to IEEE International Conference on Robotics and Automation (ICRA) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24495 2026-05-26 cs.RO 79%

Elevator-LIO: Robust LiDAR-Inertial Odometry for Multi-Floor Navigation under Elevator-Induced Non-Inertial Motion

Elevator-LIO:电梯引起的非惯性运动下多层导航的鲁棒激光雷达-惯性里程计

Yifan Zhang, Yudong Huang, Yuchong Zhang, Changze Li, Haoran Liu, Ming Yang, Tong Qin

机构 * Shanghai Jiao Tong University(上海交通大学)

专题命中 激光雷达 :LiDAR(title,abstract);分类 cs.RO

AI总结 提出Elevator-LIO框架,通过解耦状态估计模型和模式依赖的迭代误差状态卡尔曼滤波器,实现电梯内连续定位,并利用自适应体素降采样和事件触发更新抑制垂直漂移。

Comments 16 pages, 10 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.01215 2026-05-26 cs.RO 74%

Lidar Scan Registration Robust to Extreme Motions

对极端运动鲁棒的激光雷达扫描配准

Simon-Pierre Deschênes, Dominic Baril, Vladimír Kubelka, Philippe Giguère, François Pomerleau

机构 * Northern Robotics Laboratory(北方机器人实验室)

专题命中 激光雷达 :LiDAR(title);分类 cs.RO

AI总结 针对极端运动下点云畸变导致配准失败的问题,提出一种考虑轨迹运动不确定性和环境几何的去畸变方法,在200 m/s^2和800 rad/s^2的峰值加速度下,平移误差降低9.26%,旋转误差降低21.84%。

Comments 8 pages, 8 figures, published in 2021 18th Conference on Robots and Vision (CRV), Burnaby, Canada

Journal ref 2021 18th Conference on Robots and Vision (CRV), 2021, pp. 17-24

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 仿真评测 5 篇

2605.24004 2026-05-26 cs.AI cs.CV cs.LG cs.RO 82%

Reason--Imagine--Act: Closed-Loop LLM Decision Making with World Models for Autonomous Driving

推理--想象--行动:基于世界模型的闭环LLM自动驾驶决策

Zhengqi Sun, Yiwen Sun, Boxuan Liu, Tailai Chen, Tianxu Guo, Jiabin Liu

机构 * 1Department of Information Management, Peking University, Beijing 100871, China 2School of Intelligence Science Technology, Peking University, Beijing 100871, China 3State Key Laboratory of General Artificial Intelligence, BIGAI, Beijing 100080, China 4Yuanpei College, Peking University, Beijing 100871, China 5China Agricultural University, Beijing, China 6CRSC Research \& Design Institute Group Co., Ltd., Beijing, China

专题命中 仿真评测 :autonomous driving(title,abstract);分类 cs.RO、cs.CV、cs.AI

AI总结 提出Reason--Imagine--Act (RIA)闭环框架,结合LLM推理器与动作条件世界模型进行在线安全验证,在CARLA点目标协议下实现80.05%路线完成率、51.10%到达率和0.20%碰撞率。

Comments Accepted by the 2026 IEEE International Conference on Intelligent Transportation Systems (ITSC 2026). 8 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03134 2026-05-26 eess.SP cs.CV 77%

Controllable Radar Simulation with Waveform Parameter Embedding

具有波形参数嵌入的可控雷达仿真

Weiqing Xiao, Hao Huang, Chonghao Zhong, Yujie Lin, Nan Wang, Xiaoxue Chen, Zhaoxi Chen, Saining Zhang, Shuocheng Yang, Pierre Merriaux, Lei Lei, Hao Zhao

机构 * NJU(南京大学) BJTU(北京理工大学) BIT(北京理工大学) AIR, THU(空气科技,清华大学) NTU(国立台湾大学) SVM, THU(SVM,清华大学) Lightwheel AI LeddarTech

专题命中 仿真评测 :LiDAR(abstract,abstract_cn);autonomous driving(abstract);分类 cs.CV

AI总结 提出Ctrl-RS框架,通过环境反射张量、波形参数抽象和WARP-Net网络,实现可控的雷达立方体仿真,在2D/3D检测和语义分割任务中性能接近或超越真实雷达。

Comments CVPR 2026 Findings: Code: https://github.com/zhuxing0/SA-Radar Project page: https://zhuxing0.github.io/projects/SA-Radar

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.24037 2026-05-26 cs.CV cs.AI 62%

Mode-as-Sequence: Translating Multimodal Motion Prediction into Unified Sequential Mode Modeling

模式即序列:将多模态运动预测转化为统一序列模式建模

Zikang Zhou, Haibo Hu, Xinhong Chen, Yifan Zhang, Nan Guan, Yung-Hui Li, Chun Jason Xue, Jianping Wang

机构 * City University of Hong Kong(香港城市大学) City University of Hong Kong (Dongguan)(香港城市大学(东莞)) Hon Hai Research Institute(富士康研究学院) Mohamed bin Zayed University of Artificial Intelligence(莫莫丁·宾·扎耶德人工智能大学)

专题命中 仿真评测 :LiDAR(abstract);分类 cs.CV、cs.AI

AI总结 提出Mode-as-Sequence框架,将无序模式集转化为有序模式序列并显式建模模式间依赖,通过ModeSeq和Parallel ModeSeq两种实例化方法解决多模态运动预测中的模式坍塌和置信度排序问题,在Waymo数据集上取得领先性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.25947 2026-05-26 cs.CV 57%

A Pedestrian-Vehicle Interaction Benchmark and Annotation Framework for Unstructured Scenes via Uncalibrated Cameras

非标定相机下的非结构化场景行人-车辆交互基准与标注框架

Haoyang Peng, Qian Hu, Songan Zhang, Ming Yang

机构 * School of Automation and Intelligent Sensing(自动化与智能感知学院) Global Institute of Future Technology(未来技术全球研究院)

专题命中 仿真评测 :autonomous driving(abstract);分类 cs.CV

AI总结 针对非结构化场景中行人-车辆交互数据稀缺的问题,提出基于非标定监控视频的标注框架PINNS数据集,包含多国多场景的密集交互轨迹与场景信息,以促进复杂混合交通中的轨迹预测研究。

Comments 10 pages, 8 figures; project page available at https://github.com/Songan-Lab

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24876 2026-05-26 cs.CV cs.CL 57%

Agent-X: Evaluating Deep Multimodal Reasoning in Vision-Centric Agentic Tasks

Agent-X:评估视觉中心智能体任务中的深度多模态推理

Tajamul Ashraf, Amal Saqib, Hanan Ghani, Muhra AlMahri, Yuhao Li, Noor Ahsan, Umair Nawaz, Jean Lahoud, Hisham Cholakkal, Mubarak Shah, Philip Torr, Fahad Shahbaz Khan, Rao Muhammad Anwer, Salman Khan

机构 * Mohamed bin Zayed University of Artificial Intelligence(莫扎德·本·扎耶德人工智能大学) University of Central Florida(中央佛罗里达大学) University of Oxford(牛津大学)

专题命中 仿真评测 :autonomous driving(abstract);分类 cs.CV

AI总结 提出Agent-X基准,通过828个真实视觉任务和细粒度步骤评估框架,揭示当前模型在多步视觉推理中全链成功率低于50%的瓶颈。

Comments Accepted in International Conference of Learning Representations (ICLR 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他自动驾驶 1 篇

2605.25308 2026-05-26 cs.CV 57%

Stabilizing Streaming Video Geometry via Dynamic Feature Normalization

通过动态特征归一化稳定流视频几何

Xiaoyang Lyu, Muxin Liu, Xiaoshan Wu, Ruicheng Wang, Yi-Hua Huang, Yang-Tian Sun, Shaoshuai Shi, Xiaojuan Qi

机构 * The University of Hong Kong(香港大学) USTC(中国科学技术大学) Voyager Research, Didi Chuxing(滴滴出行 Voyager 研究)

专题命中 其他自动驾驶 :autonomous driving(abstract);分类 cs.CV

AI总结 针对流式RGB输入中单目几何模型的时间不一致问题(主要表现为尺度-偏移漂移),提出轻量级因果循环模块DyFN,通过动态调制特征统计量实现稳定几何估计,仅微调2%参数即可达到SOTA时间稳定性。

Comments 16 pages, 9 Figures, page: https://shawlyu.github.io/DyFN

详情

展开后加载摘要…

URL PDF HTML 收藏