arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

2026-03-04 至 2026-03-04 共收录 16 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 6 篇

2509.16654 2026-03-04 cs.CV 83%

Are VLMs Ready for Lane Topology Awareness in Autonomous Driving?

视觉-语言模型是否准备好在自动驾驶中具备车道拓扑意识?

Xin Chen, Jia He, Maozheng Li, Dongliang Xu, Tianyu Wang, Yixiao Chen, Zhixin Lin, Yue Yao

机构 * Shandong University(山东大学) MBZUAI Sems

专题命中 感知 :autonomous driving(title,abstract);BEV(abstract);分类 cs.CV

AI总结 本文评估了VLMs在自动驾驶中理解道路拓扑结构的能力,发现尽管闭源模型在部分任务表现良好,但空间推理仍是其主要瓶颈,且模型规模和推理令牌数量对性能有显著影响。

Comments 5 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02613 2026-03-04 cs.LG cs.RO 79%

Real-Time Generative Policy via Langevin-Guided Flow Matching for Autonomous Driving

通过拉格朗日引导的流匹配实现实时生成策略用于自动驾驶

Tianze Zhu, Yinuo Wang, Wenjun Zou, Tianyi Zhang, Likun Wang, Letian Tao, Feihong Zhang, Yao Lyu, Shengbo Eben Li

机构 * School of Vehicle and Mobility, Tsinghua University(车辆与移动系统学院,清华大学) College of Artificial Intelligence, Tsinghua University(人工智能学院,清华大学)

专题命中 感知 :autonomous driving(title,abstract);分类 cs.RO

AI总结 DACER-F通过流匹配实现实时生成策略,提升自动驾驶中的决策效率与性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02481 2026-03-04 cs.CV 70%

ModalPatch: A Plug-and-Play Module for Robust Multi-Modal 3D Object Detection under Modality Drop

ModalPatch: 一种插拔式模块,用于在模态缺失情况下鲁棒的多模态3D目标检测

Shuangzhi Li, Lei Ma, Xingyu Li

机构 * University of Alberta, Canada(阿尔伯塔大学) The University of Tokyo, Japan(东京大学)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 ModalPatch是一种插拔式模块,通过利用传感器数据的时间特性实现感知连续性,提升多模态3D目标检测在模态缺失情况下的鲁棒性和准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12567 2026-03-04 cs.CV 57%

GAN-Based Single-Stage Defense for Traffic Sign Classification Under Adversarial Patch

基于GAN的单阶段交通标志分类对抗攻击防御策略

Abyad Enan, Mashrur Chowdhury

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

AI总结 本文提出基于GAN的单阶段防御策略,有效提升自动驾驶系统对对抗性贴纸攻击的防御能力,显著提高交通标志分类的准确率。

Comments This work has been submitted to a peer-reviewed journal and is currently under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02560 2026-03-04 cs.CV 57%

CAWM-Mamba: A unified model for infrared-visible image fusion and compound adverse weather restoration

CAWM-Mamba:一种用于红外可见图像融合和复合恶劣天气恢复的统一模型

Huichun Liu, Xiaosong Li, Zhuangfan Huang, Tao Ye, Yang Liu, Haishu Tan

机构 * School of Physics(物理学院) Optoelectronic Engineering, Foshan University, Foshan 528225, China(光学电子工程学院,佛山大学,佛山528225,中国) Guangdong-HongKong-Macao Joint Laboratory for Intelligent Micro-Nano Optoelectronic Technology, Foshan 528225, China(粤港澳联合智能微纳光电子技术实验室,佛山528225,中国) School of Mechanical Electronic(机械电子学院) Information Engineering, China University of Mining and Technology, Beijing 100083, China(信息工程学院,中国矿业大学,北京100083,中国)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

AI总结 CAWM-Mamba提出了一种统一模型,用于红外可见图像融合和复合恶劣天气恢复,通过三个关键模块提升多退化场景下的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02532 2026-03-04 cs.CV 57%

EIMC: Efficient Instance-aware Multi-modal Collaborative Perception

EIMC: 高效实例感知多模态协作感知

Kang Yang, Peng Wang, Lantao Li, Tianci Bu, Chen Sun, Deying Li, Yongcai Wang

机构 * School of Information, Renmin University of China(中国人民大学信息学院) Sony Research and Development Center China(索尼(中国)研发有限公司) National University of Defense Technology(国防科技大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

AI总结 EIMC通过实例感知的多模态协作感知方法,提升自动驾驶安全性,减少带宽使用,实现高效且准确的3D感知。

Comments 9 pages, 8 figures, 7 tables

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 规划控制 2 篇

2603.02683 2026-03-04 cs.RO 79%

MMH-Planner: Multi-Mode Hybrid Trajectory Planning Method for UAV Efficient Flight Based on Real-Time Spatial Awareness

MMH-Planner: 多模式混合轨迹规划方法:基于实时空间感知的无人机高效飞行

Yinghao Zhao, Chenguang Dai, Liang Lyu, Zhenchao Zhang, Chaozhen Lan, Hong Xie

机构 * School of Surveying and Mapping, Information Engineering University(测绘学院,信息工程大学) School of Geodesy and Geomatics, Hubei Luojia Laboratory(地质测绘学院,湖北珞珈实验室)

专题命中 规划控制 :trajectory planning(title,abstract);分类 cs.RO

AI总结 MMH-Planner通过多模式混合轨迹规划方法,结合实时空间感知提升无人机飞行效率,优化规划效率与计算成本。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23316 2026-03-04 cs.RO cs.CV 62%

SceneStreamer: Continuous Scenario Generation as Next Token Group Prediction

SceneStreamer: 作为下一个token组预测的连续场景生成

Zhenghao Peng, Yuxin Liu, Bolei Zhou

机构 * University of California, Los Angeles(加州大学洛杉矶分校)

专题命中 规划控制 :autonomous driving(abstract);分类 cs.RO、cs.CV

AI总结 SceneStreamer通过自回归框架实现连续场景生成,支持动态代理群体的长期模拟,生成真实多样化的交通行为并提升强化学习策略的鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏

3. BEV与占用 4 篇

2603.02609 2026-03-04 cs.CV cs.RO 86%

VLMFusionOcc3D: VLM Assisted Multi-Modal 3D Semantic Occupancy Prediction

VLMFusionOcc3D: 基于VLM的多模态3D语义占位预测

A. Enes Doruk, Hasan F. Ates

机构 * Department of Artificial Intelligence and Data Engineering, Ozyegin University(人工智能与数据工程系,奥克辛大学)

专题命中 BEV与占用 :occupancy(title,abstract);autonomous driving(abstract);LiDAR(abstract);分类 cs.RO、cs.CV

AI总结 VLMFusionOcc3D通过融合视觉语言模型的语义先验,提升多模态3D语义占位预测的鲁棒性和适应性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22499 2026-03-04 cs.CV 83%

SABER: Spatially Consistent 3D Universal Adversarial Objects for BEV Detectors

SABER: 用于鸟瞰检测器的时空一致的3D通用对抗对象

Aixuan Li, Mochu Xiang, Bosen Hou, Zhexiong Wan, Jing Zhang, Yuchao Dai

机构 * School of Electronics and Information, Northwestern Polytechnical University & Shaanxi Key Laboratory of Information Acquisition and Processing, Xi’an, China(电子工程学院,西北工业大学 & 陕西省信息采集与处理重点实验室,西安,中国) School of Computing, Australian National University(计算学院,澳大利亚国立大学)

专题命中 BEV与占用 :BEV(title,abstract);autonomous driving(abstract);分类 cs.CV

AI总结 SABER提出了一种生成通用、非侵入式且3D一致的对抗对象的方法,以揭示BEV 3D目标检测器的漏洞,并提供评估自动驾驶系统鲁棒性的实用流程。

Comments Accepted to CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06770 2026-03-04 astro-ph.CO astro-ph.GA 50%

On the dependence of galaxy assembly bias on the selection criteria, number density, and redshift of galaxy samples

星系组装偏差对星系样本选择标准、数量密度和红移的依赖性

Sergio García-Moreno, Jonás Chaves-Montero

专题命中 BEV与占用 :occupancy(abstract)

AI总结 该研究探讨了银河团组装偏差对样本选择标准、数量密度和红移的依赖性,并提出了一种快速解析方法来预测银河团组装偏差。

Comments 9 pages, 8 figures, accepted by A&A

Journal ref A&A 706, A377 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01075 2026-03-04 hep-ex physics.acc-ph 50%

Beam-Beam Backgrounds for the Cool Copper Collider

铜冷却对撞机的束束背景

Dimitrios Ntounis, Laith Gordon, Lindsey Gray, Elias Mettner, Tim Barklow, Emilio A. Nanni, Caterina Vernieri

专题命中 BEV与占用 :occupancy(abstract)

AI总结 本文研究了铜冷却对撞机的束束背景,通过模拟评估了不同运行场景下的背景影响,并提出了一种灵活的背景研究工具包。

Comments 30 pages, 14 figures, 7 tables

Journal ref 2026 JINST 21 P02024

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 多传感器融合 2 篇

2603.02528 2026-03-04 cs.AI cs.RO 81%

LLM-MLFFN: Multi-Level Autonomous Driving Behavior Feature Fusion via Large Language Model

LLM-MLFFN: 通过大语言模型的多级自动驾驶行为特征融合

Xiangyu Li, Tianyi Wang, Xi Cheng, Rakesh Chowdary Machineni, Zhaomiao Guo, Sikai Chen, Junfeng Jiao, Christian Claudel

机构 * Department of Civil, Architectural, and Environmental Engineering, The University of Texas at Austin(德克萨斯大学奥斯汀分校土木、建筑与环境工程系) Systems Engineering Program, Cornell University(康奈尔大学系统工程项目) Department of Electrical and Computer Engineering, University of Michigan(密歇根大学电气与计算机工程系) Department of Civil and Environmental Engineering, University of Wisconsin-Madison(威斯康星大学麦迪逊分校土木与环境工程系) School of Architecture, The University of Texas at Austin(德克萨斯大学奥斯汀分校建筑学院)

专题命中 多传感器融合 :autonomous driving(title,abstract);分类 cs.RO、cs.AI

AI总结 LLM-MLFFN通过大语言模型的多级特征融合提升自动驾驶行为分类的准确性和鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01018 2026-03-04 cs.RO 57%

Integration of UWB Radar on Mobile Robots for Continuous Obstacle and Environment Mapping

将UWB雷达集成到移动机器人中用于连续障碍物和环境建图

Adelina Giurea, Stijn Luchie, Dieter Coppens, Jeroen Hoebeke, Eli De Poorter

机构 * Department of Information Technology, Ghent University - imec - IDLab(信息科技系,根特大学 - imec - IDLab)

专题命中 多传感器融合 :LiDAR(abstract);分类 cs.RO

AI总结 本文提出了一种基于UWB雷达的移动机器人环境建图方法,通过三步处理流程实现高精度障碍物检测,适用于无视觉特征和无固定锚点的未知环境。

Journal ref IEEE Access, vol. 14, pp. 25663-25676, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 仿真评测 1 篇

2603.02542 2026-03-04 cs.AI 57%

AnchorDrive: LLM Scenario Rollout with Anchor-Guided Diffusion Regeneration for Safety-Critical Scenario Generation

AnchorDrive: 基于锚点引导扩散再生的LLM场景回放用于安全关键场景生成

Zhulin Jiang, Zetao Li, Cheng Wang, Ziwen Wang, Chen Xiong

机构 * School of Intelligent Systems Engineering, Sun Yat-sen University, Shenzhen, China(中山大学智能系统工程学院,深圳,中国)

专题命中 仿真评测 :autonomous driving(abstract);分类 cs.AI

AI总结 AnchorDrive通过结合LLM的可控生成和扩散模型的真实轨迹生成,实现安全关键场景的高效合成,提升场景的真实性和可控性。

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 其他自动驾驶 1 篇

2603.03198 2026-03-04 cs.RO cs.CL cs.CV 62%

ACE-Brain-0: Spatial Intelligence as a Shared Scaffold for Universal Embodiments

ACE-Brain-0:空间智能作为通用具身化体系的共享框架

Ziyang Gong, Zehang Luo, Anke Tang, Zhe Liu, Shi Fu, Zhi Hou, Ganlin Yang, Weiyun Wang, Xiaofeng Wang, Jianbo Liu, Gen Luo, Haolan Kang, Shuang Luo, Yue Zhou, Yong Luo, Li Shen, Xiaosong Jia, Yao Mu, Xue Yang, Chunxiao Liu, Junchi Yan, Hengshuang Zhao, Dacheng Tao, Xiaogang Wang

机构 * Shanghai Jiao Tong University(上海交通大学) Nanyang Technological University(南洋理工大学) The Chinese University of Hong Kong(香港中文大学) The University of Hong Kong(香港大学) University of Science(科学技术大学) Fudan University(复旦大学) Xiamen University(厦门大学) East China Normal University(华东师范大学) Wuhan University(武汉大学) Sun Yat-sen University(中山大学)

专题命中 其他自动驾驶 :autonomous driving(abstract);分类 cs.RO、cs.CV

AI总结 ACE-Brain-0 通过空间智能作为共享框架,统一了自动驾驶、机器人和 UAVs 的具身化任务,采用 SSR 范式和 GRPO 方法实现跨领域泛化和领域精通的平衡。

Comments Code: https://github.com/ACE-BRAIN-Team/ACE-Brain-0 Hugging Face: https://huggingface.co/ACE-Brain/ACE-Brain-0-8B

详情

展开后加载摘要…

URL PDF HTML 收藏