arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

共收录 6049 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 6049 篇

2601.07119 2026-01-13 cs.DC cs.CV 83%

SC-MII: Infrastructure LiDAR-based 3D Object Detection on Edge Devices for Split Computing with Multiple Intermediate Outputs Integration

SC-MII:基于基础设施激光雷达的边缘设备上用于分割计算的多中间输出集成的3D目标检测

Taisuke Noguchi, Takayuki Nishio, Takuya Azumi

机构 * Graduate School of Science and Engineering(科学与工程研究生学校) Saitama University(上野大学) School of Engineering(工程学院) Institute of Science Tokyo(东京科学研究所)

专题命中 感知 :LiDAR(title,abstract);autonomous driving(abstract);分类 cs.CV

AI总结 SC-MII通过多激光雷达和边缘计算实现高效3D目标检测,降低延迟和能耗,提升隐私保护。

Comments 6 pages. This version includes minor lstlisting configuration adjustments for successful compilation. No changes to content or layout. Originally published at IEEE CCNC 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03519 2026-01-13 cs.RO 83%

A Vision-Language-Action Model with Visual Prompt for OFF-Road Autonomous Driving

一种具有视觉提示的视觉-语言-动作模型用于越野自动驾驶

Liangdong Zhang, Yiming Nie, Haoyang Li, Fanjie Kong, Baobao Zhang, Shunxin Huang, Kai Fu, Chen Min, Liang Xiao

专题命中 感知 :autonomous driving(title,abstract);trajectory planning(abstract);分类 cs.RO

AI总结 本文提出OFF-EMMA模型,通过视觉提示和COT-SC策略提升越野自动驾驶轨迹规划的准确性和鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19981 2025-12-17 cs.CV 83%

FutrTrack: A Camera-LiDAR Fusion Transformer for 3D Multiple Object Tracking

FutrTrack: 一种基于相机与激光雷达融合的变压器用于三维多目标跟踪

Martha Teiko Teye, Ori Maoz, Matthias Rottmann

机构 * University of Wuppertal(乌珀塔尔大学) Institute of Computer Science, Osnabrück University(计算机科学研究所,奥斯纳布鲁克大学)

专题命中 感知 :LiDAR(title,abstract);BEV(abstract);分类 cs.CV

AI总结 FutrTrack通过融合相机与激光雷达数据,利用基于变压器的跟踪框架,在多目标跟踪中实现了高效且鲁棒的跟踪性能。

Comments Accepted to VISAPP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12377 2025-12-16 cs.RO 83%

INDOOR-LiDAR: Bridging Simulation and Reality for Robot-Centric 360 degree Indoor LiDAR Perception -- A Robot-Centric Hybrid Dataset

INDOOR-LiDAR: 联接仿真与现实以提升机器人中心的360度室内LiDAR感知 -- 一个机器人中心的混合数据集

Haichuan Li, Changda Tian, Panos Trahanias, Tomi Westerlund

专题命中 感知 :LiDAR(title,abstract);BEV(abstract);分类 cs.RO

AI总结 INDOOR-LIDAR通过融合仿真与真实数据,为机器人感知研究提供了一个可扩展、真实且可重复的混合数据集,用于提升复杂室内环境中的感知能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09349 2025-12-11 cs.RO 83%

COVLM-RL: Critical Object-Oriented Reasoning for Autonomous Driving Using VLM-Guided Reinforcement Learning

COVLM-RL:基于VLM引导强化学习的自动驾驶中关键对象导向推理

Lin Li, Yuxin Cai, Jianwu Fang, Jianru Xue, Chen Lv

机构 * School of Mechanical and Aerospace Engineering, Nanyang Technological University(南洋理工大学机械与航空航天工程学院) National Key Laboratory of Human-Machine Hybrid Augmented Intelligence, National Engineering Research Center for Visual Information and Applications, Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(西安交通大学人机混合增强智能国家重点实验室、视觉信息与应用国家工程研究中心、人工智能与机器人研究院)

专题命中 感知 :autonomous driving(title,abstract);end-to-end driving(abstract);分类 cs.RO

AI总结 COVLM-RL通过结合关键对象导向推理与VLM引导强化学习,提升了自动驾驶系统的泛化能力和训练效率。

Comments 8 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05698 2025-12-08 cs.CV 83%

OWL: Unsupervised 3D Object Detection by Occupancy Guided Warm-up and Large Model Priors Reasoning

OWL:通过占用引导预热和大模型先验推理实现无监督3D物体检测

Xusheng Guo, Wanfa Zhang, Shijia Zhao, Qiming Xia, Xiaolong Xie, Mingming Wang, Hai Wu, Chenglu Wen

专题命中 感知 :occupancy(title,abstract);autonomous driving(abstract);分类 cs.CV

AI总结 OWL通过占用引导预热和大模型先验推理实现无监督3D物体检测,有效提升mAP性能15%以上。

Comments The 40th Annual AAAI Conference on Artificial Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.16848 2025-12-05 cs.CV 83%

A re-calibration method for object detection with multi-modal alignment bias in autonomous driving

面向自动驾驶的多模态对齐偏差校准方法

Zhihang Song, Dingyi Yao, Ruibo Ming, Lihui Peng, Danya Yao, Yi Zhang

机构 * Department of Automation Tsinghua University Beijing(自动化系清华大学北京)

专题命中 感知 :autonomous driving(title,abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文提出一种重新校准模型,通过语义分割和定制损失函数提升自动驾驶中多模态检测的鲁棒性和性能,应对校准偏差带来的影响。

Comments Accepted for publication in IST 2025. Official IEEE Xplore entry will be available once published

Journal ref 2025 IEEE International Conference on Imaging Systems and Techniques (IST)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13079 2025-11-19 cs.CV 83%

Decoupling Scene Perception and Ego Status: A Multi-Context Fusion Approach for Enhanced Generalization in End-to-End Autonomous Driving

Jiacheng Tang, Mingyue Feng, Jiachao Liu, Yaonong Wang, Jian Pu

专题命中 感知 :autonomous driving(title,abstract);BEV(abstract);分类 cs.CV

Comments Accepted to AAAI 2026 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12941 2025-11-18 cs.RO 83%

GUIDE: Gaussian Unified Instance Detection for Enhanced Obstacle Perception in Autonomous Driving

Chunyong Hu, Qi Luo, Jianyun Xu, Song Wang, Qiang Li, Sheng Yang

专题命中 感知 :autonomous driving(title,abstract);occupancy(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08374 2025-11-18 cs.CV 83%

InsFusion: Rethink Instance-level LiDAR-Camera Fusion for 3D Object Detection

Zhongyu Xia, Hansong Yang, Yongtao Wang

机构 * Wangxuan Institute of Computer Technology, Peking University(北京大学计算机技术研究院)

专题命中 感知 :LiDAR(title,abstract);autonomous driving(abstract);分类 cs.CV

Comments NeurIPS 2025 workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07925 2025-11-14 cs.CV 83%

HD$^2$-SSC: High-Dimension High-Density Semantic Scene Completion for Autonomous Driving

Zhiwen Yang, Yuxin Peng

专题命中 感知 :autonomous driving(title,abstract);occupancy(abstract);分类 cs.CV

Comments 10 pages, 6 figures, accepted by AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17685 2025-11-12 cs.CV 83%

FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving

Shuang Zeng, Xinyuan Chang, Mengwei Xie, Xinran Liu, Yifan Bai, Zheng Pan, Mu Xu, Xing Wei, Ning Guo

机构 * Xi’an Jiaotong University(西安交通大学) Amap, Alibaba Group(阿里巴巴集团) DAMO Academy, Alibaba Group(阿里巴巴达摩院)

专题命中 感知 :autonomous driving(title,abstract);end-to-end driving(abstract);分类 cs.CV

Comments Accepted to NeurIPS 2025 as Spotlight Presentation. Code: https://github.com/MIV-XJTU/FSDrive

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16500 2025-10-27 cs.CV 83%

RLGF: Reinforcement Learning with Geometric Feedback for Autonomous Driving Video Generation

Tianyi Yan, Wencheng Han, Xia Zhou, Xueyang Zhang, Kun Zhan, Cheng-zhong Xu, Jianbing Shen

机构 * SKL-IOTSC, Computer and Information Science, University of Macau(澳门大学计算机与信息科学学院) Li Auto Inc(利汽车公司)

专题命中 感知 :autonomous driving(title,abstract);occupancy(abstract);分类 cs.CV

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00744 2025-10-23 cs.CV 83%

Rethinking Backbone Design for Lightweight 3D Object Detection in LiDAR

Adwait Chandorkar, Hasan Tercan, Tobias Meisen

机构 * Institute for TMDT, University of Wuppertal, Germany(TMDT研究所,乌尔姆大学)

专题命中 感知 :LiDAR(title,abstract);autonomous driving(abstract);分类 cs.CV

Comments Best Paper Award at the Embedded Vision Workshop ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.20893 2025-10-23 cs.CV 83%

Adversarial Attacks on LiDAR-Based Tracking Across Road Users: Robustness Evaluation and Target-Aware Black-Box Method

Shengjing Tian, Xiantong Zhao, Yuhao Bian, Yinan Han, Bin Liu

机构 * School of Economics and Management, China University of Mining and Technology(经济管理学院,中国矿业大学)

专题命中 感知 :LiDAR(title,abstract);BEV(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16500 2025-10-21 cs.RO 83%

Advancing Off-Road Autonomous Driving: The Large-Scale ORAD-3D Dataset and Comprehensive Benchmarks

Chen Min, Jilin Mei, Heng Zhai, Shuai Wang, Tong Sun, Fanjie Kong, Haoyang Li, Fangyuan Mao, Fuyang Liu, Shuo Wang, Yiming Nie, Qi Zhu, Liang Xiao, Dawei Zhao, Yu Hu

机构 * Research Center for Intelligent Computing Systems, SKLP, Institute of Computing Technology, Chinese Academy of Sciences, Beijing, China, 100190(中国科学院计算技术研究所,智能计算系统研究中心,SKLP,北京,中国,100190) Tongji University, Shanghai, China, 200092(同济大学,上海,中国,200092) Xi’an Jiaotong University, Shaanxi, China, 710049(西安交通大学,陕西,中国,710049) Nanchang University, Jiangxi, China, 330047(南昌大学,江西,中国,330047) Defense Innovation Institute, Beijing, China, 100073(国防科技创新院,北京,中国,100073)

专题命中 感知 :autonomous driving(title,abstract);occupancy(abstract);分类 cs.RO

Comments Off-road robotics

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18613 2025-09-24 cs.CV 83%

MLF-4DRCNet: Multi-Level Fusion with 4D Radar and Camera for 3D Object Detection in Autonomous Driving

Yuzhi Wu, Li Xiao, Jun Liu, Guangfeng Jiang, XiangGen Xia

机构 * MoE Key Laboratory of Brain-Inspired Intelligence Perception and Cognition, University of Science and Technology of China(脑启发智能感知与认知教育部重点实验室,中国科学技术大学) Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(人工智能研究院,合肥国家科学中心) Department of Electronic Engineering and Information Science, University of Science and Technology of China(电子工程与信息科学系,中国科学技术大学) Department of Electrical and Computer Engineering, University of Delaware(电气与计算机工程系,德克萨斯大学)

专题命中 感知 :autonomous driving(title,abstract);LiDAR(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16261 2025-09-23 cs.RO 83%

RaFD: Flow-Guided Radar Detection for Robust Autonomous Driving

Shuocheng Yang, Zikun Xu, Jiahao Wang, Shahid Nawaz, Jianqiang Wang, Shaobing Xu

机构 * School of Vehicle and Mobility(车辆与移动学院) Tsinghua University(清华大学)

专题命中 感知 :autonomous driving(title,abstract);BEV(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.01317 2025-09-03 cs.CV 83%

Guided Model-based LiDAR Super-Resolution for Resource-Efficient Automotive scene Segmentation

Alexandros Gkillas, Nikos Piperigkos, Aris S. Lalos

机构 * Industrial Systems Institute, Athena Research Center, Patras Science Park, Greece(工业系统研究所,亚特兰蒂斯研究中心,帕特拉科学公园,希腊) AviSense.AI, Patras Science Park, Greece(AviSense.AI,帕特拉科学公园,希腊)

专题命中 感知 :LiDAR(title,abstract);autonomous driving(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.16873 2025-08-20 cs.CV 83%

ContrastAlign: Toward Robust BEV Feature Alignment via Contrastive Learning for Multi-Modal 3D Object Detection

Ziying Song, Hongyu Pan, Feiyang Jia, Yongchang Zhang, Lin Liu, Lei Yang, Shaoqing Xu, Peiliang Wu, Caiyan Jia, Zheng Zhang, Yadan Luo

机构 * School of Computer Science & Technology, Beijing Key Laboratory of Traffic Data Mining and Embodied Intelligence, Beijing Jiaotong University(计算机科学与技术学院,北京交通大数据挖掘与具身智能重点实验室,北京交通大学) Horizon Robotics Nanyang Technological University(南洋理工大学) University of Macau(澳门大学) School of Information Science and Engineering, Yanshan University(信息科学与工程学院,燕山大学) School of Computer Science and Technology, Harbin Institute of Technology(计算机科学与技术学院,哈尔滨理工大学) School of Information Technology and Electrical Engineering, The University of Queensland(信息技术与电气工程学院,昆士兰大学)

专题命中 感知 :BEV(title,abstract);LiDAR(abstract);分类 cs.CV

Comments 12 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07838 2025-08-12 cs.CV 83%

CBDES MoE: Hierarchically Decoupled Mixture-of-Experts for Functional Modules in Autonomous Driving

Qi Xiang, Kunsong Shi, Zhigui Lin, Lei He

专题命中 感知 :autonomous driving(title,abstract);BEV(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06529 2025-08-12 cs.CV cs.LG 83%

RMT-PPAD: Real-time Multi-task Learning for Panoptic Perception in Autonomous Driving

Jiayuan Wang, Q. M. Jonathan Wu, Katsuya Suto, Ning Zhang

机构 * Department of Electrical and Computer Engineering, University of Windsor(电气与计算机工程系,温莎大学)

专题命中 感知 :autonomous driving(title,abstract);driving perception(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.18884 2025-08-01 cs.CV 83%

HV-BEV: Decoupling Horizontal and Vertical Feature Sampling for Multi-View 3D Object Detection

Di Wu, Feng Yang, Benlian Xu, Pan Liao, Wenhui Zhao, Dingwen Zhang

机构 * the school of automation, Northwestern Polytechnical University(自动化学院,西北工业大学)

专题命中 感知 :BEV(title,abstract);autonomous driving(abstract);分类 cs.CV

Comments 13 pages, 7 figures, submitted to T-ITS

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17122 2025-07-31 cs.CV 83%

R-LiViT: A LiDAR-Visual-Thermal Dataset Enabling Vulnerable Road User Focused Roadside Perception

Jonas Mirlach, Lei Wan, Andreas Wiedholz, Hannan Ejaz Keen, Andreas Eich

机构 * XITASO GmbH(XITASO公司) Karlsruhe Institute of Technology(卡尔斯鲁厄大学) LiangDao GmbH(Liandao公司)

专题命中 感知 :LiDAR(title,abstract);autonomous driving(abstract);分类 cs.CV

Comments 11 pages, 8 figures, accepted at ICCV2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20397 2025-07-29 cs.CV 83%

VESPA: Towards un(Human)supervised Open-World Pointcloud Labeling for Autonomous Driving

Levente Tempfli, Esteban Rivera, Markus Lienkamp

机构 * Technical University of Munich(慕尼黑技术大学)

专题命中 感知 :autonomous driving(title,abstract);LiDAR(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.13899 2025-07-22 cs.CV 83%

Enhancing LiDAR Point Features with Foundation Model Priors for 3D Object Detection

Yujian Mo, Yan Wu, Junqiao Zhao, Jijun Wang, Yinghao Hu, Jun Yan

机构 * School of Computer Science and Technology, Tongji University, Shanghai 201804, China(同济大学计算机科学与技术学院) School of Electronics and Information Engineering, Tongji University, Shanghai 201804, China(同济大学电子信息工程学院)

专题命中 感知 :LiDAR(title,abstract);autonomous driving(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.10298 2025-06-27 cs.CV 83%

ROA-BEV: 2D Region-Oriented Attention for BEV-based 3D Object Detection

Jiwei Chen, Yubao Sun, Laiyan Ding, Rui Huang

机构 * School of Science and Engineering, The Chinese University of Hong Kong, Shenzhen(科学与工程学院,香港中文大学(深圳)) Engineering Research Center of Digital Forensics, Ministry of Education, Nanjing University of Information Science and Technology(数字取证工程研究中心,教育部南京信息科学技术大学)

专题命中 感知 :BEV(title,abstract);autonomous driving(abstract);分类 cs.CV

Comments accepted by IROS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17958 2025-06-24 cs.CV 83%

ELMAR: Enhancing LiDAR Detection with 4D Radar Motion Awareness and Cross-modal Uncertainty

Xiangyuan Peng, Miao Tang, Huawei Sun, Bierzynski Kay, Lorenzo Servadei, Robert Wille

机构 * Infineon Technologies AG and the Technical University of Munich(英飞凌科技有限公司和技术大学慕尼黑) China University of Geosciences(中国地质大学) Technical University of Munich(技术大学慕尼黑)

专题命中 感知 :LiDAR(title,abstract);autonomous driving(abstract);分类 cs.CV

Comments 7 pages. Accepted by IROS2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.13864 2025-06-24 cs.CV 83%

UniDrive: Towards Universal Driving Perception Across Camera Configurations

Ye Li, Wenzhao Zheng, Xiaonan Huang, Kurt Keutzer

机构 * University of Michigan, Ann Arbor(密歇根大学) University of California, Berkeley(加州大学伯克利分校)

专题命中 感知 :driving perception(title,abstract);autonomous driving(abstract);分类 cs.CV

Comments ICLR 2025; 15 pages, 7 figures, 2 tables; Code at https://github.com/ywyeli/UniDrive

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16319 2025-06-23 cs.CV 83%

RealDriveSim: A Realistic Multi-Modal Multi-Task Synthetic Dataset for Autonomous Driving

Arpit Jadon, Haoran Wang, Phillip Thomas, Michael Stanley, S. Nathaniel Cibik, Rachel Laurat, Omar Maher, Lukas Hoyer, Ozan Unal, Dengxin Dai

机构 * German Aerospace Center(德国航空航天中心) Computer Vision Lab, Huawei Research Center Zurich(华为瑞士苏黎世研究中心计算机视觉实验室) Max Planck Institute for Informatics(马克斯·普朗克信息研究所) Parallel Domain(平行领域) ETH Zurich(苏黎世联邦理工学院)

专题命中 感知 :autonomous driving(title,abstract);LiDAR(abstract);分类 cs.CV

Comments Accepted at the IEEE Intelligent Vehicles Symposium (IV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏