arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

共收录 6063 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 6063 篇

2512.08247 2025-12-10 cs.CV cs.AI 62%

Distilling Future Temporal Knowledge with Masked Feature Reconstruction for 3D Object Detection

通过掩码特征重建进行未来时间知识蒸馏以实现3D目标检测

Haowen Zheng, Hu Zhu, Lu Deng, Weihao Gu, Yang Yang, Yanyan Liang

机构 * HAOMO.AI Technology(HAOMO.AI科技公司)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本文提出FTKD方法,通过掩码特征重建和未来引导logit蒸馏,提升3D目标检测中未来帧知识的转移效果,实现性能提升。

Comments AAAI-26

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17665 2025-12-08 cs.CV cs.RO 62%

Perspective-Invariant 3D Object Detection

视角不变的3D物体检测

Ao Liang, Lingdong Kong, Dongyue Lu, Youquan Liu, Jian Fang, Huaici Zhao, Wei Tsang Ooi

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.CV

AI总结 本文提出Pi3DET数据集和跨平台适应框架,实现视角不变的3D物体检测,推动非车辆平台的3D检测研究。

Comments ICCV 2025; 54 pages, 18 figures, 22 tables; Project Page at https://pi3det.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04039 2025-12-04 cs.CV cs.AI cs.LG 62%

Fast & Efficient Normalizing Flows and Applications of Image Generative Models

高效且快速的归一化流及图像生成模型的应用

Sandeep Nagar

机构 * International Institute of Information Technology (Deemed to be University)(国际信息技术学院(认定为大学)) IIIT Hyderabad(IIIT海得拉尔)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本研究提出高效归一化流及图像生成模型的应用,包括超分辨率、农业质量评估、地质制图、隐私保护及艺术修复等

Comments PhD Thesis

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.04096 2025-12-03 cs.RO cs.CV 62%

Image-Based Relocalization and Alignment for Long-Term Monitoring of Dynamic Underwater Environments

基于图像的重定位与对齐用于动态水下环境的长期监测

Beverley Gorry, Tobias Fischer, Michael Milford, Alejandro Fontan

机构 * Queensland University of Technology(昆士兰理工大学)

专题命中 感知 :BEV(abstract);分类 cs.RO、cs.CV

AI总结 本文提出了一种结合VPR、特征匹配和图像分割的方法,用于水下环境的长期监测,并引入了首个大规模水下VPR基准测试。

Journal ref Proceedings of the 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), Hangzhou, China, pp. 10749-10756, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01755 2025-12-02 cs.CV cs.RO 62%

3EED: Ground Everything Everywhere in 3D

3EED: 在三维中万物皆 grounded

Rong Li, Yuhao Dong, Tianshuai Hu, Ao Liang, Youquan Liu, Dongyue Lu, Liang Pan, Lingdong Kong, Junwei Liang, Ziwei Liu

机构 * WorldBench Team(WorldBench团队)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.CV

AI总结 3EED提出一个大规模多平台多模态三维 grounding 基准测试,通过提供丰富的户外场景数据和跨平台学习技术,推动语言驱动的三维具身感知研究。

Comments NeurIPS 2025 DB Track; 38 pages, 17 figures, 10 tables; Project Page at https://project-3eed.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.18016 2025-12-02 cs.RO cs.AI 62%

ADA-DPM: A Neural Descriptors-based Adaptive Noise Filtering Strategy for SLAM

ADA-DPM: 基于神经描述符的自适应噪声过滤策略用于SLAM

Yongxin Shao, Aihong Tan, Binrui Wang, Yinlian Jin, Licong Guan, Peng Liao

机构 * College of Metrology Measurement and Instrument(计量测量与仪器学院) College of Mechanical and Electrical Engineering(机械与电气工程学院)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.AI

AI总结 ADA-DPM通过动态分割、全局重要性评分和跨层图卷积模块,提升SLAM中动态物体干扰和噪声的鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.01791 2025-12-02 cs.CV cs.RO 62%

A Minimal Subset Approach for Informed Keyframe Sampling in Large-Scale SLAM

大规模SLAM中用于有信息关键帧采样的最小子集方法

Nikolaos Stathoulopoulos, Christoforos Kanellakis, George Nikolakopoulos

机构 * Robotics and AI Group, Department of Computer, Electrical and Space Engineering, Luleå University of Technology(机器人与人工智能组,计算机、电气与空间工程系,卢勒奥技术大学)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.CV

AI总结 本文提出了一种基于最小子集方法的在线关键帧采样技术,通过减少冗余和保留信息来提升大规模SLAM中的闭环检测性能和定位精度。

Comments Please cite the published version. 8 pages, 9 figures

Journal ref IEEE Robotics and Automation Letters, vol. 11, no. 1, pp. 738-745, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.23311 2025-12-01 cs.CV cs.AI cs.CL 62%

Toward Automatic Safe Driving Instruction: A Large-Scale Vision Language Model Approach

迈向自动安全驾驶指令:一种大规模视觉语言模型方法

Haruki Sakajo, Hiroshi Takato, Hiroshi Tsutsui, Komei Soda, Hidetaka Kamigaito, Taro Watanabe

机构 * Nara Institute of Science and Technology(奈良科学技术研究所) Teatis inc.(Teatis公司) Queensland university of technology(昆士兰理工大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本文提出了一种大规模视觉语言模型方法,用于生成安全驾驶指令,通过构建数据集并评估模型性能,展示了微调模型在自动驾驶安全中的应用与挑战。

Comments Accepted to MMLoSo 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18668 2025-11-25 cs.CV eess.IV 62%

Data Augmentation Strategies for Robust Lane Marking Detection

用于稳健车道标记检测的数据增强策略

Flora Lian, Dinh Quang Huynh, Hector Penades, J. Stephany Berrio Perez, Mao Shan, Stewart Worrall

机构 * The University of Sydney(悉尼大学) Australian Centre for Robotics(澳大利亚机器人中心)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、eess.IV

AI总结 本文提出基于生成AI的数据增强方法,提升车道标记检测在不同视角下的鲁棒性,通过几何变换、图像修复和车辆叠加模拟部署场景,提升模型在阴影等干扰下的性能。

Comments 8 figures, 2 tables, 10 pages, ACRA, Australasian conference on robotics and automation

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17612 2025-11-25 cs.CV cs.AI 62%

Unified Low-Light Traffic Image Enhancement via Multi-Stage Illumination Recovery and Adaptive Noise Suppression

通过多阶段照明恢复和自适应噪声抑制实现统一的低光照交通图像增强

Siddiqua Namrah

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 本文提出了一种无监督的多阶段深度学习框架,通过照明恢复和噪声抑制提升低光照交通图像质量,提高自动驾驶等系统的感知可靠性。

Comments Master's thesis, Korea University, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12638 2025-11-18 cs.CV cs.AI 62%

edgeVLM: Cloud-edge Collaborative Real-time VLM based on Context Transfer

Chen Qian, Xinran Yu, Zewen Huang, Danyang Li, Qiang Ma, Fan Dang, Xuan Ding, Guangyong Shang, Zheng Yang

机构 * Tsinghua University(清华大学) Beijing Jiaotong University(北京交通大学) Inspur Yunzhou Industrial Internet Co., Ltd(Inspur云洲工业互联网有限公司)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11845 2025-11-18 cs.RO cs.AI cs.AR 62%

Autonomous Underwater Cognitive System for Adaptive Navigation: A SLAM-Integrated Cognitive Architecture

K. A. I. N Jayarathne, R. M. N. M. Rathnayaka, D. P. S. S. Peiris

机构 * Department of Computational Mathematics University of Moratuwa(计算数学系大学莫图瓦)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.AI

Comments 6 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11676 2025-11-18 cs.LG cs.AI cs.CV 62%

Learning with Preserving for Continual Multitask Learning

Hanchen David Wang, Siwoo Bae, Zirong Chen, Meiyi Ma

机构 * Vanderbilt University(范德比大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

Comments 25 pages, 16 figures, accepted at AAAI-2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11838 2025-11-18 cs.CV cs.AI 62%

Probabilistic Robustness Analysis in High Dimensional Space: Application to Semantic Segmentation Network

Navid Hashemi, Samuel Sasaki, Diego Manzanas Lopez, Lars Lindemann, Ipek Oguz, Meiyi Ma, Taylor T. Johnson

机构 * Department of Computer Science, Vanderbilt University(计算机科学系,范德比尔特大学) Automatic Control Laboratory, ETH Zürich(自动控制实验室,苏黎世联邦理工学院)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08640 2025-11-13 cs.CV cs.AI 62%

Predict and Resist: Long-Term Accident Anticipation under Sensor Noise

Xingcheng Liu, Bin Rao, Yanchen Guan, Chengyue Wang, Haicheng Liao, Jiaxun Zhang, Chengyu Lin, Meixin Zhu, Zhenning Li

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

Comments accepted by the Fortieth AAAI Conference on Artificial Intelligence (AAAI-26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07950 2025-11-12 cs.RO cs.AI 62%

USV Obstacles Detection and Tracking in Marine Environments

Yara AlaaEldin, Enrico Simetti, Francesca Odone

机构 * DIBRIS - Department of Computer Science, Bioengineering, Robotics and System Engineering(DIBRIS-计算机科学、生物工程、机器人与系统工程系) University of Genova(热那亚大学)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07238 2025-11-11 cs.CV cs.AI 62%

Leveraging Text-Driven Semantic Variation for Robust OOD Segmentation

Seungheon Song, Jaekoo Lee

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

Comments 8 pages, 5 figure references, 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS) submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.00527 2025-11-11 eess.IV cs.CV 62%

MAROON: A Dataset for the Joint Characterization of Near-Field High-Resolution Radio-Frequency and Optical Depth Imaging Techniques

Vanessa Wirth, Johanna Bräunig, Nikolai Hofmann, Martin Vossiek, Tim Weyrich, Marc Stamminger

机构 * Friedrich-Alexander-Universität Erlangen-Nürnberg(埃朗根-纽伦堡弗里德里希-亚历山大大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、eess.IV

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.07843 2025-11-11 cs.CV cs.RO 62%

Real-time Multi-view Omnidirectional Depth Estimation for Real Scenarios based on Teacher-Student Learning with Unlabeled Data

Ming Li, Xiong Yang, Chaofan Wu, Jiaheng Li, Pinzhi Wang, Xuejiao Hu, Sidan Du, Yang Li

机构 * School of Artificial Intelligence/School of Future Technology, Nanjing University of Information Science and Technology(人工智能学院/未来技术学院,信息科学与技术大学) School of Electronic Science and Engineering, Nanjing University(电子科学与工程学院,南京大学) School of Computer Engineering, Jinling Institute of Technology(计算机工程学院,金陵科技学院) Suzhou High Technology Research Institute, Nanjing University(苏州高新技术研究院,南京大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.RO、cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05404 2025-11-10 cs.CV cs.AI 62%

Multi-modal Loop Closure Detection with Foundation Models in Severely Unstructured Environments

Laura Alejandra Encinar Gonzalez, John Folkesson, Rudolph Triebel, Riccardo Giubilato

专题命中 感知 :LiDAR(abstract);分类 cs.CV、cs.AI

Comments Under review for ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26358 2025-10-31 cs.RO cs.CV 62%

AgriGS-SLAM: Orchard Mapping Across Seasons via Multi-View Gaussian Splatting SLAM

Mirko Usuelli, David Rapado-Rincon, Gert Kootstra, Matteo Matteucci

机构 * Dipartimento di Bioingegneria, Elettronica e Informazione, Politecnico di Milano(生物工程、电子与信息系,米兰理工大学) Agricultural Biosystems Engineering, Wageningen University & Research(农业生物系统工程,瓦赫宁根大学与研究)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22669 2025-10-28 cs.CV cs.AI 62%

LVD-GS: Gaussian Splatting SLAM for Dynamic Scenes via Hierarchical Explicit-Implicit Representation Collaboration Rendering

Wenkai Zhu, Xu Li, Qimin Xu, Benwu Wang, Kun Wei, Yiming Peng, Zihang Wang

机构 * School of Instrument Science and Engineering, Southeast University, Nanjing, China(仪器科学与工程学院,东南大学,南京,中国)

专题命中 感知 :LiDAR(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22243 2025-10-28 cs.CV cs.AI 62%

Real-Time Semantic Segmentation on FPGA for Autonomous Vehicles Using LMIINet with the CGRA4ML Framework

Amir Mohammad Khadem Hosseini, Sattar Mirzakuchaki

机构 * Department of Electrical Engineering, Iran University of Science and Technology (IUST)(电气工程系,伊朗科学技术大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16624 2025-10-21 cs.CV cs.RO 62%

Self-Supervised Learning to Fly using Efficient Semantic Segmentation and Metric Depth Estimation for Low-Cost Autonomous UAVs

Sebastian Mocanu, Emil Slusanschi, Marius Leordeanu

机构 * National University of Science and Technology Politehnica Bucharest(波兰技术大学布加勒斯特分校) NORCE Norwegian Research Center(挪威研究机构NORCE)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15673 2025-10-20 cs.CV cs.AI 62%

Valeo Near-Field: a novel dataset for pedestrian intent detection

Antonyo Musabini, Rachid Benmokhtar, Jagdish Bhanushali, Victor Galizzi, Bertrand Luvison, Xavier Perrotton

机构 * Valeo(瓦莱欧) BRAIN Division(BRAIN部门) Universite Paris-Saclay(巴黎-萨克雷大学) CEA(法国国家科学研究中心) List F-91120(法国)

专题命中 感知 :LiDAR(abstract);分类 cs.CV、cs.AI

Journal ref ICCV 2025 - 9th Workshop and Competition on Affective & Behavior Analysis in-the-wild (ABAW)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12260 2025-10-15 cs.CV cs.LG eess.IV 62%

AngularFuse: A Closer Look at Angle-based Perception for Spatial-Sensitive Multi-Modality Image Fusion

Xiaopeng Liu, Yupei Lin, Sen Zhang, Xiao Wang, Yukai Shi, Liang Lin

机构 * School of Information Engineering, Guangdong University of Technology(广东技术大学信息工程学院) TikTok, ByteDance Inc(字节跳动) School of Computer Science, Anhui University(安徽大学计算机科学学院) School of Computer Science and Engineering, Sun Yat-sen University(中山大学计算机科学与工程学院)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、eess.IV

Comments For the first time, angle-based perception was introduced into the multi-modality image fusion task

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11632 2025-10-14 cs.CV cs.AI cs.LG 62%

NV3D: Leveraging Spatial Shape Through Normal Vector-based 3D Object Detection

Krittin Chaowakarn, Paramin Sangwongngam, Nang Htet Htet Aung, Chalie Charoenlarpnopparut

机构 * The School of Information, Computer, and Communication Technology, Sirindhorn International Institute of Technology, Thammasat University(信息、计算机与通信技术学院,Sirindhorn国际技术学院,泰国朱拉隆梭大学) National Electronics and Computer Technology Center, National Science and Technology Development Agency(国家电子与计算机技术中心,国家科学技术发展局) Department of Electrical Engineering, Faculty of Engineering, Chulalongkorn University(电气工程系,工程学院,朱拉隆梭大学)

专题命中 感知 :LiDAR(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11072 2025-10-14 cs.RO cs.AI cs.LG cs.SY eess.SY 62%

PhysHSI: Towards a Real-World Generalizable and Natural Humanoid-Scene Interaction System

Huayi Wang, Wentao Zhang, Runyi Yu, Tao Huang, Junli Ren, Feiyu Jia, Zirui Wang, Xiaojie Niu, Xiao Chen, Jiahe Chen, Qifeng Chen, Jingbo Wang, Jiangmiao Pang

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.AI

Comments Project website: https://why618188.github.io/physhsi/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03701 2025-10-07 cs.CV cs.AI 62%

Referring Expression Comprehension for Small Objects

Kanoko Goto, Takumi Hirose, Mahiro Ukai, Shuhei Kurita, Nakamasa Inoue

机构 * Institute of Science Tokyo(东京科学研究所) National Institute of Informatics(国家信息研究所)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24927 2025-09-30 cs.AI cs.RO cs.SE 62%

When Autonomous Vehicle Meets V2X Cooperative Perception: How Far Are We?

An Guo, Shuoxiao Zhang, Enyi Tang, Xinyu Gao, Haomin Pang, Haoxiang Tian, Yanzhou Mu, Wu Wen, Chunrong Fang, Zhenyu Chen

机构 * Guangzhou University(广州大学) Nanyang Technological University(南洋理工大学) Shenzhen Research Institute of Nanjing University(南京大学深圳研究院)

专题命中 感知 :LiDAR(abstract);分类 cs.RO、cs.AI

Comments The paper has been accepted by the 40th IEEE/ACM International Conference on Automated Software Engineering, ASE 2025

Journal ref Proceedings of the 40th IEEE/ACM International Conference on Automated Software Engineering.2025

详情

展开后加载摘要…

URL PDF HTML 收藏