arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

共收录 6065 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 6065 篇

2507.10761 2025-07-16 cs.AI cs.HC 57%

Detecting AI Assistance in Abstract Complex Tasks

Tyler King, Nikolos Gurney, John H. Miller, Volkan Ustun

机构 * Cornell University(康奈尔大学) University of Southern California(南加州大学) Carnegie Mellon University(卡内基梅隆大学) Santa Fe Institute(圣菲研究所)

专题命中 感知 :autonomous driving(abstract);分类 cs.AI

Comments Accepted to HCII 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.06109 2025-07-16 cs.CV 57%

PerLDiff: Controllable Street View Synthesis Using Perspective-Layout Diffusion Models

Jinhua Zhang, Hualian Sheng, Sijia Cai, Bing Deng, Qiao Liang, Wen Li, Ying Fu, Jieping Ye, Shuhang Gu

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.10600 2025-07-16 cs.CV 57%

SparseRadNet: Sparse Perception Neural Network on Subsampled Radar Data

Jialong Wu, Mirko Meuter, Markus Schoeler, Matthias Rottmann

机构 * University of Wuppertal(乌珀塔尔大学) Aptiv Services Deutschland GmbH(Aptiv Services德国公司)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments 18 pages, 4 figures, 5 tables, with supplement

Journal ref European Conference on Computer Vision, 2024: 52-69

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.08824 2025-07-15 cs.CV 57%

Pathfinder for Low-altitude Aircraft with Binary Neural Network

Kaijie Yin, Tian Gao, Hui Kong

机构 * University of Macau(澳门大学)

专题命中 感知 :LiDAR(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.08644 2025-07-14 cs.CV 57%

OnlineBEV: Recurrent Temporal Fusion in Bird's Eye View Representations for Multi-Camera 3D Perception

Junho Koh, Youngwoo Lee, Jungho Kim, Dongyoung Lee, Jun Won Choi

机构 * Autonomous Driving Development Center, Hyundai Motors(韩世汽车自动驾驶开发中心) Department of Electrical Engineering, Hanyang University(翰阳大学电气工程系) Interdisciplinary Program in Artificial Intelligence, Seoul National University(首尔国立大学人工智能跨学科项目) Department of Electrical and Computer Engineering, Seoul National University(首尔国立大学电气与计算机工程系)

专题命中 感知 :BEV(abstract);分类 cs.CV

Comments Accepted to Transactions on Intelligent Transportation Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.07846 2025-07-11 cs.RO 57%

ROS Help Desk: GenAI Powered, User-Centric Framework for ROS Error Diagnosis and Debugging

Kavindie Katuwandeniya, Samith Rajapaksha Jayasekara Widhanapathirana

机构 * CSIRO Robotics(澳大利亚CSIRO机器人研究所)

专题命中 感知 :LiDAR(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07446 2025-07-10 cs.RO 57%

TPT-Bench: A Large-Scale, Long-Term and Robot-Egocentric Dataset for Benchmarking Target Person Tracking

Hanjing Ye, Yu Zhan, Weixi Situ, Guangcheng Chen, Jingwen Yu, Ziqi Zhao, Kuanqi Cai, Arash Ajoudani, Hong Zhang

机构 * Robotics and Computer Vision Laboratory, Southern University of Science and Technology, Shenzhen, China(机器人与计算机视觉实验室,南方科技大学,深圳,中国) Human-Robot Interfaces and Interaction Laboratory, Italian Institute of Technology, Geona, Italy(人机交互实验室,意大利理工学院,Geona,意大利) Robotics Perception and Intelligence Laboratory, Southern University of Science and Technology, Shenzhen, China(机器人感知与智能实验室,南方科技大学,深圳,中国) Chen Kar-Shun Robotics Institute, Hong Kong University of Science and Technology, Hong Kong SAR, China(陈卡尔顺机器人研究所,香港科技大学,香港特别行政区,中国)

专题命中 感知 :LiDAR(abstract);分类 cs.RO

Comments Under review. web: https://medlartea.github.io/tpt-bench/

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.08277 2025-07-10 cs.CV 57%

StixelNExT: Toward Monocular Low-Weight Perception for Object Segmentation and Free Space Detection

Marcel Vosshans, Omar Ait-Aider, Youcef Mezouar, Markus Enzweiler

机构 * Institute for Intelligent Systems(智能系统研究所) Faculty of Computer Science and Engineering, University of Applied Sciences Esslingen(计算机科学与工程学院,应用科学大学埃斯林根) Institut Pascal ISPR(帕斯卡尔研究所) Universite Clermont Auvergne INP / CNRS(克勒蒙奥弗涅大学 INP / CNRS)

专题命中 感知 :LiDAR(abstract);分类 cs.CV

Comments Accepted Conference Paper, IEEE IV 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04822 2025-07-08 cs.CV 57%

SeqGrowGraph: Learning Lane Topology as a Chain of Graph Expansions

Mengwei Xie, Shuang Zeng, Xinyuan Chang, Xinran Liu, Zheng Pan, Mu Xu, Xing Wei

机构 * Amap, Alibaba Group(阿里巴巴集团阿地图) Xi’an Jiaotong University(西安交通大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04686 2025-07-08 cs.RO 57%

MOSU: Autonomous Long-range Robot Navigation with Multi-modal Scene Understanding

Jing Liang, Kasun Weerakoon, Daeun Song, Senthurbavan Kirubaharan, Xuesu Xiao, Dinesh Manocha

机构 * University of Maryland, College Park MD, 20740, USA(马里兰大学) Goerge Mason University, Fairfax, VA, 22030, USA(乔治·马歇尔大学)

专题命中 感知 :LiDAR(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04529 2025-07-08 cs.CV 57%

A Data-Driven Novelty Score for Diverse In-Vehicle Data Recording

Philipp Reis, Joshua Ransiek, David Petri, Jacob Langner, Eric Sax

机构 * FZI Research Center for Information Technology(FZI信息科技研究中心) Karlsruhe Institute of Technology(卡尔斯鲁厄理工学院)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments 8 pages, accepted at the IEEE ITSC 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04369 2025-07-08 cs.CV 57%

MambaFusion: Height-Fidelity Dense Global Fusion for Multi-modal 3D Object Detection

Hanshi Wang, Jin Gao, Weiming Hu, Zhipeng Zhang

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), CASIA(多模态人工智能系统国家重点实验室(MAIS),中国科学院自动化所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院) Anyverse Intelligence Beijing Key Laboratory of Super Intelligent Security of Multi-Modal Information(北京多模态信息超级智能安全重点实验室) School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院)

专题命中 感知 :LiDAR(abstract);分类 cs.CV

Comments 10 pages

Journal ref ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03816 2025-07-08 cs.CV 57%

Zero Memory Overhead Approach for Protecting Vision Transformer Parameters

Fereshteh Baradaran, Mohsen Raji, Azadeh Baradaran, Arezoo Baradaran, Reihaneh Akbarifard

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05336 2025-07-08 cs.CV 57%

VideoMolmo: Spatio-Temporal Grounding Meets Pointing

Ghazi Shazan Ahmad, Ahmed Heakl, Hanan Gani, Abdelrahman Shaker, Zhiqiang Shen, Fahad Shahbaz Khan, Salman Khan

机构 * Mohamed Bin Zayed University of Artificial Intelligence(莫扎德·本·扎耶德人工智能大学) Linköping University(林奈大学) Australian National University(澳大利亚国立大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments 20 pages, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.05945 2025-07-04 cs.CV 57%

MV2DFusion: Leveraging Modality-Specific Object Semantics for Multi-Modal 3D Detection

Zitian Wang, Zehao Huang, Yulu Gao, Naiyan Wang, Si Liu

机构 * Institute of Artificial Intelligence, Beihang University(北京航空航天大学人工智能学院)

专题命中 感知 :LiDAR(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.00182 2025-07-03 cs.CV 57%

Graph-Based Deep Learning for Component Segmentation of Maize Plants

J. I. Ruiz-Martinez, A. Mendez-Vazquez, E. Rodriguez-Tello

机构 * Cinvestav(辛克维斯塔研究所)

专题命中 感知 :LiDAR(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17576 2025-07-03 cs.RO 57%

CU-Multi: A Dataset for Multi-Robot Data Association

Doncey Albin, Miles Mena, Annika Thomas, Harel Biggie, Xuefei Sun, Dusty Woods, Steve McGuire, Christoffer Heckman

机构 * Autonomous Robotics and Perception Group in the Computer Science Department at the University of Colorado Boulder(科罗拉多大学博尔德分校计算机科学系自主机器人与感知小组) Aerospace Controls Laboratory at Massachusetts Institute of Technology(麻省理工学院航空航天控制实验室) Computer Science and Artificial Intelligence Laboratory at Massachusetts Institute of Technology(麻省理工学院计算机科学与人工智能实验室) Human-Aware Robotic Exploration Lab at University of California Santa Cruz(加州大学圣克鲁兹分校人感知机器人探索实验室)

专题命中 感知 :LiDAR(abstract);分类 cs.RO

Comments 8 pages, 6 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.00570 2025-07-02 cs.CV 57%

Out-of-distribution detection in 3D applications: a review

Zizhao Li, Xueyang Kang, Joseph West, Kourosh Khoshelham

机构 * The University of Melbourne(墨尔本大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23437 2025-07-01 cs.SD cs.AI eess.AS 57%

From Large-scale Audio Tagging to Real-Time Explainable Emergency Vehicle Sirens Detection

Stefano Giacomelli, Marco Giordano, Claudia Rinaldi, Fabio Graziosi

机构 * Department of Information Engineering, Computer Science and Mathematics (DISIM), University of L’Aquila(信息工程、计算机科学与数学系(DISIM),拉奎拉大学) National Inter-University Consortium for Telecommunications (CNIT), University of L’Aquila(电信国家大学间联合体(CNIT),拉奎拉大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.AI

Comments pre-print (submitted to the IEEE/ACM Transactions on Audio, Speech, and Language Processing)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.08482 2025-07-01 cs.CV cs.LG 57%

Methodology for an Analysis of Influencing Factors on 3D Object Detection Performance

Anton Kuznietsov, Dirk Schweickard, Steven Peters

机构 * Institute of Automotive Engineering(汽车工程研究所) Technical University of Darmstadt(达姆斯塔特技术大学)

专题命中 感知 :LiDAR(abstract);分类 cs.CV

Comments IEEE International Conference on Autonomous and Trusted Computing (IEEE ATC), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22364 2025-06-30 cs.RO 57%

Robotic Multimodal Data Acquisition for In-Field Deep Learning Estimation of Cover Crop Biomass

Joe Johnson, Phanender Chalasani, Arnav Shah, Ram L. Ray, Muthukumar Bagavathiannan

机构 * Department of Soil and Crop Sciences, Texas A&M University(土壤与作物科学系,德克萨斯A&M大学) Department of Computer Science and Engineering, Texas A&M University(计算机科学与工程系,德克萨斯A&M大学) Department of Mechanical Engineering, Texas A&M University(机械工程系,德克萨斯A&M大学) College of Agriculture, Food and Natural Resources, Prairie View A&M University(农业、食品与自然资源学院,普里奥里视A&M大学)

专题命中 感知 :LiDAR(abstract);分类 cs.RO

Comments Accepted in the Extended Abstract, The 22nd International Conference on Ubiquitous Robots (UR 2025), Texas, USA

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17944 2025-06-30 cs.CV 57%

SegChange-R1: LLM-Augmented Remote Sensing Change Detection

Fei Zhou

机构 * Neusoft Institute Guangdong, China(广东新soft研究院) Airace Technology Co.,Ltd., China(Airace技术有限公司)

专题命中 感知 :BEV(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.20388 2025-06-26 cs.CV 57%

A Novel Large Vision Foundation Model (LVFM)-based Approach for Generating High-Resolution Canopy Height Maps in Plantations for Precision Forestry Management

Shen Tan, Xin Zhang, Liangxiu Han, Huaguo Huang, Han Wang

专题命中 感知 :LiDAR(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.13980 2025-06-25 cs.CV 57%

FusionSAM: Visual Multi-Modal Learning with Segment Anything

Daixun Li, Weiying Xie, Mingxiang Cao, Yunke Wang, Yusi Zhang, Leyuan Fang, Yunsong Li, Chang Xu

机构 * Xidian University(西安电子科技大学) University of Sydney(悉尼大学) Hunan University(湖南大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16262 2025-06-24 cs.CV 57%

R3eVision: A Survey on Robust Rendering, Restoration, and Enhancement for 3D Low-Level Vision

Weeyoung Kwon, Jeahun Sung, Minkyu Jeon, Chanho Eom, Jihyong Oh

机构 * Department of Metaverse Convergence, GSAIM, Chung-Ang University(元宇宙融合系、GSAIM、 Chung-Ang 大学) Department of Computer Science, Princeton University(计算机科学系、普林斯顿大学) Department of Imaging Science, GSAIM, Chung-Ang University(成像科学系、GSAIM、 Chung-Ang 大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments Please visit our project page at https://github.com/CMLab-Korea/Awesome-3D-Low-Level-Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.17457 2025-06-24 cs.CV 57%

When Every Millisecond Counts: Real-Time Anomaly Detection via the Multimodal Asynchronous Hybrid Network

Dong Xiao, Guangyao Chen, Peixi Peng, Yangru Huang, Yifan Zhao, Yongxing Dai, Yonghong Tian

机构 * National Key Laboratory for Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室,计算机科学学院,北京大学) Department of Software and Microelectronics, Peking University(软件与微电子系,北京大学) School of Electronic and Computer Engineering, Peking University(电子与计算机工程学院,北京大学) State Key Laboratory of Virtual Reality Technology and Systems, SCSE, Beihang University(虚拟现实技术与系统国家重点实验室,北航软件学院) Peng Cheng Laboratory(鹏城实验室) Baidu Inc(百度公司)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments ICML 2025 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16936 2025-06-23 cs.RO 57%

SDDiff: Boost Radar Perception via Spatial-Doppler Diffusion

Shengpeng Wang, Xin Luo, Yulong Xie, Wei Wang

机构 * Huazhong University of Science and Technology(华中科技大学) Wuhan University(武汉大学)

专题命中 感知 :occupancy(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.16219 2025-06-23 cs.RO 57%

Probabilistic Collision Risk Estimation for Pedestrian Navigation

Amine Tourki, Paul Prevel, Nils Einecke, Tim Puphal, Alexandre Alahi

机构 * Biped Robotics SA Honda Research Institute Europe GmbH(本田欧洲研究机构) Honda Research Institute Japan Co., Ltd.(本田日本研究机构) École Polytechnique Fédérale de Lausanne(洛桑联邦理工学院)

专题命中 感知 :autonomous driving(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.09682 2025-06-23 cs.CV 57%

UDA4Inst: Unsupervised Domain Adaptation for Instance Segmentation

Yachan Guo, Yi Xiao, Danna Xue, Jose L. Gomez, Antonio M. Lopez

机构 * Computer Vision Center, Universitat Autònoma de Barcelona(计算机视觉中心,巴塞罗那自治大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

Comments Accepted at IEEE Intelligent Vehicles Symposium (IV 2025) as an oral presentation

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.10815 2025-06-19 cs.CV 57%

PanopticNeRF-360: Panoramic 3D-to-2D Label Transfer in Urban Scenes

Xiao Fu, Shangzhan Zhang, Tianrun Chen, Yichong Lu, Xiaowei Zhou, Andreas Geiger, Yiyi Liao

机构 * Zhejiang University, China(浙江大学) Autonomous Vision Group (AVG) at the University of Tübingen and Tübingen AI Center, Germany(图宾根大学自主视觉组(AVG)及图宾根人工智能中心)

专题命中 感知 :self-driving(abstract);分类 cs.CV

Comments TPAMI 2025. Project page: http://fuxiao0719.github.io/projects/panopticnerf360/ Code: https://github.com/fuxiao0719/PanopticNeRF/tree/panopticnerf360

详情

展开后加载摘要…

URL PDF HTML 收藏