arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

共收录 6059 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 6059 篇

2509.16119 2026-03-17 cs.CV 70%

RadarGaussianDet3D: Gaussian Representation-based Real-time 3D Object Detection with 4D Automotive Radars

RadarGaussianDet3D: 基于高斯表示的实时3D目标检测与4D汽车雷达

Weiyi Xiong, Bing Zhu, Zewei Zheng

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 本文提出RadarGaussianDet3D,通过高斯表示提升4D雷达的3D目标检测性能,采用点高斯编码器和3D高斯点划技术生成密集特征图,结合新的高斯损失函数实现更高效的检测与推理。

Comments Accepted by IEEE Robotics and Automation Letters

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15423 2026-03-13 cs.RO cs.SY eess.SY 70%

Online Slip Detection and Friction Coefficient Estimation for Autonomous Racing

自动驾驶赛车中的在线滑动检测与摩擦系数估计

Christopher Oeltjen, Carson Sobolewski, Saleh Faghfoorian, Lorant Domokos, Giancarlo Vidal, Sriram Yerramsetty, Ivan Ruchkin

机构 * Trustworthy Engineered Autonomy (TEA) Lab, Department of Electrical and Computer Engineering, University of Florida(可信工程自主性实验室,电气与计算机工程系,佛罗里达大学)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.RO

AI总结 本文提出了一种基于IMU和LiDAR数据的轻量级方法,用于自动驾驶赛车中的在线滑动检测和摩擦系数估计,无需复杂模型或训练数据。

Comments Equal contribution by the first three authors

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.11252 2026-03-13 cs.CV 70%

Radiometric fingerprinting of object surfaces using mobile laser scanning and semantic 3D road space models

利用移动激光扫描和语义3D道路空间模型进行物体表面的辐射指纹识别

Benedikt Schwab, Thomas H. Kolbe

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文提出利用移动激光扫描和语义3D道路空间模型,通过生成物体表面的辐射指纹,实现对城市表面材料信息的自动关联与分析。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.07486 2026-03-10 cs.CV 70%

Multi-Modal Decouple and Recouple Network for Robust 3D Object Detection

多模态解耦与耦合网络用于抗干扰的3D目标检测

Rui Ding, Zhaonian Kuang, Yuzhe Ji, Meng Yang, Xinhu Zheng, Gang Hua

机构 * State Key Laboratory of Human-Machine Hybrid Augmented Intelligence, Institute of Artificial Intelligence and Robotics, Xi’an Jiaotong University(人机混合增强智能国家重点实验室,人工智能与机器人研究院,西安交通大学) Intelligent Transportation Thrust of the Systems Hub, The Hong Kong University of Science and Technology (Guangzhou)(系统枢纽智能交通方向,香港科技大学(广州)) Multimodal Experiences Research Lab, Dolby Laboratories(多模态体验研究实验室,Dolby实验室)

专题命中 感知 :BEV(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文提出多模态解耦与耦合网络,通过分离和重新耦合不同模态特征以提高在数据损坏下的3D目标检测鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02481 2026-03-04 cs.CV 70%

ModalPatch: A Plug-and-Play Module for Robust Multi-Modal 3D Object Detection under Modality Drop

ModalPatch: 一种插拔式模块,用于在模态缺失情况下鲁棒的多模态3D目标检测

Shuangzhi Li, Lei Ma, Xingyu Li

机构 * University of Alberta, Canada(阿尔伯塔大学) The University of Tokyo, Japan(东京大学)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 ModalPatch是一种插拔式模块,通过利用传感器数据的时间特性实现感知连续性,提升多模态3D目标检测在模态缺失情况下的鲁棒性和准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.23871 2026-03-02 cs.CV cs.LG 70%

Bandwidth-adaptive Cloud-Assisted 360-Degree 3D Perception for Autonomous Vehicles

带宽自适应的云辅助360度3D感知用于自动驾驶车辆

Faisal Hawladera, Rui Meireles, Gamal Elghazaly, Ana Aguiar, Raphaël Frank

机构 * Interdisciplinary Centre for Security, Reliability, and Trust (SnT), University of Luxembourg, L-1855, Luxembourg(安全、可靠性与信任跨学科研究中心(SnT),卢森堡大学) Computer Science Department, Vassar College, Poughkeepsie, NY 12604, USA(计算机科学系,瓦萨学院)

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 本文提出一种带宽自适应的云辅助3D感知方法,通过动态分割计算任务和特征压缩,降低延迟并提升检测精度,适用于复杂城市环境中的自动驾驶。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20632 2026-02-25 cs.CV 70%

Boosting Instance Awareness via Cross-View Correlation with 4D Radar and Camera for 3D Object Detection

通过4D雷达和相机的跨视图相关性提升实例意识

Xiaokai Bai, Lianqing Zheng, Si-Yuan Cao, Xiaohan Zhang, Zhe Wu, Beinan Yu, Fang Wang, Jie Bai, Hui-Liang Shen

机构 * College of Information Science and Electronic Engineering, Zhejiang University(浙江大学信息科学与电子工程学院) School of Automotive Studies, Tongji University(同济大学汽车学院) College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) Jinhua Institute of Zhejiang University(浙江大学金华研究院) School of Information and Electrical Engineering, Hangzhou City University(杭州城市学院信息与电气工程学院) Hangzhou City University Binjiang Innovation Center(杭州城市学院滨江创新中心)

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 SIFormer通过结合4D雷达和相机的跨视图相关性,提升3D目标检测中的实例意识,结合两种融合范式的优点,提高检测精度。

Comments 14 pages, 10 figures, 13 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06685 2026-02-24 cs.CV 70%

MOGS: Monocular Object-guided Gaussian Splatting in Large Scenes

MOGS:在大场景中使用单目物体引导的高斯散射

Shengkai Zhang, Yuhe Liu, Jianhua He, Xuedou Xiao, Mozi Chen, Kezhong Liu

机构 * State Key Laboratory of Maritime Technology and Safety, Wuhan University of Technology(船舶技术与安全国家重点实验室,武汉理工大学) University of Essex(埃塞克斯大学)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 MOGS通过单目视觉-惯性传感器实现大场景中高斯散射的物体引导密集深度生成,显著降低训练时间和内存消耗,同时保持高质量渲染效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22890 2026-02-16 cs.CV cs.CR 70%

CP-uniGuard: A Unified, Probability-Agnostic, and Adaptive Framework for Malicious Agent Detection and Defense in Multi-Agent Embodied Perception Systems

CP-uniGuard: 多智能体具身体验系统中恶意代理检测与防御的统一、概率无关和自适应框架

Senkang Hu, Yihang Tao, Guowen Xu, Xinyuan Qian, Yiqin Deng, Xianhao Chen, Sam Tak Wu Kwong, Yuguang Fang

机构 * Hong Kong JC STEM Lab of Smart City and Department of Computer Science, City University of Hong Kong(香港JC STEM实验室及城市大学计算机科学系) School of Computer Science and Engineering, University of Electronic Science and Technology of China(电子科技大学计算机科学与工程学院) Department of Electrical and Electronic Engineering, The University of Hong Kong(香港大学电子与电气工程系) School of Data Science, Lingnan University(岭南大学数据科学学院)

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 CP-uniGuard通过概率无关的样本共识和自适应阈值,实现多智能体系统中恶意代理的检测与防御。

Comments Accepted by IEEE Transactions on Mobile Computing (TMC)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06406 2026-02-09 cs.CV 70%

Point Virtual Transformer

点虚拟变换器

Veerain Sood, Bnalin, Gaurav Pandey

机构 * Texas A \& M University Email Engineering Technology \& Industrial Distribution Texas A \& M University Texas, USA

专题命中 感知 :BEV(abstract);LiDAR(abstract);分类 cs.CV

AI总结 PointViT通过融合真实和虚拟点,提升远距离3D目标检测性能,实现91.16%的3D AP和95.94%的BEV AP。

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05538 2026-02-09 cs.CV 70%

A Comparative Study of 3D Person Detection: Sensor Modalities and Robustness in Diverse Indoor and Outdoor Environments

三维人物检测的比较研究:传感器模态与在多样室内和室外环境中的鲁棒性

Malaz Tamim, Andrea Matic-Flierl, Karsten Roscher

机构 * Fraunhofer Institute for Cognitive Systems IKS(弗劳恩霍夫认知系统研究所)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文比较了三种三维人物检测方法,发现融合模型在复杂场景中表现最佳,但对传感器错位仍敏感。

Comments Accepted for VISAPP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.14325 2026-02-05 cs.CV 70%

Unlocking Past Information: Temporal Embeddings in Cooperative Bird's Eye View Prediction

解锁过去信息:合作鸟瞰图预测中的时间嵌入

Dominik Rößle, Jeremias Gerner, Klaus Bogenberger, Daniel Cremers, Stefanie Schmidtner, Torsten Schön

机构 * Department of Computer Science and AImotion Bavaria, Technische Hochschule Ingolstadt(计算机科学系和AImotion巴伐利亚,因戈尔施塔特技术大学) Department of Electrical Engineering and AImotion Bavaria, Technische Hochschule Ingolstadt(电气工程系和AImotion巴伐利亚,因戈尔施塔特技术大学) School of Engineering and Design, Technical University of Munich(工程与设计学院,慕尼黑技术大学) School of Computation, Information and Technology, Technical University of Munich(计算、信息与技术学院,慕尼黑技术大学)

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 TempCoBEV通过整合历史信息提升合作鸟瞰图预测的准确性和可靠性,尤其在通信故障情况下表现显著。

Comments Copyright 2024 IEEE. This is the accepted version of the paper. In 2024 IEEE Intelligent Vehicles Symposium (IV), pp. 2220-2225. Official paper available at https://doi.org/10.1109/IV55156.2024.10588608

Journal ref IEEE Intelligent Vehicles Symposium (IV), pp. 2220-2225, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02858 2026-02-04 cs.RO cs.LG cs.MA cs.NI cs.SY eess.SY 70%

IMAGINE: Intelligent Multi-Agent Godot-based Indoor Networked Exploration

IMAGINE: 基于 Godot 的智能多智能体室内网络探索

Tiago Leite, Maria Conceição, António Grilo

机构 * Institute for Systems and Robotics(系统研究所)

专题命中 感知 :LiDAR(abstract);occupancy(abstract);分类 cs.RO

AI总结 本文提出基于Godot的多智能体强化学习方法,用于解决室内未知环境中的自主协作探索问题,通过高保真模拟和课程学习提升训练效率与鲁棒性。

Comments 12 pages, submitted to a journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03906 2026-01-27 cs.CV 70%

From Filters to VLMs: Benchmarking Defogging Methods through Object Detection and Segmentation Performance

从滤波器到视觉语言模型:通过目标检测和分割性能评估去雾方法

Ardalan Aryashad, Parsa Razmara, Amin Mahjoub, Seyedarmin Azizi, Mahdi Salmani, Arad Firouzkouhi

机构 * University of Southern California(南加州大学)

专题命中 感知 :autonomous driving(abstract);driving perception(abstract);分类 cs.CV

AI总结 本文通过目标检测和分割性能评估,探讨了去雾方法在真实与合成环境中的有效性,揭示了视觉语言模型在恶劣天气下的应用潜力。

Comments Accepted at WACV 2026 Proceedings (Oral), 5th Workshop on Image, Video, and Audio Quality Assessment in Computer Vision, with a focus on VLM and Diffusion Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15545 2026-01-23 cs.CV 70%

Multi-View Projection for Unsupervised Domain Adaptation in 3D Semantic Segmentation

多视角投影用于无监督领域适应的3D语义分割

Andrew Caunes, Thierry Chateau, Vincent Fremont

机构 * Logiroad, Nantes, France(法国南特Logiroad) LS2N - Ecole Centrale de Nantes, France(法国南特LS2N-中央理工学院)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文提出了一种基于多视角投影的无监督领域适应方法,通过合成数据集训练2D分割模型并回投影生成3D标签,实现3D语义分割的领域适应和稀有类别分割。

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10962 2026-01-19 cs.CV 70%

V2X-Radar: A Multi-modal Dataset with 4D Radar for Cooperative Perception

V2X-Radar: 一种包含4D雷达的多模态数据集用于协作感知

Lei Yang, Xinyu Zhang, Jun Li, Chen Wang, Jiaqi Ma, Zhiying Song, Tong Zhao, Ziying Song, Li Wang, Mo Zhou, Yang Shen, Kai Wu, Chen Lv

机构 * School of Vehicle and Mobility, Tsinghua University(车辆与移动性学院,清华大学) Nanyang Technological University(南洋理工大学) CUMTB(中国交通车辆技术研究所) University of California, Los Angeles(加州大学洛杉矶分校) Beijing Jiaotong University(北京交通大学) ByteDance(字节跳动)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 V2X-Radar是首个包含4D雷达的多模态数据集,用于提升自动驾驶中的协作感知能力。

Comments NeurIPS 2025 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.04968 2026-01-09 cs.CV 70%

SparseLaneSTP: Leveraging Spatio-Temporal Priors with Sparse Transformers for 3D Lane Detection

SparseLaneSTP: 利用稀疏变换器结合时空先验进行3D车道检测

Maximilian Pittner, Joel Janai, Mario Faigle, Alexandru Paul Condurache

机构 * Bosch Mobility Solutions, Robert Bosch GmbH(博世移动解决方案,罗伯特·博世有限公司) Institute of Neuro- and Bioinformatics, University of Lübeck(神经与生物医学研究所,吕贝克大学) Institute for Signal Processing and System Theory, University of Stuttgart(信号处理与系统理论研究所,斯图加特大学)

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 SparseLaneSTP通过整合车道结构的几何属性和时间信息,利用稀疏变换器和时空注意力机制,提升3D车道检测的精度和一致性。

Comments Published at IEEE/CVF International Conference on Computer Vision (ICCV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03001 2026-01-07 cs.CV 70%

Towards Efficient 3D Object Detection for Vehicle-Infrastructure Collaboration via Risk-Intent Selection

为车辆-基础设施协作实现高效3D物体检测的险意选择

Li Wang, Boqi Li, Hang Chen, Xingjian Wu, Yichen Wang, Jiewen Tan, Xinyu Zhang, Huaping Liu

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

AI总结 本文提出RiSe框架,通过风险意图选择实现车辆-基础设施协作中高效3D物体检测,减少通信量同时保持高检测精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.24922 2026-01-01 cs.CV 70%

Semi-Supervised Diversity-Aware Domain Adaptation for 3D Object detection

半监督多样性感知领域适应用于3D目标检测

Bartłomiej Olber, Jakub Winter, Paweł Wawrzyński, Andrii Gamalii, Daniel Górniak, Marcin Łojek, Robert Nowak, Krystian Radlak

机构 * Warsaw University of Technology(华沙技术大学) IDEAS NCBR

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文提出了一种基于神经激活模式的半监督领域适应方法,通过少量多样化样本标注提升3D目标检测在跨领域应用中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.23215 2025-12-30 cs.CV 70%

AVOID: The Adverse Visual Conditions Dataset with Obstacles for Driving Scene Understanding

AVOID: 用于驾驶场景理解的障碍物条件数据集

Jongoh Jeong, Taek-Jin Song, Jong-Hwan Kim, Kuk-Jin Yoon

专题命中 感知 :self-driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 AVOID数据集通过模拟环境收集了多种恶劣天气条件下的道路障碍物,用于提升自动驾驶场景理解的实时障碍物检测能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22503 2025-12-30 cs.CV 70%

SCAFusion: A Multimodal 3D Detection Framework for Small Object Detection in Lunar Surface Exploration

SCAFusion: 一种针对月球表面探索的小目标多模态3D检测框架

Xin Chen, Kang Luo, Yangyi Xiao, Hesheng Wang

机构 * Department of Automation, Key Laboratory of System Control and Information Processing of Ministry of Education, Key Laboratory of Marine Intelligent Equipment and System of Ministry of Education, Shanghai Engineering Research Center of Intelligent Control and Management, Shanghai Jiao Tong University(自动化系、教育部系统控制与信息处理重点实验室、教育部海洋智能装备与系统重点实验室、上海智能控制与管理工程研究中心、上海交通大学)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 SCAFusion提出了一种针对月球表面探索的小目标多模态3D检测框架,通过改进的特征对齐和坐标注意机制提升小目标检测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19192 2025-12-29 eess.IV 70%

An on-chip Pixel Processing Approach with 2.4μs latency for Asynchronous Read-out of SPAD-based dToF Flash LiDARs

基于SPAD的dToF飞利拍雷达异步读取的芯片像素处理方法

Yiyang Liu, Rongxuan Zhang, Istvan Gyongy, Alistair Gorman, Sarrah M. Patanwala, Filip Taneski, Robert K. Henderson

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 eess.IV

AI总结 本文提出了一种低延迟的异步像素处理方法,用于基于SPAD的dToF飞利拍雷达,通过事件驱动方式提升深度采集效率,同时优化计算负载与动态响应能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20538 2025-12-24 cs.RO 70%

Uni-Mapper: Unified Mapping Framework for Multi-modal LiDARs in Complex and Dynamic Environments

Uni-Mapper:多模态激光雷达在复杂动态环境中的统一映射框架

Gilhwan Kang, Hogyun Kim, Byunghee Choi, Seokhwan Jeong, Young-Sik Shin, Younggun Cho

机构 * Hyundai Motor Company(现代汽车公司) Inha University(inha大学) Korea Institute of Machinery and Materials(韩国机械材料研究院)

专题命中 感知 :LiDAR(abstract);occupancy(abstract);分类 cs.RO

AI总结 Uni-Mapper通过动态感知和多模态激光雷达融合技术,实现复杂动态环境下的统一地图构建与回环检测。

Comments 18 pages, 14 figures

Journal ref 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.17620 2025-12-22 cs.CV 70%

StereoMV2D: A Sparse Temporal Stereo-Enhanced Framework for Robust Multi-View 3D Object Detection

StereoMV2D: 一种稀疏时间立体增强框架,用于鲁棒的多视角3D目标检测

Di Wu, Feng Yang, Wenhui Zhao, Jinwen Yu, Pan Liao, Benlian Xu, Dingwen Zhang

机构 * the school of automation, Northwestern Polytechnical University(自动化学院,西北工业大学) the school of electronic and information engineering, Suzhou University of Science and Technology(电子信息工程学院,苏州科技大学)

专题命中 感知 :autonomous driving(abstract);driving perception(abstract);分类 cs.CV

AI总结 StereoMV2D通过整合时间立体建模,提升多视角3D目标检测的鲁棒性和精度,无需显著增加计算开销。

Comments 12 pages, 4 figures. This work has been submitted to the IEEE for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11926 2025-12-16 cs.CV 70%

TransBridge: Boost 3D Object Detection by Scene-Level Completion with Transformer Decoder

TransBridge: 通过变换解码器进行场景级补全提升3D目标检测

Qinghao Meng, Chenming Wu, Liangjun Zhang, Jianbing Shen

机构 * School of Computer Science, Beijing Institute of Technology(北京理工大学计算机科学学院) Robotics and Autonomous Driving Lab (RAL), Baidu Research(百度研究机器人与自动驾驶实验室) State Key Laboratory of Internet of Things for Smart City, Department of Computer and Information Science, University of Macau(澳门大学智能城市物联网国家重点实验室,计算机与信息科学系)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

AI总结 TransBridge通过变换解码器进行场景级补全,提升3D目标检测性能,实现稀疏区域特征增强和密集点云生成

Comments 12 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17619 2025-11-25 cs.CV 70%

Rethinking the Encoding and Annotating of 3D Bounding Box: Corner-Aware 3D Object Detection from Point Clouds

重新思考3D边界框的编码与标注:从点云中基于角点的3D目标检测

Qinghao Meng, Junbo Yin, Jianbing Shen, Yunde Jia

机构 * School of Computer Science, Beijing Institute of Technology(计算机科学学院,北京理工大学) Computer Science Program, Computer, Electrical and Mathematical Sciences and Engineering (CEMSE) Division, Center of Excellence for Smart Health, and Center of Excellence for Generative AI, King Abdullah University of Science and Technology (KAUST)(计算机科学项目,计算机、电气和数学科学与工程(CEMSE)部门,智能健康卓越中心,生成人工智能卓越中心,国王阿卜杜勒·阿齐兹大学科学与技术(KAUST)) State Key Laboratory of Internet of Things for Smart City, Department of Computer and Information Science, University of Macau(智慧城市物联网国家重点实验室,澳门大学计算机与信息科学系) Guangdong Provincial Key Laboratory of Machine Perception and Intelligent Computing, Shenzhen MSU-BIT University(广东省机器感知与智能计算重点实验室,深圳MSU-BIT大学) Beijing Key Laboratory of Intelligent Information Technology, School of Computer Science, Beijing Institute of Technology(北京智能信息技术重点实验室,计算机科学学院,北京理工大学)

专题命中 感知 :BEV(abstract);LiDAR(abstract);分类 cs.CV

AI总结 本文提出基于角点的3D目标检测方法,通过角点对齐回归提升检测性能,仅需BEV角点点击即可达到接近全监督的准确率。

Comments 8 pages, 5 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12093 2025-11-18 cs.CV 70%

SFMNet: Sparse Focal Modulation for 3D Object Detection

Oren Shrout, Ayellet Tal

机构 * Technion, Israel(技术离子大学)

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

Comments WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09347 2025-11-17 cs.CV 70%

FQ-PETR: Fully Quantized Position Embedding Transformation for Multi-View 3D Object Detection

Jiangyong Yu, Changyong Shu, Sifan Zhou, Zichen Yu, Xing Hu, Yan Chen, Dawei Yang

专题命中 感知 :autonomous driving(abstract);LiDAR(abstract);分类 cs.CV

Comments I made an operational error. I intended to update the paper with Identifier arXiv:2502.15488, not submit a new paper with a different identifier. Therefore, I would like to withdraw the current submission and resubmit an updated version for Identifier arXiv:2502.15488

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.10035 2025-11-14 cs.CV 70%

DGFusion: Dual-guided Fusion for Robust Multi-Modal 3D Object Detection

Feiyang Jia, Caiyan Jia, Ailin Liu, Shaoqing Xu, Qiming Xia, Lin Liu, Lei Yang, Yan Gong, Ziying Song

机构 * School of Computer Science and Technology, Beijing Key Laboratory of Traffic Data Mining and Embodied Intelligence, Beijing Jiaotong University(计算机科学与技术学院、交通数据挖掘与具身智能北京市重点实验室、北京交通大学) State Key Laboratory of Internet of Things for Smart City and Department of Electrome chanical Engineering, University of Macau(智能城市物联网国家重点实验室、澳门大学机电工程系) Fujian Key Laboratory of Sensing and Computing for Smart Cities, Xiamen University(智能城市感知与计算福建省重点实验室、厦门大学) School of Mechanical and Aerospace Engineering, Nanyang Technological University(机械与航空航天工程学院、南洋理工大学) State Key Laboratory of Robotics and System, Harbin Institute of Technology(机器人系统国家重点实验室、哈尔滨工业大学)

专题命中 感知 :autonomous driving(abstract);driving perception(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23325 2025-11-13 cs.CV 70%

FASTopoWM: Fast-Slow Lane Segment Topology Reasoning with Latent World Models

Yiming Yang, Hongbin Lin, Yueru Luo, Suzhong Fu, Chao Zheng, Xinrui Yan, Shuqi Mei, Kun Tang, Shuguang Cui, Zhen Li

机构 * FNii, CUHK-Shenzhen(CUHK-Shenzhen研究院) SSE, CUHK-Shenzhen(CUHK-Shenzhen学院) T Lab, Tencent(腾讯实验室)

专题命中 感知 :autonomous driving(abstract);BEV(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏