arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

自动驾驶

自动驾驶感知、规划、BEV、占用预测、激光雷达和仿真评测。

2026-06-24 至 2026-06-24 共收录 4 信号源:cs.RO, cs.CV, eess.IV, cs.AI

1. 感知 4 篇

2606.24759 2026-06-24 cs.CV cs.AI 新提交 81%

UniDrive: A Unified Vision-Language and Grounding Framework for Interpretable Risk Understanding in Autonomous Driving

UniDrive: 面向自动驾驶可解释风险理解的统一视觉-语言与定位框架

Xiaowei Gao, Pengxiang Li, Yitai Cheng, Ruihan Xu, James Haworth, Stephen Law, Yun Ye

机构 * organization= Department of Earth Science \& Engineering, Imperial College London , city= London , postcode= SW7 2AZ , country= United Kingdom organization= SpaceTimeLab, Department of Civil, Environmental Geomatic Engineering, University College London , city= London , postcode= WC1E 6BT , country= United Kingdom organization= Department of Computing, The Hong Kong Polytechnic University , city= Hong Kong , country= China organization= Trinity College, University of Oxford , city= Oxford , postcode= OX1 3BH , country= United Kingdom organization= Department of Geography, University College London , city= London , postcode= WC1E 6BT , country= United Kingdom organization= Centre for Global Infrastructure Resilience, The Bartlett School of Sustainable Construction, University College London , city= London , postcode= WC1E 7HB , country= United Kingdom

专题命中 感知 :autonomous driving(title,abstract);分类 cs.CV、cs.AI

AI总结 提出UniDrive框架,通过融合时序推理与高分辨率感知分支,联合生成风险描述和边界框定位,在DRAMA-Reasoning基准上超越现有方法,提升小目标定位和可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.24096 2026-06-24 cs.CV cs.AI 新提交 62%

Beyond Bayer: Task-Optimal Sensor Co-Design for Robust Autonomous-Driving Segmentation

超越拜耳:面向鲁棒自动驾驶分割的任务最优传感器协同设计

Reeshad Khan, John Gauch

机构 * University Of Arkansas(阿肯色大学)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV、cs.AI

AI总结 提出可微RAW到任务流水线,学习光谱滤色器阵列权重提升分割mIoU,发现光学点扩散函数协同设计为负收益,噪声优化效果微弱,增大CFA块反而有害。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.03694 2026-06-24 cs.RO cs.CV cs.HC 版本更新 62%

Face versus Body Tracking for Human-Robot Interaction: An Egocentric Dataset

面向人机交互的面部与身体跟踪:一个自我中心数据集

Jessica Wenninger, Gabriel Skantze

机构 * Furhat Robotics University of Naples Federico II(那不勒斯费德里科二世大学) Division of Speech, Music and Hearing, KTH Royal Institute of Technology(语音、音乐和听觉研究所,皇家理工学院)

专题命中 感知 :autonomous driving(abstract);分类 cs.RO、cs.CV

AI总结 针对社交机器人自我中心视角下频繁身份切换问题,提出一个自定义标注的自我中心数据集,通过系统评估检测误差、对比面部与身体跟踪,并分析扩展空间记忆和外观重识别的影响,最终优化管道将身份切换减少49%。

Comments 8 pages, 5 figures, 3 tables. Camera-ready version. Accepted to the 35th IEEE International Conference on Robot and Human Interactive Communication (RO-MAN 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.24796 2026-06-24 cs.CV 新提交 57%

Pocket-SLAM: Rendering-Area-Aware Pruning for Memory-Efficient 3DGS-SLAM

Pocket-SLAM:面向内存高效的3DGS-SLAM的渲染区域感知剪枝

Leshu Li, Jie Peng, Yang Zhao

机构 * University of Minnesota, Twin Cities(明尼苏达大学双城分校) University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)

专题命中 感知 :autonomous driving(abstract);分类 cs.CV

AI总结 提出一种渲染区域感知的剪枝策略,根据高斯点对有效渲染区域的贡献进行选择性移除,在EuRoC和KITTI数据集上实现超过60%的内存减少和2倍以上的FPS提升,同时保持定位与建图精度。

Comments 2026 IEEE International Conference on Robotics and Automation(ICRA)

详情

展开后加载摘要…

URL PDF HTML 收藏