arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Robotics and Automation · 会议 · Robotics

2026-07-07 至 2026-07-07 共收录 5
2607.04378 2026-07-07 cs.RO 新提交

SurgAM: Surgical Affordance Map Prediction with Multimodal Feature Fusion for Robot Autonomy

SurgAM:用于机器人自主的多模态特征融合手术能力地图预测

Lei Song, Yonghao Long, Mengya Xu, Jiayi Geng, Xiuyuan Chen, Qi Dou

机构 * Department of Computer Science and Engineering, The Chinese University of Hong Kong(香港中文大学计算机科学与工程系) Department of Thoracic Surgery, Peking University People’s Hospital(北京大学人民医院胸外科) Thoracic Oncology Institute, Peking University People’s Hospital(北京大学人民医院胸部肿瘤研究所) Research Unit of Intelligence Diagnosis and Treatment in Early Non-small Cell Lung Cancer, Chinese Academy of Medical Sciences(中国医学科学院早期非小细胞肺癌智能诊断与治疗研究组) Institute of Advanced Clinical Medicine, Peking University(北京大学先进临床医学院) Beijing Key Laboratory of Innovative Application of Big Data in Lung Cancer, Peking University People’s Hospital(北京大学人民医院肺癌大数据创新应用北京市重点实验室)

AI总结 研究如何通过视觉数据识别手术可操作区域,提出自适应特征融合框架、分层提示学习机制和场景引导注意力解码器,建立新数据集验证,在真实模型上验证框架对下游自动化的适用性。

Comments ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.03553 2026-07-07 cs.CV cs.RO 新提交

iVISION-2DCD: A Long-Term Change Detection Dataset for Large-Scale Outdoor Construction Monitoring

iVISION-2DCD:用于大规模户外建筑监测的长期变化检测数据集

Dayou Mao, Yuchen Lin, Ashkan Ebadi, John Zelek, Alexander Wong, Yuhao Chen

机构 * National Research Council Canada(加拿大国家研究委员会) Vision and Image Processing Research Group, Systems Design Engineering, University of Waterloo(滑铁卢大学系统设计工程系视觉与图像处理研究小组)

AI总结 研究从检测建筑工地变化监测施工进度,针对跨视角变化检测算法缺开源基准数据集问题,用LiDAR点云生成iVISION-2DCD数据集,提出合成数据生成等方法,为相关领域带来新挑战。

Comments 11 pages, 7 figures, 1 table. Accepted for publication at the 2026 IEEE International Conference on Robotics and Automation (ICRA 2026). Project page: https://danielmao2019.github.io/iVISION-2DCD-dataset.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.10174 2026-07-07 cs.RO 版本更新

Autonomous Search for Sparsely Distributed Visual Phenomena through Environmental Context Modeling

通过环境上下文建模自主搜索稀疏分布的视觉现象

Eric Chen, Travis Manderson, Nare Karapetyan, Peter Edmunds, Nicholas Roy, Yogesh Girdhar

机构 * Massachusetts Institute of Technology(麻省理工学院) Woods Hole Oceanographic Institution(伍兹霍尔海洋研究所) California State University(加州州立大学)

AI总结 研究如何让自主水下航行器高效定位稀疏分布的特定珊瑚物种。利用视觉环境上下文作为信号,结合补丁级DINOv2嵌入进行一次性检测,实现自适应规划,提升自主勘测效率。

Comments Accepted to the 2026 IEEE International Conference on Robotics and Automation (ICRA 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.16738 2026-07-07 cs.RO cs.AI 版本更新

MOSAIC: Skill-Centric Manipulation Planning with Physics Simulation

MOSAIC:基于物理模拟的以技能为中心的操作规划

Itamar Mishani, Yorai Shaoul, Maxim Likhachev

机构 * Robotics Institute, School of Computer Science(机器人研究所、计算机科学学院)

AI总结 研究用预定义技能规划长期操作运动的问题,核心方法是利用物理模拟估计技能执行结果,贡献是引入MOSAIC方法,通过两个互补技能家族有效发现基于物理的解决方案。

Comments Accepted for Publication at the 2026 IEEE International Conference on Robotics and Automation (ICRA). Project page: https://skill-mosaic.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16663 2026-07-07 cs.RO cs.CV cs.LG cs.SY eess.SY 版本更新

Mitigating Covariate Shift in Imitation Learning for Autonomous Vehicles Using Latent Space Generative World Models

使用潜在空间生成世界模型减轻自动驾驶模仿学习中的协变量转移

Alexander Popov, Alperen Degirmenci, David Wehr, Shashank Hegde, Ryan Oldja, Alexey Kamenev, Bertrand Douillard, David Nistér, Urs Muller, Ruchi Bhargava, Stan Birchfield, Nikolai Smolyanskiy

机构 * NVIDIA

AI总结 提出用潜在空间生成世界模型解决自动驾驶协变量转移问题,训练时利用世界模型减轻该问题,无需大量训练数据,还引入新感知编码器,实验显示相比之前有显著改进。

Comments 8 pages, 6 figures, original September 2024, accepted at ICRA 2025 Workshop "Robots in the Wild", for associated video file, see https://youtu.be/7m3bXzlVQvU

详情

展开后加载摘要…

URL PDF HTML 收藏