arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Intelligent Robots and Systems · 会议 · Robotics

共收录 116
2603.09760 2026-07-17 cs.CV cs.RO eess.IV 版本更新

PanoAffordanceNet: Towards Holistic Affordance Grounding in 360° Indoor Environments

PanoAffordanceNet:迈向360°室内环境中的整体 affordance 地标

Guoliang Zhu, Wanjun Jia, Caoyang Shao, Yuheng Zhang, Zhiyong Li, Kailun Yang

机构 * School of Artificial Intelligence and Robotics, Hunan University, China(湖南大学人工智能与机器人学院)

AI总结 PanoAffordanceNet通过DASM和OSDH解决360°室内环境中的整体affordance地标问题,整合多级约束抑制语义漂移,构建360-AGD数据集,提升具身智能的场景感知能力。

Comments Accepted to IEEE/RSJ IROS 2026. The source code and benchmark dataset will be made publicly available at https://github.com/GL-ZHU925/PanoAffordanceNet

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05754 2026-07-17 cs.RO 版本更新

Safe-Night VLA: Seeing the Unseen via Thermal-Perceptive Vision-Language-Action Models for Safety-Critical Manipulation

Safe-Night VLA: 通过热感知视觉-语言-动作模型看见未见事物以实现安全关键操作

Dian Yu, Qingchuan Zhou, Bingkun Huang, Majid Khadiv, Zewen Yang

机构 * Munich Institute of Robotics and Machine Intelligence (MIRMI), Technical University of Munich (TUM)(慕尼黑机器人与机器智能研究所(MIRMI),技术大学慕尼黑(TUM))

AI总结 Safe-Night VLA通过整合热感知技术,实现了安全关键操作中的非可见领域感知与稳健执行。

Comments Accepted to the 2026 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08068 2026-07-17 cs.CV 版本更新

Simulating Automotive Radar with Lidar and Camera Inputs

利用激光雷达和相机输入模拟汽车雷达

Peili Song, Dezhen Song, Yifan Yang, Enfan Lan, Jingtai Liu

机构 * Institute of Robotics and Automatic Information System, Nankai University(机器人与自动信息系统研究所,南开大学) Tianjin Key Laboratory of Intelligent Robotics(天津智能机器人重点实验室) TBI center, Nankai University(南开大学TBI中心) Department of Robotics, Mohamed Bin Zayed University of Artificial Intelligence (MBZUAI)(人工智能 Mohamed Bin Zayed 大学机器人系)

AI总结 研究利用相机图像、激光雷达点云和自身速度模拟汽车雷达信号的方法,基于DIS-Net和RSS-Net两个神经网络,经实验验证能生成高保真信号,用其增强数据训练的目标检测网络性能更优,助力基于雷达的自动驾驶研发。

Comments Accepted by 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.10703 2026-07-17 cs.RO 版本更新

OGM-CBF: Occupancy Grid Map-based Control Barrier Function for Safe Mobile Robot Control with Memory of out of View Obstacles

OGM-CBF:基于占用网格地图的控制屏障函数用于具有视 out of View 障碍物记忆的移动机器人安全控制

Golnaz Raja, Miloš Prágr, Topi Reino Johannes Kärki, Teemu Mökkönen, Reza Ghabcheloo

机构 * Faculty of Engineering and Natural Sciences, Tampere University(坦佩雷大学工程与自然科学学院)

AI总结 本文提出OGM-CBF方法,结合占用网格地图与控制屏障函数,解决移动机器人在未知环境中对已离开视野的障碍物记忆问题,提升安全性。

Comments Submitted to IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.29846 2026-07-16 cs.RO cs.HC 版本更新

Legible Shared Autonomy: Implicit Communication of Robot Belief through Motion

可理解的共享自主:通过运动隐式传达机器人信念

Jinwei Liu, Pengfei Li, Shaofeng Chen, Tao Wang, Yun-Bo Zhao

机构 * Department of Automation, University of Science and Technology of China(自动化系,中国科学技术大学) Department of Automation, Hefei University of Technology(自动化系,合肥工业大学) National Key Laboratory of Autonomous Intelligent Unmanned Systems(自主智能无人系统国家重点实验室)

AI总结 针对共享自主系统中机器人意图不透明导致用户控制效率低的问题,提出通过可理解运动隐式传达机器人信念,并基于置信度自适应分配控制权,实验证明该方法显著提升用户对机器人信念的理解并减少控制负担。

Comments Accepted at IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.11811 2026-07-16 cs.RO cs.AI cs.CV 版本更新

RADAR: Closed-Loop Robotic Data Generation via Semantic Planning and Autonomous Causal Environment Reset

RADAR: 通过语义规划和自主因果环境重置实现闭环机器人数据生成

Yongzhong Wang, Keyu Zhu, Yong Zhong, Liqiong Wang, Jinyu Yang, Feng Zheng

机构 * Southern University of Science and Technology(南方科技大学) Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)) Spatialtemporal AI(时空人工智能)

AI总结 RADAR通过语义规划和自主因果环境重置实现闭环机器人数据生成,具备高适应性和可扩展性,在仿真和现实部署中均表现出卓越的性能。

Comments 8 pages, 4 figures. Accepted to the 2026 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2026). Project page: https://radar-iros.netlify.app/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08503 2026-07-15 cs.CV cs.GR cs.RO eess.IV 版本更新

Spherical-GOF: Geometry-Aware Panoramic Gaussian Opacity Fields for 3D Scene Reconstruction

球面-GOF:面向3D场景重建的几何感知全景高斯不透明场

Zhe Yang, Guoqiang Zhao, Sheng Wu, Kai Luo, Kailun Yang

机构 * School of Artificial Intelligence and Robotics, Hunan University(人工智能与机器人学院,湖南大学) National Engineering Research Center of Robot Visual Perception and Control Technology, Hunan University(机器人视觉感知与控制技术国家工程研究中心,湖南大学)

AI总结 Spherical-GOF通过球面高斯不透明场实现全景图像的高效渲染,提升3D场景重建的几何一致性和光度质量。

Comments Accepted to IEEE/RSJ IROS 2026. The source code and dataset will be released at https://github.com/1170632760/Spherical-GOF

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01700 2026-07-15 cs.RO 版本更新

Tilt-Ropter: A Fully Actuated Hybrid Aerial-Terrestrial Vehicle with Tilt Rotors and Passive Wheels

Tilt-Ropter: 一种带有倾转旋翼和被动轮的全驱动混合空中-地面车辆

Ruoyu Wang, Xuchen Liu, Zongzhou Wu, Zixuan Guo, Wendi Ding, Ben M. Chen

机构 * Department of Mechanical and Automation Engineering, The Chinese University of Hong Kong(机械与自动化工程系,香港中文大学) Faculty of Engineering, The University of Hong Kong(工程学院,香港大学) Peng Cheng Laboratory(鹏城实验室)

AI总结 提出全驱动混合空中-地面车辆Tilt-Ropter,通过倾转旋翼和被动轮实现高效多模态运动,并设计统一非线性模型预测控制器实现低跟踪误差和地面运动功耗降低92.8%。

Comments 8 pages, 10 figures. Accepted by the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06796 2026-07-15 cs.RO 版本更新

RoboDesign1M: A Large-scale Dataset for Robot Design Understanding

RoboDesign1M:用于机器人设计理解的大规模数据集

Tri Le, Toan Nguyen, Quang Tran, Quang Nguyen, Baoru Huang, Hoan Nguyen, Minh Nhat Vu, Tung D. Ta, Anh Nguyen

机构 * FPT Software AI Center(FPT软件人工智能中心) University of Liverpool(利物浦大学) University of Information Technology(信息技术大学) Automation & Control Institute(自动化与控制研究所) Department of Creative Informatics(创意信息系) Faculty of Environment and Information Studies(环境与信息科学系)

AI总结 针对机器人设计领域缺乏大规模数据集的问题,介绍了含100万个样本的RoboDesign1M数据集,提出半自动数据收集管道,经多项实验评估,该数据集可推动机器人设计理解研究及相关自动化发展。

Comments 8 pages, accepted to IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.03987 2026-07-14 cs.RO 版本更新

Fast Asymptotically Optimal Kinodynamic Planning via Vectorization

通过向量化实现快速渐近最优运动动力学规划

Yitian Gao, Andrew Lu, Zachary Kingston

机构 * Department of Computer Science, Purdue University(普渡大学计算机科学系)

AI总结 研究复杂运动动力学系统的规划问题,提出基于JAX和XLA编译器的并行渐近最优运动动力学RRT规划器,结合AO - x元算法实现渐近最优,经理论分析和实验验证其性能优势。

Comments 8 pages, 5 figures, 4 tables. Accepted to IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.15890 2026-07-14 cs.RO 版本更新

Robust Fleet Sizing for Multi-UAV Inspection Missions under Synchronized Replacement Demand

多无人机巡检任务下同步替换需求的鲁棒编队

Vishal Ramesh, Antony Thomas

机构 * Robotics Research Center, IIIT Hyderabad(IIIT海得拉巴机器人研究中心)

AI总结 本文针对多无人机巡检任务中同步替换需求的鲁棒编队问题,提出一种闭式充分编队规则,通过增加缓冲无人机数量提升任务可靠性。

Comments Accepted to IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.15865 2026-07-14 cs.RO 版本更新

DTEA: A Dual-Topology Elastic Actuator Enabling Real-Time Switching Between Series and Parallel Compliance

DTEA:一种双拓扑弹性执行器,实现系列和并联顺应性之间的实时切换

Vishal Ramesh, Aman Singh, Shishir Kolathaya

AI总结 本文提出了一种新型双拓扑弹性执行器DTEA,能够实现在操作过程中系列执行器(SEA)与并联执行器(PEA)拓扑结构之间的动态切换,通过实验验证了其在负载下的鲁棒性和切换时间,展示了其在两种模式下的静态刚度和扰动抑制性能。

Comments Accepted to IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06279 2026-07-14 cs.CV cs.RO eess.IV 版本更新

Can we Trust Unreliable Voxels? Exploring 3D Semantic Occupancy Prediction under Label Noise

我们能信任不可靠的体素吗?在标签噪声下的3D语义占用预测探索

Wenxin Li, Kunyu Peng, Di Wen, Junwei Zheng, Jiale Wei, Mengfei Duan, Yuheng Zhang, Rui Fan, Kailun Yang

机构 * School of Artificial Intelligence and Robotics(人工智能与机器人学院) National Engineering Research Center of Robot Visual Perception and Control Technology(机器人视觉感知与控制技术国家工程研究中心) Institute for Anthropomatics and Robotics(人机一体化与机器人研究所) Karlsruhe Institute of Technology(卡尔斯鲁厄技术大学) INSAIT College of Electronic and Information Engineering(电子与信息学院) Shanghai Institute of Intelligent Science and Technology(智能科学与技术上海研究院) Shanghai Research Institute for Intelligent Autonomous Systems(智能自主系统上海研究院) Shanghai Key Laboratory of Intelligent Autonomous Systems(智能自主系统上海重点实验室) State Key Laboratory of Autonomous Intelligent Unmanned Systems(自主智能无人系统国家重点实验室) Frontiers Science Center for Intelligent Autonomous Systems(智能自主系统前沿科学中心)

AI总结 本文提出DPR-Occ框架,通过双源部分标签推理解决3D语义占用预测中的标签噪声问题,实验表明其在高噪声环境下仍能保持性能。

Comments Accepted to IROS 2026. The benchmark and source code will be made publicly available at https://github.com/mylwx/OccNL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10979 2026-07-14 cs.RO cs.SY eess.SY 版本更新

Autonomous Close-Proximity Photovoltaic Panel Coating Using a Quadcopter

使用四旋翼无人机进行近距离光伏板自主涂层作业

Dimitri Jacquemont, Carlo Bosio, Teaya Yang, Ruiqi Zhang, Ozgur Orun, Shuai Li, Reza Alam, Thomas M. Schutzius, Simo A. Makiharju, Mark W. Mueller

机构 * High Performance Robotics Laboratory, Department of Mechanical Engineering, University of California Berkeley(加州大学伯克利分校机械工程系高性能机器人实验室)

AI总结 研究利用四旋翼无人机对光伏板进行近距离自主涂层作业,提出基于视觉惯性里程计等的定位系统及考虑地面效应等的控制方法,经室内外实验验证了系统自主能力,为光伏板涂层作业提供新途径。

Comments 7 pages, 11 figures. Accepted to IEEE IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15254 2026-07-13 cs.RO 版本更新

OIPP: Object-Adaptive Impact Point Predictor for Catching Diverse In-Flight Objects

OIPP: 用于捕捉多样飞行物体的物体自适应冲击点预测器

Ngoc Huy Nguyen, Kazuki Shibata, Takamitsu Matsubara

机构 * Division of Information Science, Graduate School of Science and Technology, Nara Institute of Science and Technology (NAIST)(信息科学系,科技研究生学校,科学技术国立研究所(NAIST))

AI总结 OIPP通过物体自适应编码器和冲击点预测器,提升飞行物体捕捉的准确性和多样性。

Comments Accepted to IEEE/RSJ IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18999 2026-07-10 cs.RO cs.AI cs.CV 版本更新

OREN: Octree Residual Network for Real-Time Euclidean Signed Distance Mapping

OREN:八叉树残差网络用于实时欧几里得符号距离映射

Zhirui Dai, Qihao Qian, Tianxing Fan, Nikolay Atanasov

机构 * Department of Electrical and Computer Engineering, University of California San Diego(加州大学圣地亚哥分校电子与计算机工程系)

AI总结 OREN结合八叉树插值的显式先验和神经网络回归的隐式残差,实现高效且可微的非截断SDF重建,优于现有方法的准确性和效率。

Comments Accepted to IEEE/RSJ International Conference Intelligent Robots & Systems (IROS) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.02036 2026-07-10 cs.RO 版本更新

TurboMap: GPU-Accelerated Local Mapping for Visual SLAM

TurboMap: 面向视觉SLAM的GPU加速局部建图

Parsa Hosseininejad, Kimia Khabiri, Shishir Gopinath, Soudabeh Mohammadhashemi, Karthik Dantu, Steven Y. Ko

机构 * Simon Fraser University(西蒙弗雷泽大学) University at Buffalo(布法罗大学)

AI总结 针对视觉SLAM中局部建图延迟问题,提出GPU并行化与CPU优化结合的TurboMap后端,通过重构地图点创建、融合及关键帧管理,实现1.3-1.6倍加速且保持精度。

Comments Accepted for presentation at IROS 2026, preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10757 2026-07-10 cs.RO cs.DC 版本更新

FastTrack: GPU-Accelerated Tracking for Visual SLAM

FastTrack:用于视觉同步定位与地图构建的GPU加速跟踪

Kimia Khabiri, Parsa Hosseininejad, Shishir Gopinath, Karthik Dantu, Steven Y. Ko

机构 * Simon Fraser University(西蒙弗雷泽大学) University at Buffalo(布法罗大学)

AI总结 研究视觉惯性SLAM系统跟踪模块,提出利用GPU加速立体特征匹配和局部地图跟踪等耗时组件的方法,在ORB-SLAM3中用CUDA实现,经实验在特定数据集和设备上跟踪性能提升达2.8倍。

Comments Published at IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2025)

Journal ref 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), Hangzhou, China, pp. 15165-15172, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.29358 2026-07-09 cs.RO 版本更新

LAMP: Long-Horizon Adaptive Manipulation Planning for Multi-Robot Collaboration in Cluttered Space

LAMP: 杂乱空间中多机器人协作的长时域自适应操作规划

Shuai Zhou, Yorai Shaoul, Jiaoyang Li

机构 * Robotics Institute, Carnegie Mellon University(卡内基梅隆大学机器人研究所)

AI总结 提出LAMP框架,结合学习生成的操作模型与两种规划器(LAMPA*和LAMP-Lazy),实现多机器人在密集杂乱环境中的长时域协作操作规划,解决现有方法无法处理的复杂任务。

Comments IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.23800 2026-07-09 cs.RO cs.AI cs.LG 版本更新

Object Search in Partially-Known Environments via LLM-informed Model-based Planning and Prompt Selection

通过LLM引导的基于模型的规划与提示选择进行部分已知环境中的物体搜索

Abhishek Paudel, Abhish Khanal, Raihan I. Arnob, Shahriar Hossain, Gregory J. Stein

机构 * Department of Computer Science, George Mason University(乔治·马歇尔大学计算机科学系)

AI总结 本文提出了一种基于LLM的模型规划框架和提示选择方法,用于部分已知环境中的物体搜索。利用LLM估计不同位置搜索目标物体的可能性,并结合环境地图提取的旅行成本,以实现有效的搜索性能。

Comments 10 pages, 8 figures. Accepted to IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.06667 2026-07-09 cs.RO cs.AI 版本更新

Rapidly Learning Soft Robot Control via Implicit Time-Stepping

通过隐式时间步长快速学习软机器人控制

Andrew Choi, Dezhong Tong, Xiaonan Huang

机构 * Horizon Robotics University of Michigan(密歇根大学)

AI总结 研究针对软机器人模拟框架稀缺及策略学习难的问题,采用隐式时间步长,以DisMech模拟器及增量自然曲率控制方法,经与Elastica对比实验,证明该方法能在不牺牲准确性的情况下显著加速软机器人策略学习。

Comments Accepted to IROS 2026. Code: https://github.com/QuantuMope/dismech-rl

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06924 2026-07-08 cs.RO 版本更新

LIPP: Load-Aware Informative Path Planning with Physical Sampling

LIPP:基于物理采样的负载感知信息路径规划

Hojune Kim, Guangyao Shi, Gaurav S. Sukhatme

机构 * University of Southern California(南加州大学)

AI总结 研究针对物理样本采集场景中信息增益与负载相关遍历成本耦合问题,提出负载感知信息路径规划(LIPP),将其制定为混合整数二次规划,推导理论界限,经模拟验证其随样本质量增加能提升单位能量的不确定性降低能力。

Comments Accepted for presentation at IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.01642 2026-07-08 cs.RO 版本更新

FailSafe: Reasoning and Recovery from Failures in Vision-Language-Action Models

FailSafe: 视觉-语言-动作模型中的失败推理与恢复

Zijun Lin, Jiafei Duan, Haoquan Fang, Dieter Fox, Ranjay Krishna, Cheston Tan, Bihan Wen

机构 * Nanyang Technological University(南洋理工大学) Centre for Frontier AI Research, A*STAR(A*STAR前沿人工智能研究中心) Allen Institute for AI(艾伦人工智能研究所) University of Washington(华盛顿大学)

AI总结 提出FailSafe系统,自动生成多样化失败案例及可执行恢复动作,微调LLaVA-OV-7B构建FailSafe-VLM,使机器人检测并恢复失败,在ManiSkill任务上平均提升三个VLA模型性能达22.6%。

Comments IROS 2026. Project Page: https://jimntu.github.io/FailSafe

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07514 2026-07-08 cs.RO 版本更新

HJCD-IK: GPU-Accelerated Inverse Kinematics through Batched Hybrid Jacobian Coordinate Descent

HJCD-IK:通过批处理混合雅可比坐标下降实现GPU加速的逆运动学

Cael Yasutake, Andrew H. Liu, Zachary Kingston, Brian Plancher

机构 * Columbia University(哥伦比亚大学) Purdue University(普渡大学) Barnard College, Columbia University(巴纳德学院,哥伦比亚大学)

AI总结 研究机器人逆运动学问题,提出基于GPU加速、采样的混合求解器HJCD-IK,通过结合方向感知贪婪坐标下降初始化、雅可比优化和并行碰撞滤波器,在速度和精度上比现有求解器有显著提升,且能找到无碰撞解决方案并开源代码。

Comments Accepted to IROS 2026. 8 pages, 6 figures, 3 tables, 4 algorithms

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24966 2026-07-08 cs.CV 版本更新

Social 3D Scene Graphs: Modeling Human Actions and Relations for Interactive Service Robots

社交3D场景图:为交互式服务机器人建模人类行为与关系

Ermanno Bartoli, Dennis Rotondi, Buwei He, Patric Jensfelt, Kai O. Arras, Iolanda Leite

机构 * Faculty of Robotics Perception and Learning, KTH Royal Institute of Technology(机器人感知与学习学院,皇家理工学院) Institute for Artificial Intelligence, University of Stuttgart(人工智能研究所,斯图加特大学) International Max Planck Research School for Intelligent Systems (IMPRS-IS)(智能系统国际马克斯·普朗克研究学校)

AI总结 研究为使机器人能以符合社会规范和情境感知的方式行动,引入社交3D场景图,用开放词汇框架捕获环境中人类及其属性、活动和关系,还引入新基准,实验表明其改进了人类活动预测和人-环境关系推理。

Comments Equal contribution from E. Bartoli and D. Rotondi. Paper accepted at IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.28720 2026-07-07 cs.RO 版本更新

CubifyGS: Object-Centric 3D Gaussian Splatting for Lifelong Dynamic Scene Maintenance

CubifyGS: 面向对象的3D高斯泼溅用于终身动态场景维护

Bohan Ren, Dianyi Yang, Shiyang Liu, Yu Gao, Jiadong Tang, Zhilin Lai, Yi Yang, Mengyin Fu

机构 * School of Automation, Beijing Institute of Technology, Beijing, China(北京理工大学自动化学院,北京,中国) Guangzhou Saite Intelligent Technology Co., Ltd.(广州赛泰智能科技有限公司)

AI总结 提出CubifyGS,一种面向对象的映射框架,通过将可移动实例建模为可重用高斯资产,并采用事件触发自适应优化,实现刚性物体重排下的高效动态场景维护。

Comments Accepted to IROS 2026. 8 pages, 5 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.21241 2026-07-07 cs.RO cs.AI 版本更新

CorridorVLA: Explicit Spatial Constraints for Generative Action Heads via Sparse Anchors

CorridorVLA:通过稀疏锚点实现生成动作头的显式空间约束

Dachong Li, ZhuangZhuang Chen, Jin Zhang, Jianqiang Li

机构 * College of Computer Science and Software Engineering(计算机科学与软件工程学院) National Engineering Laboratory for Big Data System Computing Technology(大数据系统计算技术国家工程实验室)

AI总结 CorridorVLA通过稀疏锚点提供显式空间约束,提升动作生成性能,在LIBERO-Plus基准上改进成功率3.4%-12.4%。

Comments Accepted to the 2026 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.20659 2026-07-07 cs.RO 版本更新

StageCraft: Execution Aware Mitigation of Distractor and Obstruction Failures in VLA Models

StageCraft: 通过执行意识缓解VLA模型中干扰和障碍故障

Kartikay Milind Pangaonkar, Prabin Rath, Omkar Patil, Nakul Gopalan

机构 * Arizona State University(亚利桑那州立大学)

AI总结 StageCraft通过利用大规模视觉语言模型进行推理,改进预训练VLA策略性能,通过操控环境初始状态避免执行故障,实现在三个真实任务领域中性能提升40%。

Comments Accepted to IEEE International Conference on Intelligent Robots and Systems (IROS) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15061 2026-07-07 cs.RO cs.CV 版本更新

Ask-to-Clarify: Resolving Instruction Ambiguity through Multi-turn Dialogue

Ask-to-Clarify: 通过多轮对话解决指令歧义

Xingyao Lin, Xinghao Zhu, Tianyi Lu, Guojin Zhong, Sicheng Xie, Hui Zhang, Xipeng Qiu, Zuxuan Wu, Yu-Gang Jiang

机构 * College of Computer Science and Artificial Intelligence, Fudan University, Shanghai, China(复旦大学计算机科学与人工智能学院) Shanghai Innovation Institute, Shanghai, China(上海创新研究院) Mechanical Systems Control Lab, UC Berkeley, California, USA(伯克利机械系统控制实验室)

AI总结 本文提出Ask-to-Clarify框架,通过多轮对话解决指令歧义问题,结合视觉语言模型和扩散模型,采用两阶段知识绝缘策略训练,实现多任务中更高效的协作式具身代理。

Comments Accepted by IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11891 2026-07-03 cs.RO cs.SY eess.SY 版本更新

VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer

VLSA: 具有即插即用安全约束层的视觉-语言-动作模型

Songqiao Hu, Zeyi Liu, Shuang Liu, Jun Cen, Zihan Meng, Shihefeng Wang, Xiang Li, Xiao He

机构 * Department of Automation, Tsinghua University(清华大学自动化系) Institute for Embodied Intelligence and Robotics, Tsinghua University(清华大学智能与机器人研究院) TetraBOT DAMO Academy, Alibaba Group(阿里巴巴集团达摩院)

AI总结 提出AEGIS架构,通过控制屏障函数构建即插即用安全约束层,集成到VLA模型中保证安全性与任务性能,在SafeLIBERO基准上障碍物避免率提升超50%,任务成功率提升近10%。

Comments Accepted by IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏