arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

2026-03-03 至 2026-03-03 共收录 147 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 机器人操作 47 篇

2603.01751 2026-03-03 cs.RO cs.AI cs.LG cs.SY eess.SY 67%

Shape-Interpretable Visual Self-Modeling Enables Geometry-Aware Continuum Robot Control

具有形状可解释性的视觉自建模使几何感知的连续体机器人控制成为可能

Peng Yu, Xin Wang, Ning Tan

机构 * School of Computer Science and Engineering, Sun Yat-sen University(计算机科学与工程学院,中山大学)

专题命中 机器人操作 :manipulation(abstract);分类 cs.RO、cs.AI、cs.LG

AI总结 本文提出一种基于形状可解释性的视觉自建模框架,通过贝塞尔曲线表示法实现连续体机器人几何感知的控制,实验验证了其在复杂环境中的鲁棒性和精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01151 2026-03-03 cs.RO cs.CV cs.GR 62%

D-REX: Differentiable Real-to-Sim-to-Real Engine for Learning Dexterous Grasping

D-REX:用于学习灵巧抓取的可微现实到仿真到现实引擎

Haozhe Lou, Mingtong Zhang, Haoran Geng, Hanyang Zhou, Sicheng He, Zhiyuan Gao, Siheng Zhao, Jiageng Mao, Pieter Abbeel, Jitendra Malik, Daniel Seita, Yue Wang

机构 * Physical Superintelligence (PSI) Lab, University of Southern California(南加州大学物理超智能实验室) Viterbi School of Engineering, University of Southern California(南加州大学韦伯尔工程学院) Department of EECS, University of California, Berkeley(加州大学伯克利分校电子工程与计算机科学系)

专题命中 机器人操作 :robotic(abstract);分类 cs.RO、cs.CV

AI总结 D-REX通过可微引擎实现现实到仿真到现实的抓取学习,利用高斯点表示自动构建数字双胞胎并优化质量识别,从而提升抓取策略的力感知性能。

Comments ICLR 2026 Poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05619 2026-03-03 cs.AI cs.LG 62%

Beyond RLHF and NLHF: Population-Proportional Alignment under an Axiomatic Framework

超越RLHF和NLHF:在公理框架下的人口比例对齐

Kihyun Kim, Jiawei Zhang, Asuman Ozdaglar, Pablo A. Parrilo

机构 * MIT LIDS(麻省理工学院LIDS) University of Wisconsin-Madison(威斯康星大学麦迪逊分校)

专题命中 机器人操作 :manipulation(abstract);分类 cs.AI、cs.LG

AI总结 本文提出了一种基于公理框架的人口比例对齐方法,通过社会选择理论解决传统偏好学习中的偏差和操纵问题,并在推荐任务和语言模型对齐中验证了其有效性。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00481 2026-03-03 cs.LG cs.CV 62%

Analyzing Physical Adversarial Example Threats to Machine Learning in Election Systems

分析对机器学习在选举系统中物理对抗示例威胁

Khaleque Md Aashiq Kamal, Surya Eada, Aayushi Verma, Subek Acharya, Adrian Yemin, Benjamin Fuller, Kaleel Mahmood

机构 * Department of Electrical, Computer and Biomedical Engineering, University of Rhode Island, Kingston, RI, USA(电气、计算机与生物医学工程系,罗德岛大学,金斯敦,RI,USA) Department of Computer Science and Statistics, University of Rhode Island, Kingston, RI, USA(计算机科学与统计学系,罗德岛大学,金斯敦,RI,USA) Voting Technology Center, University of Connecticut, Storrs, CT, USA(选举技术中心,康涅狄格大学,斯托尔斯,CT,USA)

专题命中 机器人操作 :manipulation(abstract);分类 cs.CV、cs.LG

AI总结 研究分析了对抗示例对选举系统中机器学习模型的物理威胁,通过实验揭示数字与物理领域对抗攻击效果的差异。

Comments 20 pages, 8 figures, 28 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00436 2026-03-03 cs.LG cs.AI 62%

ROKA: Robust Knowledge Unlearning against Adversaries

ROKA: 面对对抗者的鲁棒知识反学习

Jinmyeong Shin, Joshua Tapia, Nicholas Ferreira, Gabriel Diaz, Moayed Daneshyari, Hyeran Jeon

机构 * University of California, Merced(加州大学梅尔塞德斯分校) California State University, East Bay(加州州立大学东湾分校)

专题命中 机器人操作 :manipulation(abstract);分类 cs.AI、cs.LG

AI总结 ROKA通过神经愈合机制实现鲁棒的知识反学习,有效对抗间接反学习攻击,同时保护保留数据的准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16557 2026-03-03 cs.CV cs.ET cs.HC cs.LG 62%

Person Identification from Egocentric Human-Object Interactions using 3D Hand Pose

基于第一人称视角人类-物体交互的人员识别

Muhammad Hamza, Danish Hamid, Muhammad Tahir Akram

机构 * Department of Creative Technologies(创意技术系) Air University(空军大学)

专题命中 机器人操作 :manipulation(abstract);分类 cs.CV、cs.LG

AI总结 I2S通过3D手姿态分析实现基于第一人称视角的无感用户识别,结合多阶段特征增强和新颖描述符,达到97.52%的高识别率。

Comments 21 pages, 8 figures, 7 tables. Preprint of a manuscript to appear in CCF Trans. Pervasive Comp. Interact. (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01942 2026-03-03 cs.HC cs.AI 57%

Ignore All Previous Instructions: Jailbreaking as a de-escalatory peace building practise to resist LLM social media bots

忽略此前所有指令:将‘禁锢’视为一种缓和和平建设的实践,以抵抗大语言模型社交媒体机器人

Huw Day, Adrianna Jezierska, Jessica Woodgate

机构 * School of Engineering Maths & Technology University of Bristol(工程数学与科技学院 英国布里斯托尔大学) Business School University of Bristol(商学院 英国布里斯托尔大学) School of Computer Science University of Bristol(计算机科学学院 英国布里斯托尔大学)

专题命中 机器人操作 :manipulation(abstract);分类 cs.AI

AI总结 本文提出将‘禁锢’作为一种非暴力的缓和和平建设实践,通过用户与疑似LLM账号的互动来对抗大语言模型社交媒体机器人。

Comments Accepted to ICLR 2026 AI for peace workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14492 2026-03-03 cs.RO 57%

UNCLE-Grasp: Uncertainty-Aware Grasping of Leaf-Occluded Strawberries

UNCLE-Grasp: 有不确定性意识的叶片遮挡草莓抓取

Malak Mansour, Ali Abouzeid, Zezhou Sun, Qinbo Sun, Dezhen Song, Abdalla Swikir

机构 * Department of Robotics, Mohamed bin Zayed University of Artificial Intelligence(机器人系,Mohamed bin Zayed人工智能大学)

专题命中 机器人操作 :robotic(abstract);分类 cs.RO

AI总结 UNCLE-Grasp通过建模遮挡和形状补全的几何不确定性,提出了一种在部分遮挡下可靠抓取草莓的方法,通过多假设评估和保守置信界标准提升抓取可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25175 2026-03-03 cs.CL cs.AI 57%

EasySteer: A Unified Framework for High-Performance and Extensible LLM Steering

EasySteer: 一种高绩效且可扩展的LLM引导统一框架

Haolei Xu, Xinyu Mei, Yuchen Yan, Rui Zhou, Wenqi Zhang, Weiming Lu, Yueting Zhuang, Yongliang Shen

机构 * Zhejiang University(浙江大学)

专题命中 机器人操作 :manipulation(abstract);分类 cs.AI

AI总结 EasySteer通过模块化架构和预计算引导向量,实现高效且可扩展的LLM引导,提升推理效率并支持多种应用场景。

Comments Functionality upgrade. Code: https://github.com/ZJU-REAL/EasySteer Demo: https://www.youtube.com/watch?v=3rRGzZmhrXg

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11917 2026-03-03 cs.RO 57%

OneTwoVLA: A Unified Vision-Language-Action Model with Adaptive Reasoning

OneTwoVLA:一种具有自适应推理能力的统一视觉-语言-动作模型

Fanqi Lin, Ruiqian Nai, Yingdong Hu, Jiacheng You, Junming Zhao, Yang Gao

机构 * Tsinghua University(清华大学) Shanghai Qi Zhi Institute(上海启智研究院)

专题命中 机器人操作 :manipulation(abstract);分类 cs.RO

AI总结 OneTwoVLA通过自适应推理机制实现视觉-语言-动作统一,提升机器人在复杂任务中的规划、交互与执行能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27492 2026-03-03 cs.CV 57%

ThinkMorph: Emergent Properties in Multimodal Interleaved Chain-of-Thought Reasoning

ThinkMorph:多模态交错链式推理中的涌现特性

Jiawei Gu, Yunzhuo Hao, Huichen Will Wang, Linjie Li, Michael Qizhe Shieh, Yejin Choi, Ranjay Krishna, Yu Cheng

机构 * National University of Singapore(新加坡国立大学) Zhejiang University(浙江大学) University of Washington(华盛顿大学) Stanford University(斯坦福大学) absolute AI The Chinese University of Hong Kong(香港中文大学)

专题命中 机器人操作 :manipulation(abstract);分类 cs.CV

AI总结 ThinkMorph通过统一模型提升多模态推理性能,展现视觉操控与模式切换等新兴能力。

Comments project page: https://thinkmorph.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.08007 2026-03-03 cs.CV 57%

Velocity Disambiguation for Video Frame Interpolation

视频帧插值中的速度歧义消除

Zhihang Zhong, Yiming Zhang, Wei Wang, Xiao Sun, Yu Qiao, Gurunandan Krishnan, Sizhuo Ma, Jian Wang

机构 * School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院) Cornell University(康奈尔大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) OtoNexus Medical Technologies(OtoNexus医疗科技公司) Snap Inc(Snap公司)

专题命中 机器人操作 :manipulation(abstract);分类 cs.CV

AI总结 本文提出距离索引方法,通过显式提示对象移动距离来提升视频帧插值的精度和质量。

Comments ECCV2024 Oral; TPAMI

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00351 2026-03-03 cs.RO cs.AI cs.LG cs.SD 56%

Acoustic Sensing for Universal Jamming Grippers

用于通用阻塞夹持器的声学传感

Lion Weber, Theodor Wienert, Martin Splettstößer, Alexander Koenig, Oliver Brock

机构 * Robotics and Biology Laboratory, Technische Universität Berlin(技术大学柏林机器人与生物学实验室) Science of Intelligence, Research Cluster of Excellence, Berlin(柏林智能科学卓越研究中心) Robotics Institute Germany(德国机器人研究所)

专题命中 机器人操作 :分类 cs.RO、cs.AI、cs.LG;robotics(journal_ref)

AI总结 本文提出一种基于声学传感的通用阻塞夹持器,利用夹持器自身作为传感器,实现高精度物体识别与抓取性能。

Comments Accepted at ICRA 2026, supplementary material under https://rbo.gitlab-pages.tu-berlin.de/papers/acoustic-jamming-icra26/

Journal ref IEEE International Conference on Robotics and Automation (ICRA) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01609 2026-03-03 physics.optics physics.app-ph 50%

Doubly resonant nonlinear metasurfaces enabling NIR-to-UV upconversion for reconfigurable Fourier optical processing

双共振非线性超材料实现近红外到紫外上转换用于可重构傅里叶光学处理

Jumin Qiu, Meibao Qin, Tingting Liu, Lun Qu, Xintong Shi, Feng Wu, Tianbao Yu, Qiegen Liu, Shuyuan Xiao

专题命中 机器人操作 :manipulation(abstract)

AI总结 双共振非线性超材料实现近红外到紫外上转换,用于可重构傅里叶光学处理和全光图像处理。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01207 2026-03-03 cond-mat.supr-con cond-mat.mes-hall 50%

Superconducting diode effect in multichannel Majorana wires

多重通道马约拉纳线中的超导二极管效应

Sagar Santra, Dibyendu Samanta, Sudeep Kumar Ghosh

专题命中 机器人操作 :manipulation(abstract)

AI总结 多重通道 Rashba 纳米线中,通过自洽 Bogoliubov-de Gennes 形式主义研究了超导二极管效应,揭示了非对称配对和拓扑相稳定机制,实现了高效率非 reciprocity 传输和拓扑超导体操控。

Comments 15 pages and 12 figures. Comments are welcome

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00879 2026-03-03 cond-mat.mes-hall 50%

Temperature-driven enhancement and sign reversal of field-like torque in Py/FePS$_3$ bilayers

温度驱动的Py/FePS₃双层中场类扭矩增强与符号反转

Dhananjaya Mahapatra, Anudeepa Ghosh, Harekrishna Bhunia, Bipul Pal, Partha Mitra

专题命中 机器人操作 :manipulation(abstract)

AI总结 研究揭示Py/FePS₃双层中温度驱动的场类扭矩增强与符号反转,揭示反铁磁绝缘体在调控自旋轨道扭矩对称性与效率中的作用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00622 2026-03-03 cond-mat.mtrl-sci cond-mat.mes-hall 50%

Sliding Ferroelectricity Induced and Switched Altermagnetism in GaSe-VPSe3-GaSe Sandwiched Heterostructure with Strong Magnetoelectric Effect

滑动铁电性诱导和切换的GaSe-VPSe3-GaSe三明治异质结中的交替磁性与强磁电效应

Pengqiang Dong, Hanbo Sun, Chao Wu, Ping Li

专题命中 机器人操作 :manipulation(abstract)

AI总结 通过滑动铁电性操控GaSe-VPSe3-GaSe异质结实现强磁电耦合的交替磁体,揭示了磁相变的微观机制,为多铁电存储器设计提供新途径。

Comments 12 pages, 6 figures, Accepted Acta Materialia

Journal ref Acta Materialia (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16379 2026-03-03 cond-mat.mtrl-sci 50%

Stacking-tunable multiferroic states in bilayer ScI2

双层ScI₂中可调的多铁性态

Yaxin Pan, Chongze Wang, Shuyuan Liu, Fengzhu Ren, Chang Liu, Bing Wang, Jun-Hyung Cho

专题命中 机器人操作 :manipulation(abstract)

AI总结 双层ScI₂通过调控堆叠构型实现磁性、铁电性和谷极化的可调性。

Comments 7 figures

Journal ref Appl. Phys. Lett. 127, 221601 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.08078 2026-03-03 cs.SI cs.CY 50%

Political audience diversity and news reliability in algorithmic ranking

政治受众多样性与算法排名中的新闻可靠性

Saumya Bhadani, Shun Yamaya, Alessandro Flammini, Filippo Menczer, Giovanni Luca Ciampaglia, Brendan Nyhan

专题命中 机器人操作 :manipulation(abstract)

AI总结 本文提出利用政治受众多样性作为新闻可靠性信号,改进算法推荐,提升用户获取信息的可信度。

Comments 47 pages, 23 figures, 5 tables (including supplementary materials). Nat Hum Behav (2022)

Journal ref Nat Hum Behav 6, 495--505 (2022)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 具身导航 28 篇

2504.08806 2026-03-03 cs.AI cs.RO 87%

Endowing Embodied Agents with Spatial Reasoning Capabilities for Vision-and-Language Navigation

赋予具身体验智能体空间推理能力以实现视觉-语言导航

Qianqian Bai, Zhongpu Chen, Ling Luo, Huaming Du, Yuqian Lei, Ziyun Jiao

机构 * Southwestern University of Finance(西南财经大学) Department of Informatics,Universitat Hamburg, Hamburg, Germany(汉堡大学信息学院) University of Electronic Science(电子科技大学)

专题命中 具身导航 :navigation(title,abstract);embodied agent(title);分类 cs.RO、cs.AI

AI总结 BrainNav通过模仿生物认知功能,提升具身体验智能体的空间推理能力,实现更高效的视觉-语言导航。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02004 2026-03-03 cs.RO 85%

CHOP: Counterfactual Human Preference Labels Improve Obstacle Avoidance in Visuomotor Navigation Policies

CHOP:反事实人类偏好标签改进视觉运动导航策略中的障碍物避让

Gershom Seneviratne, Jianyu An, Vaibhav Shende, Sahire Ellahy, Yaxita Amin, Kondapi Manasanjani, Samarth Chopra, Jonathan Deepak Kannan, Dinesh Manocha

机构 * University of Maryland, College Park(马里兰大学学院公园分校)

专题命中 具身导航 :navigation(title,abstract);robotics(abstract);embodied agent(abstract);分类 cs.RO

AI总结 CHOP通过反事实人类偏好标签改进视觉运动导航策略的障碍物避让能力,显著提升安全性和导航效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18041 2026-03-03 cs.CV cs.RO 84%

Openfly: A comprehensive platform for aerial vision-language navigation

Openfly:面向空中视觉-语言导航的综合性平台

Yunpeng Gao, Chenhui Li, Zhongrui You, Junli Liu, Zhen Li, Pengan Chen, Qizhi Chen, Zhonghan Tang, Liansheng Wang, Penghui Yang, Yiwen Tang, Yuhang Tang, Shuai Liang, Songyi Zhu, Ziqin Xiong, Yifei Su, Xinyi Ye, Jianan Li, Yan Ding, Dong Wang, Xuelong Li, Zhigang Wang, Bin Zhao

机构 * Shanghai AI Laboratory(上海人工智能实验室) Northwestern Polytechnical University(西北工业大学) Beihang University(北航) Shanghai Jiao Tong University(上海交通大学) The University of Hong Kong(香港大学) Zhejiang University(浙江大学) University of Science and Technology of China(中国科学技术大学) East China University of Science and Technology(东华大学) Fudan University(复旦大学) Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) TeleAI

专题命中 具身导航 :navigation(title,abstract);embodied AI(abstract);分类 cs.RO、cs.CV

AI总结 OpenFly平台通过整合多种渲染引擎和自动化工具链,构建大规模空中VLN数据集,并提出关键帧感知的VLN模型,提升户外空中视觉-语言导航的研究与应用。

Comments accepted by ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01477 2026-03-03 cs.RO 80%

SFCo-Nav: Efficient Zero-Shot Visual Language Navigation via Collaboration of Slow LLM and Fast Attributed Graph Alignment

SFCo-Nav: 通过慢LLM与快速属性图对齐的协作实现高效的零样本视觉语言导航

Chaoran Xiong, Litao Wei, Xinhao Hu, Kehui Ma, Ziyi Xia, Zixin Jiang, Zhen Sun, Ling Pei

机构 * Shanghai Key Laboratory of Navigation and Location Based Services, Shanghai Jiao Tong University(上海导航与位置基于服务重点实验室,上海交通大学)

专题命中 具身导航 :navigation(title,abstract);分类 cs.RO;robotics(comments)

AI总结 SFCo-Nav通过慢LLM与快速属性图对齐的协作,实现了高效的零样本视觉语言导航,显著提升效率并降低计算成本。

Comments Accepted by 2026 IEEE International Conference on Robotics and Automation (ICRA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02079 2026-03-03 cs.CV 79%

MMNavAgent: Multi-Magnification WSI Navigation Agent for Clinically Consistent Whole-Slide Analysis

MMNavAgent: 多倍率WSI导航代理用于临床一致的全滑片分析

Zhengyang Xu, Han Li, Jingsong Liu, Linrui Xie, Xun Ma, Xin You, Shihui Zu, Ayako Ito, Xinyu Hao, Hongming Xu, Shaohua Kevin Zhou, Nassir Navab, Peter J. Schüffler

机构 * Institute of Pathology, Technical University of Munich, Germany(慕尼黑技术大学病理研究所) Munich Data Science Institute (MDSI), Munich, Germany(慕尼黑数据科学研究所) Munich Center for Machine Learning (MCML), Munich, Germany(慕尼黑机器学习中心) Computer Aided Medical Procedures (CAMP), TU Munich, Munich, Germany(计算机辅助医疗程序(CAMP),慕尼黑技术大学) Dalian University of Technology(大连理工大学) Cancer Hospital of Dalian University of Technology, Shenyang(大连理工大学肿瘤医院) Department of Human Pathology, Juntendo University Graduate School of Medicine(立命大学医学研究生院人类病理部门) University of Science and Technology of China(中国科学技术大学) Institute of Medical Robotics, Shanghai Jiao Tong University(上海交通大学医学机器人研究所) Northwest University of China(中国西北大学)

专题命中 具身导航 :navigation(title,abstract);分类 cs.CV

AI总结 MMNavAgent通过多倍率交互建模和自适应倍率选择,提升全滑片图像诊断性能,实验显示AUC和BACC均有所提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01898 2026-03-03 cs.RO 79%

SaferPath: Hierarchical Visual Navigation with Learned Guidance and Safety-Constrained Control

SaferPath: 基于学习引导和安全约束控制的分层视觉导航

Lingjie Zhang, Zeyu Jiang, Changhao Chen

机构 * PEAK-Lab, The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)的PEAK实验室)

专题命中 具身导航 :navigation(title,abstract);分类 cs.RO

AI总结 SaferPath通过结合学习引导和安全约束控制,提升移动机器人在复杂环境中的视觉导航能力,有效减少碰撞并提高成功率。

Comments ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01143 2026-03-03 cs.RO 79%

StraightTrack: Towards Mixed Reality Navigation System for Percutaneous K-wire Insertion

StraightTrack:面向混合现实导航系统的穿刺K线植入系统

Han Zhang, Benjamin D. Killeen, Yu-Chun Ku, Lalithkumar Seenivasan, Yuxuan Zhao, Mingxu Liu, Yue Yang, Suxi Gu, Alejandro Martin-Gomez, Russell H. Taylor, Greg Osgood, Mathias Unberath

机构 * Johns Hopkins University(约翰霍普金斯大学) Johns Hopkins Medicine(约翰霍普金斯医学)

专题命中 具身导航 :navigation(title,abstract);分类 cs.RO

AI总结 StraightTrack通过混合现实导航系统提高经皮K线植入的精度,减少线弯曲问题,提升骨折固定的可靠性。

Journal ref Healthcare Technology Letters Vol. 11 Issue 6 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01813 2026-03-03 cs.RO 79%

SSMG-Nav: Enhancing Lifelong Object Navigation with Semantic Skeleton Memory Graph

SSMG-Nav: 通过语义骨架记忆图增强终身物体导航

Haochen Niu, Lantao Zhang, Xingwu Ji, Rendong Ying, Peilin Liu, Fei Wen

机构 * Brain-Inspired Application Technology Center (BATC), Shanghai Jiao Tong University(脑启发应用技术中心(BATC),上海交通大学)

专题命中 具身导航 :navigation(title,abstract);分类 cs.RO

AI总结 SSMG-Nav通过语义骨架记忆图提升终身物体导航,整合多模态信息与持久记忆,优化路径效率与任务成功率。

Comments Accepted by 2026 ICRA

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.00507 2026-03-03 cs.RO 79%

Optimal-Horizon Social Robot Navigation in Heterogeneous Crowds

异质人群中的最优时间跨度社会机器人导航

Jiamin Shi, Haolin Zhang, Yuchen Yan, Shitao Chen, Jingmin Xin, Nanning Zheng

机构 * National Key Laboratory of Human-Machine Hybrid Augmented Intelligence(人机混合增强智能国家重点实验室) Nation Engineering Research Center for Visual Information and Applications(视觉信息与应用国家工程研究中心) Institute of Artificial Intelligence and Robotics(人工智能与机器人研究所) Xi’an Jiaotong University(西安交通大学)

专题命中 具身导航 :navigation(title,abstract);分类 cs.RO

AI总结 本文提出了一种基于社会条件优化的机器人导航框架,通过时空Transformer和强化学习策略动态调整预测时间跨度,提升在异质人群中的导航效率与安全性。

Comments 7 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01199 2026-03-03 cs.RO 72%

Closed-loop Control of Steerable Balloon Endoscopes for Robot-assisted Transcatheter Intracardiac Procedures

闭环控制可操控气囊内窥镜用于机器人辅助经导管心内操作

Max McCandless, Jonathan Hamid, Sammy Elmariah, Nathaniel Langer, Pierre E. Dupont

机构 * Department of Cardiac Surgery, Boston Children’s Hospital, Harvard Medical School, Boston, MA, USA(心脏外科部,波士顿儿童医院,哈佛医学院,马萨诸塞州波士顿,美国) Department of Medicine, Cardiology Division, University of California San Francisco, San Francisco, CA, USA(医学部,心脏病学分会,加州大学旧金山分校,旧金山,加利福尼亚州,美国) Department of Cardiac Surgery, Massachusetts General Hospital, Boston, MA, USA(心脏外科部,麻省总医院,波士顿,马萨诸塞州,美国)

专题命中 具身导航 :navigation(abstract);robotic(abstract);分类 cs.RO;robotics(journal_ref)

AI总结 本文提出了一种可操控气囊内窥镜,通过闭环控制实现心内手术中工具的精确导航与稳定方向控制。

Comments 8 pages, 11 figures

Journal ref IEEE Robotics and Automation Letters, Volume 11, Issue 4, Pages 4211-4218, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01584 2026-03-03 eess.IV eess.SP physics.med-ph 71%

MR-Compass: Inertial Navigation-Driven Motion Correction for Brain MRI

MR-Compass:基于惯性导航的脑部MRI运动校正

Musa Tunc Arslan, Fatih Calakli, Joshua Auger, Hongli Fan, Alan J Macy, Simon K Warfield

专题命中 具身导航 :navigation(title)

AI总结 MR-Compass利用MRI系统磁场和重力场直接估计头部运动,通过相位相关性校正运动,提升图像质量。

详情

展开后加载摘要…

URL PDF HTML 收藏