arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

2026-05-12 至 2026-05-12 共收录 47 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 机器人操作 47 篇

2605.08774 2026-05-12 cs.RO cs.LG 88%

ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation

ProcVLM:基于学习的程序导向进度奖励用于机器人操作

Youhe Feng, Hansen Shi, Haoyang Li, Xinlei Guo, Yang Wang, Chengyang Zhang, Jinkai Zhang, Xiaohan Zhang, Jie Tang, Jing Zhang

机构 * School of Information, Renmin University of China(中国人民大学信息学院) Zhipu AI(智谱AI) Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)

专题命中 机器人操作 :manipulation(title,abstract);robotic(title,abstract);分类 cs.RO、cs.LG

AI总结 ProcVLM通过程序结构和阶段内视觉变化学习密集的进度奖励,提升机器人操作的程序推理能力,优于现有基线方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01008 2026-05-12 cs.RO 88%

DexWrist: A Robotic Wrist for Constrained and Dynamic Manipulation

DexWrist:一种用于受限和动态操作的机械腕

Martin Peticco, Gabriella Ulloa, John Marangola, Nitish Dashora, Pulkit Agrawal

机构 * Improbable AI Lab, Massachusetts Institute of Technology(Improbable AI实验室,麻省理工学院)

专题命中 机器人操作 :manipulation(title,abstract);robotic(title,abstract);分类 cs.RO

AI总结 DexWrist通过结合准直接驱动和解耦并行运动机制,在紧凑设计中实现高扭矩和动态接触任务,提升了受限环境中的操作性能。

Comments 9 pages, 8 figures. Submitted to RA-L 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22088 2026-05-12 cs.RO 87%

Force Policy: Learning Hybrid Force-Position Control Policy under Interaction Frame for Contact-Rich Manipulation

力政策:在交互框架下学习混合力-位置控制策略以进行密集接触 manipulation

Hongjie Fang, Shirun Tang, Mingyu Mei, Haoxiang Qin, Zihao He, Jingjing Chen, Ying Feng, Chenxi Wang, Wanxi Liu, Zaixing He, Cewu Lu, Shiquan Wang

机构 * Noematrix Flexiv Shanghai Jiao Tong University(上海交通大学) Zhejiang University(浙江大学) Shanghai Innovation Institute(上海创新研究院)

专题命中 机器人操作 :manipulation(title,title_cn);分类 cs.RO

AI总结 本文提出力政策,通过视觉和力反馈实现全局-局部控制,提升接触稳定性和执行质量,适用于复杂接触任务。

Comments accepted by RSS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.11674 2026-05-12 cs.RO cs.AI 87%

AffordSim: A Scalable Data Generator and Benchmark for Affordance-Aware Robotic Manipulation

AffordSim:一种可扩展的数据生成器和基准,用于面向 affordance 的机器人操作

Mingyang Li, Haofan Xu, Haowen Sun, Xinzhe Chen, Sihua Ren, Liqi Huang, Xinyang Sui, Chenyang Miao, Jiawei Ye, Qiongjie Cui, Zeyang Liu, Xingyu Chen, Xuguang Lan

机构 * School of Artificial Intelligence, Xi’an Jiaotong University(西安交通大学人工智能学院)

专题命中 机器人操作 :manipulation(title,abstract);robotic(title);分类 cs.RO、cs.AI

AI总结 AffordSim 通过整合开放词汇 3D affordance 预测,解决了机器人操作中接触信息获取的挑战,实现了高成功率的轨迹生成和跨仿真到现实的迁移。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.09789 2026-05-12 cs.RO 86%

Zero-Shot Sim-to-Real Robot Learning: A Dexterous Manipulation Study on Reactive Catching

零样本仿真到现实机器人学习:关于反应性接住的灵巧操作研究

Kejia Ren, Gaotian Wang, Andrew S. Morgan, Kaiyu Hang

机构 * Department of Computer Science, Rice University(计算机科学系,里士大学) Robotics and AI Institute(机器人与人工智能研究所)

专题命中 机器人操作 :manipulation(title,abstract);robot learning(title);分类 cs.RO

AI总结 本文提出DRIS方法,通过同时表示和传播多个随机实例,提升策略鲁棒性,减少现实世界微调需求,在反应性接住任务中实现零样本仿真到现实迁移。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14022 2026-05-12 cs.RO cs.AI 86%

Language Conditioned Multi-Finger Dexterous Manipulation Enabled by Physical Compliance and Switching of Controllers

通过物理合规性和控制器切换实现语言条件下的多指灵巧操作

Cheng Pan, Kai Junge, Benhui Dai, Qinghua Guan, Josie Hughes

机构 * Embodied AI(具身人工智能)

专题命中 机器人操作 :manipulation(title,abstract);robotics(abstract);robotic(abstract);分类 cs.RO、cs.AI

AI总结 本文提出一种结合高层视觉语言动作模型与小型控制模型的切换控制器,通过事件驱动机制实现多指灵巧操作的鲁棒控制,验证了硬件级合规性在提升接触稳定性方面的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.14125 2026-05-12 cs.CV cs.AI cs.RO 85%

HiVLA: A Visual-Grounded-Centric Hierarchical Embodied Manipulation System

HiVLA:一种以视觉为基础的分层具身操作系统

Tianshuo Yang, Guanyu Chen, Yutian Chen, Zhixuan Liang, Yitian Liu, Zanxin Chen, Chunpu Xu, Haotian Liang, Jiangmiao Pang, Yao Mu, Ping Luo

机构 * The University of Hong Kong(香港大学) Shanghai AI Laboratory(上海人工智能实验室) Shanghai Jiao Tong University(上海交通大学) The Chinese University of Hong Kong(香港中文大学)

专题命中 机器人操作 :manipulation(title,abstract);robotic(abstract);分类 cs.RO、cs.AI、cs.CV

AI总结 HiVLA提出一种分层框架,将高层语义规划与底层动作控制解耦,通过视觉 grounding 和 Diffusion Transformer 专家提升机器人操作能力,实验证明其在长时序技能组合和精细操作中表现优异。

Comments Project Page: https://tianshuoy.github.io/HiVLA-page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.21971 2026-05-12 cs.RO cs.AI cs.LG 84%

Supervised Mixture-of-Experts for Surgical Grasping and Retraction

监督混合专家架构用于手术抓取与牵开

Lorenzo Mazza, Ariel Rodriguez, Rayan Younis, Martin Lelis, Ortrun Hellig, Chenpan Li, Sebastian Bodenstedt, Martin Wagner, Stefanie Speidel

机构 * Department of Translational Surgical Oncology, NCT/UCC Dresden, Faculty of Medicine and University Hospital Carl Gustav Carus, TUD, Dresden, Germany(translational Surgical Oncology部门,NCT/UCC Dresden,医学院和Carl Gustav Carus大学医院,TUD,德累斯顿,德国) National Center for Tumor Diseases (NCT), NCT/UCC Dresden, a partnership between DKFZ, Faculty of Medicine and University Hospital Carl Gustav Carus, TUD, and HZDR, Dresden, Germany(肿瘤疾病国家中心(NCT),NCT/UCC德累斯顿,DKFZ、医学院和Carl Gustav Carus大学医院、TUD以及HZDR之间的合作,德累斯顿,德国) Faculty of Computer Science, TUD, Dresden, Germany(计算机科学系,TUD,德累斯顿,德国) The Center for Tactile Internet with Human-in-the-Loop (CeTI), TUD, Dresden, Germany(带有Human-in-the-Loop的触觉互联网中心(CeTI),TUD,德累斯顿,德国) Department of Visceral, Thoracic and Vascular Surgery, Faculty of Medicine and University Hospital Carl Gustav Carus, TUD, Dresden, Germany(visceral、胸腔和血管外科部门,医学院和Carl Gustav Carus大学医院,TUD,德累斯顿,德国)

专题命中 机器人操作 :robotics(abstract,comments);robot learning(abstract);manipulation(abstract);robotic(abstract)

AI总结 本文提出一种监督混合专家架构,用于结构化手术任务,通过轻量动作解码器政策实现复杂操作,展示了在少样本下优于传统方法的性能。

Comments Accepted at Robotics:Science and Systems 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.08799 2026-05-12 cs.RO 83%

ElasticFlow: One-Step Physics-Consistent Policy with Elastic Time Horizons for Language-Guided Manipulation

ElasticFlow: 语言引导操作中的弹性时间 horizon 的一步物理一致策略

Kewei Chen, Yayu Long, Shuai Li, Mingsheng Shang

机构 * Chongqing Institute of Green and Intelligent Technology, Chinese Academy of Sciences(中国科学院重庆绿色智能技术研究所) Chongqing School, University of Chinese Academy of Sciences(中国科学院大学重庆校区) Faculty of Information Technology and Electrical Engineering, University of Oulu, Finland(芬兰奥卢大学信息科技与电气工程学院)

专题命中 机器人操作 :manipulation(title);embodied AI(abstract);robotic(abstract);分类 cs.RO

AI总结 ElasticFlow通过弹性时间 horizon 机制和直接建模平均速度场,实现高效的一步映射,提升语言引导操作的物理一致性和效率。

Comments Accepted to Findings of ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.08727 2026-05-12 cs.CV cs.AI cs.LG 82%

Control Your View: High-Resolution Global Semantic Manipulation in Learned Image Compression

掌控你的视角:高分辨率全球语义操控在学习图像压缩中的应用

Jiaming Liang, Chi-Man Pun, Weisi Lin, Greta Seng Peng Mok

机构 * University of Macau(澳门大学) Nanyang Technological University(南洋理工大学)

专题命中 机器人操作 :manipulation(title,abstract);分类 cs.AI、cs.CV、cs.LG

AI总结 本文提出了一种新的方法,用于在学习图像压缩中实现高分辨率全球语义操控,通过改进的周期几何衰减调度策略,克服了现有方法在高分辨率语义操控中的局限性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.21926 2026-05-12 cs.RO 79%

Information Filtering via Variational Regularization for Robot Manipulation

通过变分正则化进行机器人操作的信息过滤

Jinhao Zhang, Wenlong Xia, Yaojia Wang, Zhexuan Zhou, Huizhe Li, Yichen Lai, Haoming Song, Youmin Gong, Jie Mei

机构 * Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)) Shanghai Jiao Tong University(上海交通大学)

专题命中 机器人操作 :manipulation(title);robotic(abstract);分类 cs.RO

AI总结 本文提出变分正则化方法,通过在U-Net和DiT中随机掩码骨干特征或跳过中间层,减少中间特征中的无关噪声,提升机器人操作性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12878 2026-05-12 cs.CV cs.RO 79%

Uni-Hand: Universal Hand Motion Forecasting in Egocentric Views

Uni-Hand:面向亲身体验视角的通用手部运动预测

Junyi Ma, Wentao Bao, Jingyi Xu, Guanzhong Sun, Yu Zheng, Erhang Zhang, Xieyuanli Chen, Hesheng Wang

机构 * IRMV Lab, Shanghai Jiao Tong University(上海交通大学IMV实验室) Meta Reality Labs(Meta现实实验室) Department of Electronic Engineering, Shanghai Jiao Tong University(上海交通大学电子工程系) China University of Mining and Technology(中国矿业大学) College of Intelligence Science and Technology, National University of Defense Technology(国防科技大学智能科学与技术学院)

专题命中 机器人操作 :manipulation(abstract);robot policy(abstract);robotic(abstract);分类 cs.RO、cs.CV

AI总结 本文提出Uni-Hand框架,通过多模态输入、多维多目标预测和多任务赋能,实现手部在2D和3D空间的精准预测,并引入目标指标预测手腕和指尖关节点,同时预测手-物体交互状态,提升下游任务性能。

Comments Accepted by T-PAMI 2026. Code and data: https://github.com/IRMVLab/UniHand

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11604 2026-05-12 cs.CL 78%

Talk to Your Slides: High-Efficiency Slide Editing via Language-Driven Structured Data Manipulation

与幻灯片对话:通过语言驱动的结构化数据操作实现高效幻灯片编辑

Kyudan Jung, Hojun Cho, Jooyeol Yun, Soyoung Yang, Jaehyeok Jang, Jaegul Choo

机构 * KAIST AI(韩国科学技术院人工智能研究所) Chung-Ang University(Chung-Ang 大学)

专题命中 机器人操作 :manipulation(title,abstract)

AI总结 本文提出Talk-to-Your-Slides,通过语言驱动的结构化数据操作实现高效幻灯片编辑,相较于基于图像的GUI方法,其在文本处理和格式任务中更快、更准确且成本更低。

Comments 30 pages, Accepted at ACL2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.09066 2026-05-12 cond-mat.mes-hall 78%

Manipulation of magnetic skyrmions by non-uniform electric fields

通过非均匀电场操控磁单畴

N. I. Simchuck, I. S. Burmistrov, S. S. Apostoloff

专题命中 机器人操作 :manipulation(title,abstract)

AI总结 本文提出利用局部电场操控Néel型磁单畴的理论方法,通过电场实现单畴的生成、驱动与湮灭,以及更复杂的纹理结构,建立相图并讨论实际应用可行性。

Comments 16 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.00458 2026-05-12 cs.SE 78%

LDMDroid: Leveraging LLMs for Detecting Data Manipulation Errors in Android Apps

LDMDroid:利用LLM检测Android应用中的数据操纵错误

Xiangyang Xiao, Huaxun Huang, Rongxin Wu

专题命中 机器人操作 :manipulation(title,abstract)

AI总结 本文提出利用LLM开发LDMDroid框架,通过状态感知过程生成UI事件序列以提高数据操纵功能触发成功率,并利用视觉特征识别数据状态变化,有效检测数据操纵错误。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.01632 2026-05-12 physics.optics 78%

Orbit-orbit photonics: Harnessing vortex-trajectory interplay for light manipulation

轨道-轨道光子学:利用涡旋轨迹相互作用实现光操控

Raghvendra P. Chaudhary, Imon Kalyan, Nir Shitrit

专题命中 机器人操作 :manipulation(title,abstract)

AI总结 研究揭示了光的内禀轨道角动量与外在轨道角动量相互作用,通过等离子体椭圆腔实现涡旋依赖的光束位移,拓展了光操控工具箱。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.04541 2026-05-12 cs.CV 77%

Angle-I2P: Angle-Consistent-Aware Hierarchical Attention for Cross-Modality Outlier Rejection

Angle-I2P:基于角度一致性的层次注意力交叉模态异常剔除

Muyao Peng, Shun Zou, Pei An, You Yang, Qiong Liu

机构 * School of Electronic Information and Communications, Huazhong University of Science and Technology(华中科技大学电子信息与通信学院)

专题命中 机器人操作 :manipulation(abstract,abstract_cn);robotic(abstract);分类 cs.CV

AI总结 本文提出Angle-I2P方法,通过角度一致性的几何约束和层次注意力机制提升跨模态异常剔除性能,实验表明在多个数据集上均取得最佳效果。

Comments Accepted by ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.06225 2026-05-12 cs.LG cs.AI 76%

Memory Inception: Latent-Space KV Cache Manipulation for Steering LLMs

记忆 inception:用于引导 LLMs 的潜在空间 KV 缓存操作

Andy Zeyi Liu, Michael Zhang, Ilana Greenberg, Adam Alnasser, Lucas Baker, John Sous

机构 * Yale University(耶鲁大学) Princeton University(普林斯顿大学) Jump Trading

专题命中 机器人操作 :manipulation(title);分类 cs.AI、cs.LG

AI总结 本文提出 memory inception 方法,通过在选定层插入文本衍生的键值对(KV)银行,在潜在注意力空间中引导 LLMs,实现更高效的控制与更少的存储消耗。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.09672 2026-05-12 cs.RO 74%

MVB-Grasp: Minimum-Volume-Box Filtering of Diffusion-based Grasps for Frontal Manipulation

MVB-Grasp:基于扩散模型的抓取生成中最小体积盒过滤用于正前方操作

Bibek Poudel, Abdul Basit, Muhammad Shafique

机构 * Unitree(单位树) Intel(英特尔)

专题命中 机器人操作 :manipulation(title);分类 cs.RO

AI总结 本文提出MVB-Grasp方法,通过引入最小体积包围盒几何先验,提升低成本机械臂在受限工作空间中的正前方抓取成功率,结合几何过滤与重评分函数,验证了其在Z1机械臂上的有效性。

Comments 8 pages, 12 figures, accepted to IJCNN 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00548 2026-05-12 cs.CV cs.GR 74%

Colorful-Noise: Training-Free Low-Frequency Noise Manipulation for Color-Based Conditional Image Generation

彩色噪声:无需训练的低频噪声操控用于基于颜色的条件图像生成

Nadav Z. Cohen, Ofir Abramovich, Ariel Shamir

机构 * Reichman University(雷曼大学)

专题命中 机器人操作 :manipulation(title);分类 cs.CV

AI总结 本文研究了扩散模型输入噪声的特性,发现低频成分主导图像全局结构和颜色,高频频成分控制细节。通过低频图像先验操控低频噪声,实现无需训练的条件生成,控制整体结构和颜色,保留高频细节多样性。

Comments SIGGRAPH 2026 Conference Paper. Project Page at: https://nadavc220.github.io/colorful-noise/

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.09954 2026-05-12 cs.RO cs.CV 73%

JODA: Composable Joint Dynamics for Articulated Objects

JODA:可组合的联合动力学用于连杆物体

Tianhong Gao, Cheng Yu, Yinghao Xu, Mengyu Chu

机构 * Peking University(北京大学) Ant Group, Robbyant(蚂蚁集团,Robbyant)

专题命中 机器人操作 :embodied AI(abstract);manipulation(abstract);分类 cs.RO、cs.CV

AI总结 JODA通过结构化三通道场生成关节级动力学,结合视觉和语言模型推断并优化关节动力学,实现可控的连杆物体动力学建模。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.06222 2026-05-12 cs.RO cs.AI 73%

When to Trust Imagination: Adaptive Action Execution for World Action Models

何时相信想象:为世界动作模型的自适应动作执行

Rui Wang, Yue Zhang, Jiehong Lin, Kuncheng Luo, Jianan Wang, Zhongrui Wang, Xiaojuan Qi

机构 * Southern University of Science and Technology(南方科技大学) The University of Hong Kong(香港大学)

专题命中 机器人操作 :manipulation(abstract);robotic(abstract);分类 cs.RO、cs.AI

AI总结 本文提出FFDC和混合时间 horizon训练,通过预测-观察一致性实现自适应动作执行,提升机器人在复杂场景中的鲁棒性和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.09999 2026-05-12 cs.RO cs.PF cs.SY eess.SY 72%

Muninn: Your Trajectory Diffusion Model But Faster

Muninn:你的轨迹扩散模型但更快

Gokul Puthumanaillam, Hao Jiang, Ruben Hernandez, Jose Fuentes, Paulo Padrao, Leonardo Bobadilla, Melkior Ornik

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Florida International University(佛罗里达国际大学) Providence College(普罗维登斯学院)

专题命中 机器人操作 :manipulation(abstract);navigation(abstract);分类 cs.RO;robotics(comments)

AI总结 本文提出Muninn,一种无需训练的缓存包装器,通过利用扩散轨迹规划器的两个信号,减少去噪器计算,实现轨迹扩散模型的实时应用,提升速度并保证性能和安全性。

Comments Accepted to Robotics: Science and Systems 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.10734 2026-05-12 cs.LG 70%

XQCfD: Accelerating Fast Actor-Critic Algorithms with Prior Data and Prior Policies

XQCfD:利用先验数据和先验策略加速快速Actor-Critic算法

Daniel Palenicek, Florian Vogt, Joe Watson, Ingmar Posner, Danica Kragic, Jan Peters

机构 * Technical University of Darmstadt(技术大学达姆施塔特) KTH Royal Institute of Technology(皇家理工学院) University of Oxford(牛津大学) German Research Center for AI (DFKI)(德国人工智能研究中心(DFKI)) Robotics Institute Germany (RIG)(德国机器人研究所)

专题命中 机器人操作 :manipulation(abstract);robotic(abstract);分类 cs.LG

AI总结 本文提出XQCfD算法,通过先验数据、预训练策略和静止策略架构提升样本效率,在稀疏奖励任务中取得最佳性能。

Comments 22 pages, 10 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.09487 2026-05-12 cs.LG 70%

Kintsugi: Learning Policies by Repairing Executable Knowledge Bases

Kintsugi: 通过修复可执行知识库学习策略

Teng Cao, Yu Deng, Hikaru Shindo, Quentin Delfosse, Lanxi Wen, Suli Wang, Jannis Blüml, Christopher Tauchmann, Kristian Kersting

机构 * Artificial Intelligence and Machine Learning Lab, Technical University of Darmstadt, Germany(德累斯顿技术大学人工智能与机器学习实验室) Hessian Center for Artificial Intelligence (hessian.AI), Germany(黑森人工智能中心) Department of Computer Science, Technical University of Darmstadt, Germany(德累斯顿技术大学计算机科学系) Department of Computer Science, Technical University of Munich (TUM), Germany(慕尼黑技术大学计算机科学系) German Research Center for Artificial Intelligence (DFKI), Germany(德国人工智能研究中心) Centre for Cognitive Science, Technical University of Darmstadt, Germany(德累斯顿技术大学认知科学中心)

专题命中 机器人操作 :embodied agent(abstract);manipulation(abstract);分类 cs.LG

AI总结 Kintsugi提出一种白盒策略学习框架,通过验证器门控构建可执行知识库,以可组合的类型条目表示任务级策略知识,并通过局部类型编辑提升知识库,实现可检查、可编辑和可部署的策略改进。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.00884 2026-05-12 cs.CV 70%

LiteVLA-H: Dual-Rate Vision-Language-Action Inference for Onboard Aerial Guidance and Semantic Perception

LiteVLA-H: 双速率视觉-语言-动作推断用于机载空中引导与语义感知

Justin williams, Kishor Datta Gupta, Roy George, Mrinmoy Sarkar

机构 * Department of Cyber Physical Systems, Clark Atlanta University(克劳克阿特拉大学网络物理系统系)

专题命中 机器人操作 :manipulation(abstract,abstract_cn);分类 cs.CV

AI总结 LiteVLA-H通过双速率操作在边缘设备上实现高效视觉-语言-动作推断,兼顾快速动作输出与语义理解,提升空中引导与场景感知性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.08268 2026-05-12 cs.MA cs.AI 70%

Insider Attacks in Multi-Agent LLM Consensus Systems

多智能体大语言模型共识系统中的内部攻击

Xiaolin Sun, Zixuan Liu, Yibin Hu, Zizhan Zheng

机构 * Department of XXX, University of YYY, Location, Country(XXX系,YYY大学,Location,Country) School of ZZZ, Institute of WWW, Location, Country(ZZZ学院,WWW研究所,Location,Country) Department of Computer Science, Tulane University, New Orleans, United States of America(计算机科学系, Tulane大学,新奥尔良,美国)

专题命中 机器人操作 :manipulation(abstract);world model(abstract);分类 cs.AI

AI总结 研究多智能体大语言模型共识系统中的内部攻击问题,提出基于世界模型的框架,通过学习良性智能体的潜在行为状态并利用强化学习训练攻击者,有效降低共识率并延长分歧时间。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.09998 2026-05-12 cs.LG cs.AI 62%

Continual Harness: Online Adaptation for Self-Improving Foundation Agents

持续Harness:面向自我改进基础代理的在线适应

Seth Karten, Joel Zhang, Tersoo Upaa, Ruirong Feng, Wenzhe Li, Chengshuai Shi, Chi Jin, Kiran Vodrahalli

机构 * Princeton University(普林斯顿大学) ARISE Foundation(ARISE基金会) Google DeepMind(谷歌DeepMind)

专题命中 机器人操作 :embodied agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出持续Harness,一种无需人工干预的自我改进机制,通过在线适应提升具身代理在长视界部分可观察决策中的表现,实验证明其在Pokémon游戏中显著降低操作成本并接近专家水平。

Comments 28 pages, 19 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.09258 2026-05-12 cs.CV cs.AI 62%

Monocular Biomechanical Tracking of Fingers with Inverse Kinematics to Foundation Models

单目生物力学追踪手指:结合逆运动学与基础模型

R. James Cotton, Pouyan Firouzabadi, Wendy Murray

机构 * Shirley Ryan AbilityLab Department of PM\&R Northwestern University Shirley Ryan AbilityLab Department of Biomedical Engineering Northwestern University

专题命中 机器人操作 :manipulation(abstract);分类 cs.AI、cs.CV

AI总结 本文提出结合SAM 3D Body基础模型与逆运动学优化的方法,实现单目视频中手指关节角度的生物力学追踪,验证结果表明在多视角重建中误差较小,扩展了单目生物力学分析的应用范围。

Comments Accepted to EMBC 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01774 2026-05-12 cs.RO cs.SY eess.SY 61%

MOBIUS: A Multi-Modal Bipedal Robot that can Walk, Crawl, Climb, and Roll

MOBIUS:一种能够行走、爬行、攀爬和滚动的多模态双足机器人

Alexander Schperberg, Yusuke Tanaka, Stefano Di Cairano, Dennis Hong

机构 * Mitsubishi Electric Research Laboratories(三菱电机研究实验室) Robotic Systems Lab(机器人系统实验室) Robotics and Mechanisms Laboratory(机器人与机构实验室) Department of Mechanical and Aerospace Engineering, University of California, Los Angeles(加州大学洛杉矶分校机械与航空航天工程系)

专题命中 机器人操作 :manipulation(abstract);分类 cs.RO;robotics(comments)

AI总结 MOBIUS机器人通过四条肢体实现多种运动模式切换,结合强化学习与混合规划架构,实现动态攀爬与负载支撑,拓展了移动操作与抓取能力。

Comments Paper is accepted at the Robotics: Science and Systems conference, held in Sydney, Australia, July 13th-17th, 2026. Alexander Schperberg and Yusuke Tanaka are co-first authors. Both were at the Robotics and Mechanisms Laboratory (RoMeLa) at UCLA when the work started, and are now with Mitsubishi Electric Research Laboratories and ETH Zurich (RSL) respectively

详情

展开后加载摘要…

URL PDF HTML 收藏