arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

共收录 4115 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 模仿学习与强化学习 4115 篇

2505.10482 2025-09-30 cs.LG cs.AI 62%

Fine-tuning Diffusion Policies with Backpropagation Through Diffusion Timesteps

Ningyuan Yang, Jiaxuan Gao, Feng Gao, Yi Wu, Chao Yu

机构 * Tsinghua University(清华大学)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments 9 pages for main text, 23 pages in total, submitted to Neurips, 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12243 2025-09-30 cs.RO cs.AI 62%

RISE: Robust Imitation through Stochastic Encoding

Mumuksh Tayal, Manan Tayal, Ravi Prakash

机构 * Cyber Physical Systems, Indian Institute of Science (IISc), Bengaluru(印度科学研究院计算机物理系统,班加罗尔)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.AI

Comments 15 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22130 2025-09-29 cs.MA cs.AI cs.LG 62%

Multi-Agent Path Finding via Offline RL and LLM Collaboration

Merve Atasever, Matthew Hong, Mihir Nitin Kulkarni, Qingpei Li, Jyotirmoy V. Deshmukh

机构 * Department of Computer Science, University of Southern California(计算机科学系,南加州大学)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19301 2025-09-29 cs.RO cs.LG 62%

Residual Off-Policy RL for Finetuning Behavior Cloning Policies

Lars Ankile, Zhenyu Jiang, Rocky Duan, Guanya Shi, Pieter Abbeel, Anusha Nagabandi

机构 * Amazon FAR (Frontier AI & Robotics)(亚马逊前沿人工智能与机器人实验室) Stanford University(斯坦福大学) Carnegie Mellon University(卡内基梅隆大学) UC Berkeley(加州大学伯克利分校)

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.RO、cs.LG

Comments Project website: https://residual-offpolicy-rl.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15173 2025-09-24 cs.CV cs.AI 62%

AvatarShield: Visual Reinforcement Learning for Human-Centric Synthetic Video Detection

Zhipei Xu, Xuanyu Zhang, Qing Huang, Xing Zhou, Jian Zhang

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.17274 2025-09-23 cs.RO cs.LG math.OC 62%

Learning and Optimization with 3D Orientations

Alexandros Ntagkas, Constantinos Tsakonas, Chairi Kiourt, Konstantinos Chatzilygeroudis

机构 * Computational Intelligence Laboratory (CILab), Department of Mathematics, University of Patras(计算智能实验室(CILab),数学系,帕特拉大学) Laboratory of Automation and Robotics (LAR) in the Department of Electrical & Computer Engineering, University of Patras(自动化与机器人实验室(LAR),电气与计算机工程系,帕特拉大学) Archimedes/Athena RC, Greece(阿基米德/雅典RC,希腊) Athena - Research and Innovation Center in Information, Communication and Knowledge Technologies, Xanthi, Greece(雅典-信息、通信和知识技术研究与创新中心,辛提,希腊)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

Comments 9 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11367 2025-09-16 cs.LG cs.AI 62%

Detecting Model Drifts in Non-Stationary Environment Using Edit Operation Measures

Chang-Hwan Lee, Alexander Shim

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments 28 pages, 3 figures, 17 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.08257 2025-09-12 cs.RO cs.AI 62%

Symmetry-Guided Multi-Agent Inverse Reinforcement Learning

Yongkai Tian, Yirong Qi, Xin Yu, Wenjun Wu, Jie Luo

机构 * State Key Laboratory of Complex & Critical Software Environment, Beihang University, Beijing, China(复杂与关键软件环境国家重点实验室,北京航空航天大学)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.AI

Comments 8pages, 6 figures. Accepted for publication in the Proceedings of the 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2025) as oral presentation

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.05735 2025-09-09 cs.LG cs.AI 62%

Offline vs. Online Learning in Model-based RL: Lessons for Data Collection Strategies

Jiaqi Chen, Ji Shi, Cansu Sancaktar, Jonas Frey, Georg Martius

专题命中 模仿学习与强化学习 :world model(abstract);分类 cs.AI、cs.LG

Comments Accepted at Reinforcement Learning Conference (RLC 2025); Code available at: https://github.com/swsychen/Offline_vs_Online_in_MBRL

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.04628 2025-09-08 cs.RO cs.AI 62%

Action Chunking with Transformers for Image-Based Spacecraft Guidance and Control

Alejandro Posadas-Nava, Andrea Scorsoglio, Luca Ghilardi, Roberto Furfaro, Richard Linares

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO、cs.AI

Comments 12 pages, 6 figures, 2025 AAS/AIAA Astrodynamics Specialist Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.22524 2025-09-05 cs.RO cs.AI 62%

Robust Offline Imitation Learning Through State-level Trajectory Stitching

Shuze Wang, Yunpeng Mei, Hongjie Cao, Yetian Yuan, Gang Wang, Jian Sun, Jie Chen

机构 * National Key Lab of Autonomous Intelligent Unmanned Systems, Beijing Institute of Technology, Beijing 100081, China(自主智能无人系统国家重点实验室,北京理工大学,北京100081,中国) Department of Control Science and Engineering, Harbin Institute of Technology, Harbin 150001, China(控制科学与工程学院,哈尔滨工业大学,哈尔滨150001,中国)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00574 2025-09-03 cs.RO cs.LG 62%

Learning Dolly-In Filming From Demonstration Using a Ground-Based Robot

Philip Lorimer, Alan Hunter, Wenbin Li

机构 * Department of Computer Science, University of Bath, UK(计算机科学系,巴斯大学)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

Comments Preprint; under double-anonymous review. 6 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.08955 2025-09-03 cs.LG cs.AI math.OC 62%

Towards Efficient Risk-Sensitive Policy Gradient: An Iteration Complexity Analysis

Rui Liu, Anish Gupta, Erfaun Noorani, Pratap Tokekar

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21101 2025-09-01 cs.LG cs.AI 62%

Beyond Prediction: Reinforcement Learning as the Defining Leap in Healthcare AI

Dilruk Perera, Gousia Habib, Qianyi Xu, Daniel J. Tan, Kai He, Erik Cambria, Mengling Feng

机构 * National University of Singapore(新加坡国立大学) Nanyang Technological University(南洋理工大学)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

Comments 40 pages in total (including appendix)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.15876 2025-08-28 cs.RO cs.AI 62%

Bidirectional Task-Motion Planning Based on Hierarchical Reinforcement Learning for Strategic Confrontation

Qizhen Wu, Lei Chen, Kexin Liu, Jinhu Lu

机构 * School of Automation Science and Electrical Engineering, Beihang University(自动化科学与电气工程学院,北京航空航天大学) Advanced Research Institute of Multidisciplinary Sciences and State Key Laboratory of CNS/ATM, Beijing Institute of Technology(多学科科学先进研究院和 CNS/ATM 国家重点实验室,北京理工大学)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.12244 2025-08-28 cs.RO cs.LG 62%

To the Noise and Back: Diffusion for Shared Autonomy

Takuma Yoneda, Luzhe Sun, Ge Yang, Bradly Stadie, Matthew Walter

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

Comments https://diffusion-for-shared-autonomy.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.04920 2025-08-26 cs.RO cs.LG cs.SY eess.SY 62%

Sim-to-Real Transfer of Deep Reinforcement Learning Agents for Online Coverage Path Planning

Arvi Jonnarth, Ola Johansson, Jie Zhao, Michael Felsberg

机构 * Department of Electrical Engineering, Linköping University, Sweden(瑞典林雪平大学电气工程系) Department of Information and Communication Engineering, Dalian University of Technology, China(中国大连理工大学信息与通信工程学院) School of Engineering, University of KwaZulu-Natal, Durban, South Africa(南非夸祖鲁-纳塔尔大学工程学院)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

Comments Published in IEEE Access

Journal ref IEEE Access, 2025, volume 13, pages 106883-106905

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07071 2025-08-14 cs.LG cs.AI 62%

Retrieval-Augmented Decision Transformer: External Memory for In-context RL

Thomas Schmied, Fabian Paischer, Vihang Patil, Markus Hofmarcher, Razvan Pascanu, Sepp Hochreiter

机构 * ELLIS Unit, LIT AI Lab, Institute for Machine Learning, JKU Linz, Austria(ELLIS单位、LIT人工智能实验室、机器学习研究所、JKU林茨、奥地利) Extensity AI Google DeepMind(谷歌DeepMind) Mila - Québec AI Institute(魁北克人工智能研究所) NXAI

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.07826 2025-08-12 cs.LG cs.AI 62%

On the Sample Efficiency of Abstractions and Potential-Based Reward Shaping in Reinforcement Learning

Giuseppe Canonaco, Leo Ardon, Alberto Pozanco, Daniel Borrajo

机构 * AI Research Dept. JPMorganChase Madrid, ES(摩根大通西班牙马德里人工智能研究部) AI Research Dept. JPMorganChase London, UK(摩根大通英国伦敦人工智能研究部)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04280 2025-08-07 cs.LG cs.AI 62%

Enhancing Vision-Language Model Training with Reinforcement Learning in Synthetic Worlds for Real-World Success

George Bredis, Stanislav Dereka, Viacheslav Sinii, Ruslan Rakhimov, Daniil Gavrilov

机构 * George Bredis(无) Stanislav Dereka(无) Viacheslav Sinii(无) Ruslan Rakhimov(无) Daniil Gavrilov(无)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02219 2025-08-05 cs.RO cs.LG 62%

CO-RFT: Efficient Fine-Tuning of Vision-Language-Action Models through Chunked Offline Reinforcement Learning

Dongchi Huang, Zhirui Fang, Tianle Zhang, Yihang Li, Lin Zhao, Chunhe Xia

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.05833 2025-08-05 cs.RO cs.LG 62%

Refined Policy Distillation: From VLA Generalists to RL Experts

Tobias Jülg, Wolfram Burgard, Florian Walter

机构 * Department of Computer Science & Artificial Intelligence, University of Technology Nuremberg(计算机科学与人工智能系,图恩大学)

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.RO、cs.LG

Comments accepted for publication at IROS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15897 2025-08-01 cs.CV cs.LG 62%

Learning 3D Scene Analogies with Neural Contextual Scene Maps

Junho Kim, Gwangtak Bae, Eun Sun Lee, Young Min Kim

机构 * Dept. of Electrical and Computer Engineering, Seoul National University(电子与计算机工程系,首尔国立大学) Interdisciplinary Program in Artificial Intelligence and INMC, Seoul National University(人工智能跨学科项目及INMC,首尔国立大学)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.CV、cs.LG

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18883 2025-07-28 cs.AI cs.RO 62%

Success in Humanoid Reinforcement Learning under Partial Observation

Wuhao Wang, Zhiyong Chen

机构 * School of Engineering, The University of Newcastle(工程学院)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.AI

Comments 11 pages, 3 figures, and 4 tables. Not published anywhere else

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17070 2025-07-24 cs.LG cs.AI 62%

Advancing Robustness in Deep Reinforcement Learning with an Ensemble Defense Approach

Adithya Mohan, Dominik Rößle, Daniel Cremers, Torsten Schön

机构 * AImotion Bavaria, Technische Hochschule Ingolstadt(巴伐利亚AImotion、英格尔施泰特技术大学) School of Computation, Information and Technology, TU Munich(计算、信息与技术学院,慕尼黑技术大学)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments 6 pages, 4 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17055 2025-07-24 cs.RO cs.LG 62%

Shared Control of Holonomic Wheelchairs through Reinforcement Learning

Jannis Bähler, Diego Paez-Granados, Jorge Peña-Queralta

机构 * Swiss Paraplegic Research, SPF(瑞士瘫痪研究机构) SCAI Lab, D-HEST, Swiss Federal School of Technology in Zurich - ETH Zurich(SCAI实验室,瑞士联邦理工学院-苏黎世-ETH Zurich) Centre for Artificial Ingelligece, Zurich University of Applied Sciences - ZHAW. Switzerland(人工智能中心,瑞士应用科学大学-ZHAW)

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.11566 2025-07-17 cs.NE cs.AI cs.RO 62%

Emergent Heterogeneous Swarm Control Through Hebbian Learning

Fuda van Diggelen, Tugay Alperen Karagüzel, Andres Garcia Rincon, A. E. Eiben, Dario Floreano, Eliseo Ferrante

机构 * Computer Science(计算机科学) École Polytechnique Fédérale de Lausanne(瑞士联邦理工学院) New York University Abu Dhabi(纽约大学阿布扎克分校)

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09495 2025-07-15 cs.AI cs.ET cs.HC cs.RO cs.SY eess.SY 62%

GenAI-based Multi-Agent Reinforcement Learning towards Distributed Agent Intelligence: A Generative-RL Agent Perspective

Hang Wang, Junshan Zhang

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.AI

Comments Position paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.08707 2025-07-14 cs.LG cs.RO 62%

SPLASH! Sample-efficient Preference-based inverse reinforcement learning for Long-horizon Adversarial tasks from Suboptimal Hierarchical demonstrations

Peter Crowley, Zachary Serlin, Tyler Paine, Makai Mann, Michael Benjamin, Calin Belta

机构 * Boston University(波士顿大学) MIT Lincoln Laboratory(麻省理工学院林肯实验室) MIT Pavlab(麻省理工学院 Pavlab 实验室) Woods Hole Oceanographic Institution(伍兹霍尔海洋研究所) University of Maryland, College Park(马里兰大学学院公园分校)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06615 2025-07-10 cs.LG cs.AI 62%

Efficient Multi-Task Reinforcement Learning with Cross-Task Policy Guidance

Jinmin He, Kai Li, Yifan Zang, Haobo Fu, Qiang Fu, Junliang Xing, Jian Cheng

机构 * Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Tsinghua University(清华大学) Tencent AI Lab(腾讯AI实验室)

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

Comments NeurIPS2024

Journal ref 38th Conference on Neural Information Processing Systems (NeurIPS 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏