arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

共收录 4115 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 模仿学习与强化学习 4115 篇

2305.16317 2023-10-17 cs.LG cs.AI 62%

Parallel Sampling of Diffusion Models

Andy Shih, Suneel Belkhale, Stefano Ermon, Dorsa Sadigh, Nima Anari

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments 37th Conference on Neural Information Processing Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.13424 2023-10-17 cs.LG cs.AI cs.NE 62%

Dealing with Sparse Rewards Using Graph Neural Networks

Matvey Gerasyov, Ilya Makarov

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

Journal ref IEEE Access, vol. 11, pp. 89180-89187, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.07899 2023-10-13 cs.AI cs.RO 62%

RoboCLIP: One Demonstration is Enough to Learn Robot Policies

Sumedh A Sontakke, Jesse Zhang, Sébastien M. R. Arnold, Karl Pertsch, Erdem Bıyık, Dorsa Sadigh, Chelsea Finn, Laurent Itti

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.03161 2023-10-06 cs.LG cs.AI 62%

Neural architecture impact on identifying temporally extended Reinforcement Learning tasks

Victor Vadakechirayath George

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments Master's thesis at Albert-Ludwigs-University, Freiburg Faculty of Engineering, Department of Computer Science Chair for Machine Learning. Advisor: Raghu Rajan, Examiners: Prof. Dr. Frank Hutter, Prof. Dr. Thomas Brox

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.04066 2023-10-02 eess.SY cs.LG cs.RO cs.SY 62%

Stable and Safe Reinforcement Learning via a Barrier-Lyapunov Actor-Critic Approach

Liqun Zhao, Konstantinos Gatsis, Antonis Papachristodoulou

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

Comments Accepted by IEEE CDC 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.16074 2023-09-29 cs.RO cs.LG 62%

Infer and Adapt: Bipedal Locomotion Reward Learning from Demonstrations via Inverse Reinforcement Learning

Feiyang Wu, Zhaoyuan Gu, Hanran Wu, Anqi Wu, Ye Zhao

专题命中 模仿学习与强化学习 :robot learning(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.08690 2023-09-28 cs.LG cs.AI 62%

Replay Buffer with Local Forgetting for Adapting to Local Environment Changes in Deep Model-Based Reinforcement Learning

Ali Rahimi-Kalahroudi, Janarthanan Rajendran, Ida Momennejad, Harm van Seijen, Sarath Chandar

专题命中 模仿学习与强化学习 :world model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.14727 2023-09-27 eess.SY cs.AI cs.LG cs.SY 62%

Effective Multi-Agent Deep Reinforcement Learning Control with Relative Entropy Regularization

Chenyang Miao, Yunduan Cui, Huiyun Li, Xinyu Wu

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.14096 2023-09-26 cs.LG cs.RO 62%

Tracking Control for a Spherical Pendulum via Curriculum Reinforcement Learning

Pascal Klink, Florian Wolf, Kai Ploeger, Jan Peters, Joni Pajarinen

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.05178 2023-09-26 cs.RO cs.LG 62%

Pre-Training for Robots: Offline RL Enables Learning New Tasks from a Handful of Trials

Aviral Kumar, Anikait Singh, Frederik Ebert, Mitsuhiko Nakamoto, Yanlai Yang, Chelsea Finn, Sergey Levine

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.11089 2023-09-21 eess.SY cs.AI cs.LG cs.SY 62%

Practical Probabilistic Model-based Deep Reinforcement Learning by Integrating Dropout Uncertainty and Trajectory Sampling

Wenjun Huang, Yunduan Cui, Huiyun Li, Xinyu Wu

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.16979 2023-09-21 cs.AI cs.LG 62%

Adaptive PD Control using Deep Reinforcement Learning for Local-Remote Teleoperation with Stochastic Time Delays

Luc McCutcheon, Saber Fallah

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

Comments 7 pages + 1 references, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.06692 2023-09-18 cs.LG cs.AI cs.CL 62%

Guiding Pretraining in Reinforcement Learning with Large Language Models

Yuqing Du, Olivia Watkins, Zihan Wang, Cédric Colas, Trevor Darrell, Pieter Abbeel, Abhishek Gupta, Jacob Andreas

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

Comments ICML 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.02976 2023-09-08 cs.RO cs.LG 62%

Natural and Robust Walking using Reinforcement Learning without Demonstrations in High-Dimensional Musculoskeletal Models

Pierre Schumacher, Thomas Geijtenbeek, Vittorio Caggiano, Vikash Kumar, Syn Schmitt, Georg Martius, Daniel F. B. Haeufle

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.14919 2023-09-04 cs.LG cs.AI 62%

On Reward Structures of Markov Decision Processes

Falcon Z. Dai

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

Comments This PhD thesis draws heavily from arXiv:1907.02114 and arXiv:2002.06299; minor edits

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.12438 2023-08-25 cs.LG cs.AI cs.SE 62%

Deploying Deep Reinforcement Learning Systems: A Taxonomy of Challenges

Ahmed Haj Yahmed, Altaf Allah Abbassi, Amin Nikanjam, Heng Li, Foutse Khomh

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments Accepted for publication in The International Conference on Software Maintenance and Evolution (ICSME 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.12270 2023-08-24 cs.LG cs.AI 62%

Language Reward Modulation for Pretraining Reinforcement Learning

Ademi Adeniji, Amber Xie, Carmelo Sferrazza, Younggyo Seo, Stephen James, Pieter Abbeel

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

Comments Code available at https://github.com/ademiadeniji/lamp

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.10797 2023-08-23 cs.LG cs.AI 62%

Stabilizing Unsupervised Environment Design with a Learned Adversary

Ishita Mediratta, Minqi Jiang, Jack Parker-Holder, Michael Dennis, Eugene Vinitsky, Tim Rocktäschel

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

Comments CoLLAs 2023 - Oral; Second and third authors contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.06236 2023-08-22 cs.MA cs.LG cs.RO 62%

iPLAN: Intent-Aware Planning in Heterogeneous Traffic via Distributed Multi-Agent Reinforcement Learning

Xiyang Wu, Rohan Chandra, Tianrui Guan, Amrit Singh Bedi, Dinesh Manocha

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.08520 2023-08-17 cs.CV cs.LG 62%

Painter: Teaching Auto-regressive Language Models to Draw Sketches

Reza Pourreza, Apratim Bhattacharyya, Sunny Panchal, Mingu Lee, Pulkit Madan, Roland Memisevic

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.CV、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.06569 2023-08-16 cs.LG cs.AI 62%

Policy Regularization with Dataset Constraint for Offline Reinforcement Learning

Yuhang Ran, Yi-Chen Li, Fuxiang Zhang, Zongzhang Zhang, Yang Yu

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

Comments Accepted to ICML 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.04958 2023-08-10 cs.AI cs.LG 62%

Improving Autonomous Separation Assurance through Distributed Reinforcement Learning with Attention Networks

Marc W. Brittain, Luis E. Alvarez, Kara Breeden

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

Comments 7 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.07813 2023-08-08 cs.SE cs.AI cs.LG 62%

A Search-Based Testing Approach for Deep Reinforcement Learning Agents

Amirhossein Zolfagharian, Manel Abdellatif, Lionel Briand, Mojtaba Bagherzadeh, Ramesh S

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Journal ref in IEEE Transactions on Software Engineering, vol. 49, no. 7, pp. 3715-3735, July 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.15393 2023-08-04 cs.RO cs.AI 62%

Programmable Control of Ultrasound Swarmbots through Reinforcement Learning

Matthijs Schrage, Mahmoud Medany, Daniel Ahmed

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO、cs.AI

Journal ref 2023 The Authors. Advanced Materials Technologies published by Wiley-VCH GmbH

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.14048 2023-08-03 cs.NE cs.AI cs.LG q-bio.NC 62%

Short-Term Plasticity Neurons Learning to Learn and Forget

Hector Garcia Rodriguez, Qinghai Guo, Timoleon Moraitis

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments Accepted at ICML 2022

Journal ref Proceedings of the 39th International Conference on Machine Learning, 162:18704-18722 (2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.15671 2023-07-31 cs.CV cs.RO 62%

TrackAgent: 6D Object Tracking via Reinforcement Learning

Konstantin Röhrl, Dominik Bauer, Timothy Patten, Markus Vincze

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.CV

Comments International Conference on Computer Vision Systems (ICVS) 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.14339 2023-07-31 cs.RO cs.AI 62%

Efficient Exploration Using Extra Safety Budget in Constrained Policy Optimization

Haotian Xu, Shengjie Wang, Zhaolei Wang, Yunzhe Zhang, Qing Zhuo, Yang Gao, Tao Zhang

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.AI

Comments 7 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.09246 2023-07-28 cs.RO cs.LG 62%

Experimental Study on Reinforcement Learning-based Control of an Acrobot

Leo Dostal, Alexej Bespalko, Daniel A. Duecker

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.13008 2023-07-25 cs.LG cs.HC cs.RO 62%

Imitation Learning with Human Eye Gaze via Multi-Objective Prediction

Ravi Kumar Thakur, MD-Nazmus Samin Sunbeam, Vinicius G. Goecks, Ellen Novoseller, Ritwik Bera, Vernon J. Lawhern, Gregory M. Gremillion, John Valasek, Nicholas R. Waytowich

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO、cs.LG

Comments Paper accepted and selected as an oral presentation at Interactive Learning with Implicit Human Feedback Workshop at ICML 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.05012 2023-07-19 cs.LG cs.AI 62%

PoPS: Policy Pruning and Shrinking for Deep Reinforcement Learning

Dor Livne, Kobi Cohen

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments This paper has been accepted for publication in the IEEE Journal of Selected Topics in Signal Processing

详情

展开后加载摘要…

URL PDF HTML 收藏