arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

共收录 4115 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 模仿学习与强化学习 4115 篇

2005.01643 2020-11-03 cs.LG cs.AI stat.ML 62%

Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Sergey Levine, Aviral Kumar, George Tucker, Justin Fu

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.14575 2020-10-29 cs.RO cs.LG cs.SY eess.SY 62%

Learning Time Reduction Using Warm Start Methods for a Reinforcement Learning Based Supervisory Control in Hybrid Electric Vehicle Applications

Bin Xu, Jun Hou, Junzhe Shi, Huayi Li, Dhruvang Rathod, Zhe Wang, Zoran Filipi

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.14274 2020-10-28 cs.AI cs.LG 62%

Behavior Priors for Efficient Reinforcement Learning

Dhruva Tirumala, Alexandre Galashov, Hyeonwoo Noh, Leonard Hasenclever, Razvan Pascanu, Jonathan Schwarz, Guillaume Desjardins, Wojciech Marian Czarnecki, Arun Ahuja, Yee Whye Teh, Nicolas Heess

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

Comments Submitted to Journal of Machine Learning Research (JMLR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.06506 2020-10-27 cs.LG cs.AI stat.ML 62%

First Order Constrained Optimization in Policy Space

Yiming Zhang, Quan Vuong, Keith W. Ross

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.10740 2020-10-22 cs.LG cs.RO cs.SY eess.SY 62%

Safety Verification of Model Based Reinforcement Learning Controllers

Akshita Gupta, Inseok Hwang

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.09933 2020-10-21 cs.LG cs.AI 62%

Proximal Policy Gradient: PPO with Policy Gradient

Ju-Seung Byun, Byungmoon Kim, Huamin Wang

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments 7 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.09108 2020-10-20 cs.LG cs.AI cs.MA q-fin.PM 62%

Bridging the gap between Markowitz planning and deep reinforcement learning

Eric Benhamou, David Saltiel, Sandrine Ungari, Abhishek Mukhopadhyay

专题命中 模仿学习与强化学习 :robot learning(abstract);分类 cs.AI、cs.LG

Comments 10 pages, ICAPS PRL. arXiv admin note: substantial text overlap with arXiv:2009.14136, arXiv:2010.08497

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.07467 2020-10-16 cs.RO cs.LG 62%

Human-guided Robot Behavior Learning: A GAN-assisted Preference-based Reinforcement Learning Approach

Huixin Zhan, Feng Tao, Yongcan Cao

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.04816 2020-10-13 cs.LG cs.AI 62%

Characterizing Policy Divergence for Personalized Meta-Reinforcement Learning

Michael Zhang

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

Comments Deep Reinforcement Learning Workshop, NeurIPS 2019; Workshop on Meta-Learning (Meta-Learn), NeurIPS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.11212 2020-10-08 cs.RO cs.LG 62%

Robust Reinforcement Learning-based Autonomous Driving Agent for Simulation and Real World

Péter Almási, Róbert Moni, Bálint Gyires-Tóth

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

Comments \c{opyright} 2020 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.01825 2020-09-24 cs.LG cs.MA cs.RO stat.ML 62%

Robust Reinforcement Learning using Adversarial Populations

Eugene Vinitsky, Yuqing Du, Kanaad Parvate, Kathy Jang, Pieter Abbeel, Alexandre Bayen

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.04063 2020-09-24 cs.RO cs.LG 62%

Goal-Conditioned Variational Autoencoder Trajectory Primitives with Continuous and Discrete Latent Codes

Takayuki Osa, Shuhei Ikemoto

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

Comments 8 pages, SN Computer Science

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.09593 2020-09-22 cs.LG cs.AI 62%

Dynamic Horizon Value Estimation for Model-based Reinforcement Learning

Junjie Wang, Qichao Zhang, Dongbin Zhao, Mengchen Zhao, Jianye Hao

专题命中 模仿学习与强化学习 :world model(abstract);分类 cs.AI、cs.LG

Comments 9 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.12001 2020-09-17 cs.LG cs.AI stat.ML 62%

AutoFS: Automated Feature Selection via Diversity-aware Interactive Reinforcement Learning

Wei Fan, Kunpeng Liu, Hao Liu, Pengyang Wang, Yong Ge, Yanjie Fu

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

Comments Accepted by ICDM 2020. In this version, we revised some typos or mistakes for camera-ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.13712 2020-09-01 cs.RO cs.AI cs.SY eess.SY 62%

Control of a Nature-inspired Scorpion using Reinforcement Learning

Aakriti Agrawal, V S Rajashekhar, Rohitkumar Arasanipalai, Debasish Ghose

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO、cs.AI

Comments 4 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.13193 2020-09-01 cs.CV cs.RO 62%

Learn by Observation: Imitation Learning for Drone Patrolling from Videos of A Human Navigator

Yue Fan, Shilei Chu, Wei Zhang, Ran Song, Yibin Li

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO、cs.CV

Comments Accepted by IROS 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.12693 2020-08-31 cs.LG cs.AI stat.ML 62%

Sample Efficiency in Sparse Reinforcement Learning: Or Your Money Back

Trevor A. McInroe

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.12647 2020-08-31 cs.LG cs.AI 62%

ADAIL: Adaptive Adversarial Imitation Learning

Yiren Lu, Jonathan Tompson

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.05088 2020-08-13 cs.LG cs.AI 62%

An ocular biomechanics environment for reinforcement learning

Julie Iskander, Mohammed Hossny

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments 5 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.01593 2020-08-05 cs.LG cs.RO stat.ML 62%

Learning Transition Models with Time-delayed Causal Relations

Junchi Liang, Abdeslam Boularias

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.10504 2020-07-22 cs.AI cs.LG stat.ML 62%

Battlesnake Challenge: A Multi-agent Reinforcement Learning Playground with Human-in-the-loop

Jonathan Chung, Anna Luo, Xavier Raffin, Scott Perry

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.05196 2020-07-13 cs.LG cs.AI stat.ML 62%

Pre-trained Word Embeddings for Goal-conditional Transfer Learning in Reinforcement Learning

Matthias Hutsebaut-Buysse, Kevin Mets, Steven Latré

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

Comments Paper accepted to the ICML 2020 Language in Reinforcement Learning (LaReL) Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.15085 2020-06-29 cs.LG cs.AI stat.ML 62%

What can I do here? A Theory of Affordances in Reinforcement Learning

Khimya Khetarpal, Zafarali Ahmed, Gheorghe Comanici, David Abel, Doina Precup

专题命中 模仿学习与强化学习 :embodied agent(abstract);分类 cs.AI、cs.LG

Comments Thirty-seventh International Conference on Machine Learning (ICML 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.11751 2020-06-24 cs.LG cs.AI stat.ML 62%

Sample Factory: Egocentric 3D Control from Pixels at 100000 FPS with Asynchronous Reinforcement Learning

Aleksei Petrenko, Zhehui Huang, Tushar Kumar, Gaurav Sukhatme, Vladlen Koltun

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments Paper published in ICML2020. Visualizations of trained policies can be found at https://sites.google.com/view/sample-factory

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.06061 2020-06-24 cs.LG cs.AI stat.ML 62%

TOMA: Topological Map Abstraction for Reinforcement Learning

Zhao-Heng Yin, Wu-Jun Li

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.09939 2020-06-18 cs.LG cs.AI 62%

Forgetful Experience Replay in Hierarchical Reinforcement Learning from Demonstrations

Alexey Skrynnik, Aleksey Staroverov, Ermek Aitygulov, Kirill Aksenov, Vasilii Davydov, Aleksandr I. Panov

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.06636 2020-06-17 cs.AI cs.RO 62%

Catch & Carry: Reusable Neural Controllers for Vision-Guided Whole-Body Tasks

Josh Merel, Saran Tunyasuvunakool, Arun Ahuja, Yuval Tassa, Leonard Hasenclever, Vu Pham, Tom Erez, Greg Wayne, Nicolas Heess

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.07043 2020-06-15 cs.LG cs.AI cs.CL stat.ML 62%

Language-Conditioned Goal Generation: a New Approach to Language Grounding for RL

Cédric Colas, Ahmed Akakzia, Pierre-Yves Oudeyer, Mohamed Chetouani, Olivier Sigaud

专题命中 模仿学习与强化学习 :embodied agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.03970 2020-06-08 cs.LG cs.AI stat.ML 62%

Reinforcement Learning in Non-Stationary Environments

Sindhu Padakandla, Prabuchandran K. J, Shalabh Bhatnagar

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI、cs.LG

Journal ref Applied Intelligence 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.05838 2020-05-28 cs.LG cs.AI cs.NE stat.ML 62%

Goal-conditioned Imitation Learning

Yiming Ding, Carlos Florensa, Mariano Phielipp, Pieter Abbeel

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI、cs.LG

Comments Published at NeurIPS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏