arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

共收录 4116 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 模仿学习与强化学习 4116 篇

2009.03349 2020-09-10 cs.LG cs.SY eess.SY stat.ML 57%

Deep Learning and Reinforcement Learning for Autonomous Unmanned Aerial Systems: Roadmap for Theory to Deployment

Jithin Jagannath, Anu Jagannath, Sean Furman, Tyler Gwin

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

Comments Preprint of Book Chapter to be published in Springer

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.05441 2020-09-01 cs.LG cs.MA stat.ML 57%

Delay-Aware Multi-Agent Reinforcement Learning for Cooperative and Competitive Environments

Baiming Chen, Mengdi Xu, Zuxin Liu, Liang Li, Ding Zhao

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.11811 2020-08-28 cs.LG math.OC stat.ML 57%

Constrained Markov Decision Processes via Backward Value Functions

Harsh Satija, Philip Amortila, Joelle Pineau

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.11543 2020-08-28 cs.RO cs.MA 57%

Continuous Deep Hierarchical Reinforcement Learning for Ground-Air Swarm Shepherding

Hung The Nguyen, Tung Duy Nguyen, Vu Phi Tran, Matthew Garratt, Kathryn Kasmarik, Sreenatha Anavatti, Michael Barlow, Hussein A. Abbass

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.09062 2020-08-18 stat.ML cs.CR cs.LG 57%

Adversarial Reinforcement Learning under Partial Observability in Autonomous Computer Network Defence

Yi Han, David Hubczenko, Paul Montague, Olivier De Vel, Tamas Abraham, Benjamin I. P. Rubinstein, Christopher Leckie, Tansu Alpcan, Sarah Erfani

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.LG

Comments 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.02521 2020-08-07 cs.RO 57%

Deep Reinforcement Learning based Local Planner for UAV Obstacle Avoidance using Demonstration Data

Lei He, Nabil Aouf, James F. Whidborne, Bifeng Song

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO

Comments Please find video demos at https://www.youtube.com/watch?v=4Zj49QtDdks

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.02435 2020-08-04 cs.LG stat.ML 57%

A Nonparametric Off-Policy Policy Gradient

Samuele Tosatto, Joao Carvalho, Hany Abdulsamad, Jan Peters

专题命中 模仿学习与强化学习 :robot learning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.03827 2020-07-22 cs.LG cs.CR cs.SY eess.SY math.OC stat.ML 57%

Manipulating Reinforcement Learning: Poisoning Attacks on Cost Signals

Yunhan Huang, Quanyan Zhu

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.LG

Comments This chapter is written for the forthcoming book "Game Theory and Machine Learning for Cyber Security" (Wiley-IEEE Press), edited by Charles Kamhoua et. al. arXiv admin note: text overlap with arXiv:1906.10571

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.09540 2020-07-21 cs.AI 57%

Multi-Principal Assistance Games

Arnaud Fickinger, Simon Zhuang, Dylan Hadfield-Menell, Stuart Russell

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.06123 2020-07-14 cs.SD cs.LG eess.AS stat.ML 57%

OtoWorld: Towards Learning to Separate by Learning to Move

Omkar Ranadive, Grant Gasser, David Terpay, Prem Seetharaman

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

Comments Published in Self Supervision in Audio and Speech Workshop, 37th International Conference on Machine Learning, Vienna, Austria (ICML 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.00463 2020-07-02 cs.AI 57%

A Generalized Reinforcement Learning Algorithm for Online 3D Bin-Packing

Richa Verma, Aniruddha Singhal, Harshad Khadilkar, Ansuma Basumatary, Siddharth Nayak, Harsh Vardhan Singh, Swagat Kumar, Rajesh Sinha

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI

Comments 9 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.00425 2020-07-02 cs.LG stat.ML 57%

Interaction-limited Inverse Reinforcement Learning

Martin Troussard, Emmanuel Pignat, Parameswaran Kamalaruban, Sylvain Calinon, Volkan Cevher

专题命中 模仿学习与强化学习 :robot learning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.06800 2020-06-30 cs.LG stat.ML 57%

Context-aware Dynamics Model for Generalization in Model-Based Reinforcement Learning

Kimin Lee, Younggyo Seo, Seunghyun Lee, Honglak Lee, Jinwoo Shin

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

Comments Accepted in ICML2020. First two authors contributed equally, website: https://sites.google.com/view/cadm code: https://github.com/younggyoseo/CaDM

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.08383 2020-06-30 math.OC cs.LG cs.SY eess.SY math.ST stat.TH 57%

Global Convergence of Policy Gradient Methods to (Almost) Locally Optimal Policies

Kaiqing Zhang, Alec Koppel, Hao Zhu, Tamer Başar

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

Comments Initially submitted in Jan. 2019. Accepted to SIAM Journal on Control and Optimization (SICON)

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.03499 2020-06-29 cs.LG stat.ML 57%

Online Constrained Model-based Reinforcement Learning

Benjamin van Niekerk, Andreas Damianou, Benjamin Rosman

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.LG

Comments Conf. Uncertainty in Artificial Intelligence (UAI). 2017

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.10810 2020-06-22 cs.LG stat.ML 57%

Reparameterized Variational Divergence Minimization for Stable Imitation

Dilip Arumugam, Debadeepta Dey, Alekh Agarwal, Asli Celikyilmaz, Elnaz Nouri, Bill Dolan

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.07815 2020-06-16 cs.LG math.OC stat.ML 57%

Optimistic Distributionally Robust Policy Optimization

Jun Song, Chaoyue Zhao

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.07777 2020-06-16 cs.LG cs.HC stat.ML 57%

Active Imitation Learning from Multiple Non-Deterministic Teachers: Formulation, Challenges, and Algorithms

Khanh Nguyen, Hal Daumé

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.07041 2020-06-15 stat.ML cs.LG 57%

Mutual Information Based Knowledge Transfer Under State-Action Dimension Mismatch

Michael Wan, Tanmay Gangwani, Jian Peng

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.LG

Comments Conference on Uncertainty in Artificial Intelligence (UAI 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.09725 2020-06-11 cs.AI 57%

Conservative Agency via Attainable Utility Preservation

Alexander Matt Turner, Dylan Hadfield-Menell, Prasad Tadepalli

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.AI

Comments Published in AI, Ethics, and Society 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.10570 2020-06-08 cs.SI cs.LG 57%

Detecting Troll Behavior via Inverse Reinforcement Learning: A Case Study of Russian Trolls in the 2016 US Election

Luca Luceri, Silvia Giordano, Emilio Ferrara

专题命中 模仿学习与强化学习 :manipulation(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.08786 2020-05-26 cs.LG stat.ML 57%

Maximum Entropy-Regularized Multi-Goal Reinforcement Learning

Rui Zhao, Xudong Sun, Volker Tresp

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.LG

Comments Published in International Conference on Machine Learning (ICML 2019), Long Beach, USA. arXiv admin note: text overlap with arXiv:1902.08039

Journal ref PMLR 97:7553-7562, 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.00981 2020-05-19 cs.AI 57%

Benchmarking End-to-End Behavioural Cloning on Video Games

Anssi Kanervisto, Joonas Pussinen, Ville Hautamäki

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.AI

Comments To appear in IEEE Conference on Games 2020. Experiment code available at https://github.com/joonaspu/video-game-behavioural-cloning and https://github.com/joonaspu/ViControl

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.00935 2020-05-05 cs.LG cs.MA cs.SY eess.SP eess.SY stat.ML 57%

Deep Reinforcement Learning for Intelligent Transportation Systems: A Survey

Ammar Haydari, Yasin Yilmaz

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.12909 2020-05-01 cs.LG stat.ML 57%

Evolutionary Stochastic Policy Distillation

Hao Sun, Xinyu Pan, Bo Dai, Dahua Lin, Bolei Zhou

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.02380 2020-04-07 cs.LG stat.ML 57%

Intrinsic Exploration as Multi-Objective RL

Philippe Morere, Fabio Ramos

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.10026 2020-03-24 cs.NE cs.RO cs.SY eess.SY 57%

Learning to Walk: Spike Based Reinforcement Learning for Hexapod Robot Central Pattern Generation

Ashwin Sanjay Lele, Yan Fang, Justin Ting, Arijit Raychowdhury

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.RO

Comments 5 pages, 7 figures, to be published in proceeding of IEEE AICAS

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.05325 2020-03-12 cs.LG stat.ML 57%

Meta-learning curiosity algorithms

Ferran Alet, Martin F. Schneider, Tomas Lozano-Perez, Leslie Pack Kaelbling

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.LG

Comments Published in ICLR 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1712.06489 2020-03-10 cs.RO 57%

Multi-Fidelity Reinforcement Learning with Gaussian Processes

Varun Suryan, Nahush Gondhalekar, Pratap Tokekar

专题命中 模仿学习与强化学习 :navigation(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.00179 2020-03-04 cs.LG stat.ML 57%

TAdam: A Robust Stochastic Gradient Optimizer

Wendyam Eric Lionel Ilboudo, Taisuke Kobayashi, Kenji Sugimoto

专题命中 模仿学习与强化学习 :robotics(abstract);分类 cs.LG

Comments 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏