IN-RIL: Interleaved Reinforcement and Imitation Learning for Policy Fine-Tuning
Dechen Gao, Hang Wang, Hanchu Zhou, Nejib Ammar, Shatadal Mishra, Ahmadreza Moradipari, Iman Soltani, Junshan Zhang
机构
*
Department of Computer Science(计算机科学系)
;
University of California Davis(加州大学戴维斯分校)
;
Department of Electrical and Computer Engineering(电气与计算机工程系)
;
Toyota InfoTech Labs(丰田信息科技实验室)
;
Department of Mechanical and Aerospace Engineering(机械与航空航天工程系)
Improving Behavioural Cloning with Positive Unlabeled Learning
Qiang Wang, Robert McCarthy, David Cordova Bulens, Kevin McGuinness, Noel E. O'Connor, Nico Gürtler, Felix Widmaier, Francisco Roldan Sanchez, Stephen J. Redmond