机构
*
Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学北京校区人工智能学院)
;
Beijing Key Laboratory of Research on Large Models and Intelligent Governance(北京大型模型与智能治理研究重点实验室)
;
Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(教育部下一代智能搜索与推荐工程研究中心)
;
Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)
;
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)
Rewarding DINO: Predicting Dense Rewards with Vision Foundation Models
奖励DINO:基于视觉基础模型预测密集奖励
Pierre Krack, Tobias Jülg, Wolfram Burgard, Florian Walter
机构
*
Department of Computer Science & Artificial Intelligence, University of Technology Nuremberg(技术大学纽伦堡计算机科学与人工智能系)
;
TUM School of Computation, Information and Technology, Technical University of Munich(慕尼黑技术大学计算、信息与技术学院)
ViSA: Visited-State Augmentation for Generalized Goal-Space Contrastive Reinforcement Learning
ViSA:基于已访问状态的通用目标空间对比强化学习增强
Issa Nakamura, Tomoya Yamanokuchi, Yuki Kadokawa, Jia Qu, Shun Otsub, Ken Miyamoto, Shotaro Miwa, Takamitsu Matsubara
机构
*
Graduate School of Information Science, Nara Institute of Science and Technology (NAIST)(信息科学研究生院,奈良科学与技术研究所)
;
Advanced Technology R&D Center, Mitsubishi Electric Corporation(三菱电机先进技术研发中心)