arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

世界模型

面向环境建模、时序预测、仿真规划、具身智能和自动驾驶的世界模型方法与应用。

共收录 1124 信号源:cs.AI, cs.LG, cs.CV, cs.RO, cs.MA

1. 模型式强化学习 1124 篇

2106.05969 2022-11-01 cs.CV cs.AI cs.LG cs.RO 56%

Dynamics-Regulated Kinematic Policy for Egocentric Pose Estimation

Zhengyi Luo, Ryo Hachiuma, Ye Yuan, Kris Kitani

专题命中 模型式强化学习 :分类 cs.AI、cs.LG、cs.CV;dynamics model(abstract)

Comments NeurIPS 2021. Project page: https://zhengyiluo.github.io/projects/kin_poly/

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.05922 2022-10-13 cs.LG cs.AI 56%

A Unified Framework for Alternating Offline Model Training and Policy Learning

Shentao Yang, Shujian Zhang, Yihao Feng, Mingyuan Zhou

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

Comments 36th Conference on Neural Information Processing Systems (NeurIPS 2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.01805 2022-10-06 cs.LG cs.AI 56%

CostNet: An End-to-End Framework for Goal-Directed Reinforcement Learning

Per-Arne Andersen, Morten Goodwin, Ole-Christoffer Granmo

专题命中 模型式强化学习 :model-based RL(abstract);分类 cs.AI、cs.LG

Comments 14 pages, 5 figures, In Proceedings of the International Conference on Innovative Techniques and Applications of Artificial Intelligence, SGAI2020

Journal ref 2020 Springer Nature Switzerland AG

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.00368 2022-10-04 cs.LG 56%

Parameter-varying neural ordinary differential equations with partition-of-unity networks

Kookjin Lee, Nathaniel Trask

专题命中 模型式强化学习 :latent dynamics(abstract);分类 cs.LG;dynamics model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.12387 2022-09-27 cs.LG cs.AI stat.ML 56%

Neural State-Space Modeling with Latent Causal-Effect Disentanglement

Maryam Toloubidokhti, Ryan Missel, Xiajun Jiang, Niels Otani, Linwei Wang

专题命中 模型式强化学习 :latent dynamics(abstract);分类 cs.AI、cs.LG

Comments presented at the 13th Machine Learning in Medical Imaging (MLMI 2022) Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.07123 2022-08-16 cs.RO cs.AI 56%

Online 3D Bin Packing Reinforcement Learning Solution with Buffer

Aaron Valero Puche, Sukhan Lee

专题命中 模型式强化学习 :model-based RL(abstract);分类 cs.AI、cs.RO

Comments Accepted in IROS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.14003 2022-07-29 cs.CL cs.AI cs.CY cs.HC cs.LG 56%

Raising Student Completion Rates with Adaptive Curriculum and Contextual Bandits

Robert Belfer, Ekaterina Kochmar, Iulian Vlad Serban

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

Comments 6 pages, 1 figure, To appear in the Proceedings of the 23rd International Conference on Artificial Intelligence in Education (AIED 2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.06543 2022-07-26 cs.AI cs.LG 56%

Policy Optimization in Dynamic Bayesian Network Hybrid Models of Biomanufacturing Processes

Hua Zheng, Wei Xie, Ilya O. Ryzhov, Dongming Xie

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

Comments 36 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.04424 2022-07-13 cs.RO cs.LG 56%

Accelerated Reinforcement Learning for Temporal Logic Control Objectives

Yiannis Kantaros

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.LG、cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.04153 2022-07-01 cs.LG cs.AI 56%

Model-Value Inconsistency as a Signal for Epistemic Uncertainty

Angelos Filos, Eszter Vértes, Zita Marinho, Gregory Farquhar, Diana Borsa, Abram Friesen, Feryal Behbahani, Tom Schaul, André Barreto, Simon Osindero

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

Comments The first three authors contributed equally. Accepted at ICML 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.13452 2022-06-28 cs.LG 56%

Causal Dynamics Learning for Task-Independent State Abstraction

Zizhao Wang, Xuesu Xiao, Zifan Xu, Yuke Zhu, Peter Stone

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.LG;dynamics model(abstract)

Comments International Conference on Machine Learning 2022 (ICML)

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.00602 2022-06-20 stat.ML cs.AI cs.LG 56%

Meta-Learning Hypothesis Spaces for Sequential Decision-making

Parnian Kassraie, Jonas Rothfuss, Andreas Krause

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

Comments 23 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.10736 2022-05-24 cs.LG cs.AI stat.ML 56%

Should Models Be Accurate?

Esra'a Saleh, John D. Martin, Anna Koop, Arash Pourzarabi, Michael Bowling

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

Comments The 5th Multidisciplinary Conference on Reinforcement Learning and Decision Making ( RLDM 2022 )

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.11236 2022-03-29 cs.LG cs.AI q-bio.NC 56%

Variational Predictive Routing with Nested Subjective Timescales

Alexey Zakharov, Qinghai Guo, Zafeirios Fountas

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

Comments 18 pages, 13 figures

Journal ref Tenth International Conference on Learning Representations (ICLR 2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2201.01353 2022-03-22 cs.LG cs.CV cs.SY eess.SY 56%

Linear Variational State-Space Filtering

Daniel Pfrommer, Nikolai Matni

专题命中 模型式强化学习 :latent dynamics(abstract);分类 cs.LG、cs.CV

Comments 18 pages, 6 figures. Fixed proof in appendix. For associated code, see https://github.com/pfrommerd/variational_state_space_models

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.05079 2022-03-11 cs.LG cs.AI 56%

SAGE: Generating Symbolic Goals for Myopic Models in Deep Reinforcement Learning

Andrew Chester, Michael Dann, Fabio Zambetta, John Thangarajah

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

Comments 11 pages, 8 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.01810 2022-03-04 cs.LG cs.AI 56%

Integrating Contrastive Learning with Dynamic Models for Reinforcement Learning from Images

Bang You, Oleg Arenz, Youping Chen, Jan Peters

专题命中 模型式强化学习 :latent dynamics(abstract);分类 cs.AI、cs.LG

Comments 28 pages, 11 figures, 5 tables

Journal ref Neurocomputing 476(2022)102-114

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.06036 2022-02-15 cs.LG cs.AI 56%

Neural NID Rules

Luca Viano, Johanni Brea

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

Comments Physical Reasoning and Inductive Biases for the Real World at NeurIPS 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.11735 2022-01-11 cs.LG cs.AI cs.LO 56%

Lifted Model Checking for Relational MDPs

Wen-Chi Yang, Jean-François Raskin, Luc De Raedt

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.10504 2021-12-21 cs.LG cs.AI 56%

Sample-Efficient Reinforcement Learning via Conservative Model-Based Actor-Critic

Zhihai Wang, Jie Wang, Qi Zhou, Bin Li, Houqiang Li

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

Comments Accepted to AAAI22

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.10316 2021-12-14 cs.AI cs.LG 56%

Proper Value Equivalence

Christopher Grimm, André Barreto, Gregory Farquhar, David Silver, Satinder Singh

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

Journal ref NeurIPS 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.08253 2021-11-30 cs.LG cs.AI stat.ML 56%

When to Trust Your Model: Model-Based Policy Optimization

Michael Janner, Justin Fu, Marvin Zhang, Sergey Levine

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

Comments NeurIPS 2019. Code at https://github.com/JannerM/mbpo, project page at: https://jannerm.github.io/mbpo-www/

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.00715 2021-11-03 cs.LG cs.AI stat.ML 56%

Efficient Reinforcement Learning for StarCraft by Abstract Forward Models and Transfer Learning

Ruo-Ze Liu, Haifeng Guo, Xiaozhong Ji, Yang Yu, Zhen-Jia Pang, Zitai Xiao, Yuzhou Wu, Tong Lu

专题命中 模型式强化学习 :model-based RL(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.15489 2021-11-01 cs.LG cs.AI 56%

GalilAI: Out-of-Task Distribution Detection using Causal Active Experimentation for Safe Transfer RL

Sumedh A Sontakke, Stephen Iota, Zizhao Hu, Arash Mehrjou, Laurent Itti, Bernhard Schölkopf

专题命中 模型式强化学习 :model-based RL(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.14524 2021-10-28 cs.LG cs.MA 56%

Model based Multi-agent Reinforcement Learning with Tensor Decompositions

Pascal Van Der Vaart, Anuj Mahajan, Shimon Whiteson

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.LG、cs.MA

Journal ref 2nd Workshop on Quantum Tensor Networks in Machine Learning (NeurIPS 2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.12840 2021-10-26 cs.LG cs.AI stat.ML 56%

Self-Consistent Models and Values

Gregory Farquhar, Kate Baumli, Zita Marinho, Angelos Filos, Matteo Hessel, Hado van Hasselt, David Silver

专题命中 模型式强化学习 :model-based RL(abstract);分类 cs.AI、cs.LG

Comments NeurIPS 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.07910 2021-10-18 cs.LG cs.AI 56%

SaLinA: Sequential Learning of Agents

Ludovic Denoyer, Alfredo de la Fuente, Song Duong, Jean-Baptiste Gaya, Pierre-Alexandre Kamienny, Daniel H. Thompson

专题命中 模型式强化学习 :model-based RL(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.10242 2021-09-23 cs.CL cs.AI cs.LG 56%

Towards Automatic Evaluation of Dialog Systems: A Model-Free Off-Policy Evaluation Approach

Haoming Jiang, Bo Dai, Mengjiao Yang, Tuo Zhao, Wei Wei

专题命中 模型式强化学习 :model-based reinforcement learning(abstract);分类 cs.AI、cs.LG

Comments Conference on Empirical Methods in Natural Language Processing (EMNLP), 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.03214 2021-09-08 cs.LG cs.AI 56%

Robust Predictable Control

Benjamin Eysenbach, Ruslan Salakhutdinov, Sergey Levine

专题命中 模型式强化学习 :model-based RL(abstract);分类 cs.AI、cs.LG

Comments Project site with videos and code: https://ben-eysenbach.github.io/rpc

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.01852 2021-08-27 cs.LG cs.RO stat.ML 56%

q-VAE for Disentangled Representation Learning and Latent Dynamical Systems

Taisuke Kobayashi

专题命中 模型式强化学习 :latent dynamics(abstract);分类 cs.LG、cs.RO

Comments 8 pages, 8 figures

Journal ref IEEE Robotics and Automation Letters, 2020

详情

展开后加载摘要…

URL PDF HTML 收藏