arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 4758 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 记忆与上下文管理 4758 篇

2010.14680 2021-06-22 cs.LG stat.ML 57%

Learning to Represent Action Values as a Hypergraph on the Action Vertices

Arash Tavakoli, Mehdi Fatemi, Petar Kormushev

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments ICLR 2021, code: https://github.com/atavakol/action-hypergraph-networks

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.08832 2021-06-17 cs.LG 57%

Solving Continuous Control with Episodic Memory

Igor Kuznetsov, Andrey Filchenkov

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments To appear in the 30th International Joint Conference on Artificial Intelligence (IJCAI 2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.08477 2021-06-17 cs.LG stat.ML 57%

Fundamental Limits of Reinforcement Learning in Environment with Endogeneous and Exogeneous Uncertainty

Rongpeng Li

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Manuscript has been submitted to an IEEE journal. Copyright may be transferred without further notice

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.06868 2021-06-15 cs.LG stat.AP 57%

Short-term forecasting of global solar irradiance with incomplete data

Laura S. Hoyos-Gómez, Jose F. Ruiz-Muñoz, Belizza J. Ruiz-Mendoza

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.13658 2021-06-15 cs.LG 57%

Locally Persistent Exploration in Continuous Control Tasks with Sparse Rewards

Susan Amin, Maziar Gomrokchi, Hossein Aboutalebi, Harsh Satija, Doina Precup

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments To be published in ICML, 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.03050 2021-06-08 cs.LG 57%

Efficient Continuous Control with Double Actors and Regularized Critics

Jiafei Lyu, Xiaoteng Ma, Jiangpeng Yan, Xiu Li

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 21 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.02396 2021-06-07 eess.SY cs.LG cs.SY 57%

A Learning-based Optimal Market Bidding Strategy for Price-Maker Energy Storage

Mathilde D. Badoual, Scott J. Moura

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Presented at the 2021 American Control Conference (ACC), New Orleans, USA, May 25-28, 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.02204 2021-06-07 cs.AI 57%

Detecting and Adapting to Novelty in Games

Xiangyu Peng, Jonathan C. Balloch, Mark O. Riedl

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 10 pages, 5 figures, Accepted to the AAAI21 Workshop on on Reinforcement Learning in Games

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.13565 2021-06-07 cs.LG 57%

Low-Precision Reinforcement Learning: Running Soft Actor-Critic in Half Precision

Johan Bjorck, Xiangyu Chen, Christopher De Sa, Carla P. Gomes, Kilian Q. Weinberger

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.09992 2021-06-01 cs.LG 57%

Don't Do What Doesn't Matter: Intrinsic Motivation with Action Usefulness

Mathieu Seurin, Florian Strub, Philippe Preux, Olivier Pietquin

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Accepted at Internationnal Joint Conference on Artificial Intelligence (IJCAI'21) and Self-Supervision for Reinforcement Learning Workshop (SSL-RL @ICLR'21)

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.13912 2021-06-01 math.OC cs.LG cs.MA 57%

Unified Reinforcement Q-Learning for Mean Field Game and Control Problems

Andrea Angiuli, Jean-Pierre Fouque, Mathieu Laurière

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.10577 2021-05-25 cs.AI cs.NE 57%

Modelling the development of counting with memory-augmented neural networks

Zack Dulberg, Taylor Webb, Jonathan Cohen

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Accepted talk at Proceedings of the 42nd Annual Meeting of the Cognitive Science Society

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.07284 2021-05-24 q-bio.NC cs.AI 57%

A brain basis of dynamical intelligence for AI and computational neuroscience

Joseph D. Monaco, Kanaka Rajan, Grace M. Hwang

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Perspective article: 24 pages, 3 figures, 1 display box

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.09059 2021-05-20 cs.CY cs.AI 57%

The State of AI Ethics Report (January 2021)

Abhishek Gupta, Alexandrine Royer, Connor Wright, Falaah Arif Khan, Victoria Heath, Erick Galinkin, Ryan Khurana, Marianna Bergamaschi Ganapini, Muriam Fancy, Masa Sweidan, Mo Akif, Renjie Butalid

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI

Comments 188 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2101.09536 2021-05-07 cs.CV cs.LG 57%

Memory-Efficient Semi-Supervised Continual Learning: The World is its Own Replay Buffer

James Smith, Jonathan Balloch, Yen-Chang Hsu, Zsolt Kira

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Accepted by the 2021 International Joint Conference on Neural Networks (IJCNN 2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.04546 2021-04-26 cs.LG cs.CV stat.ML 57%

Wandering Within a World: Online Contextualized Few-Shot Learning

Mengye Ren, Michael L. Iuzzolino, Michael C. Mozer, Richard S. Zemel

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments ICLR 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.08060 2021-04-19 cs.LG 57%

MEG: Generating Molecular Counterfactual Explanations for Deep Graph Networks

Danilo Numeroso, Davide Bacciu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 8 pages, 5 figures, to appear in the Proceedings of the 2021 International Joint Conference on Neural Networks (IJCNN 2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.02899 2021-04-08 cs.LG 57%

Recognizing and Verifying Mathematical Equations using Multiplicative Differential Neural Units

Ankur Mali, Alexander Ororbia, Daniel Kifer, C. Lee Giles

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.00249 2021-04-02 cs.CV cs.LG 57%

LaPred: Lane-Aware Prediction of Multi-Modal Future Trajectories of Dynamic Agents

ByeoungDo Kim, Seong Hyeon Park, Seokhwan Lee, Elbek Khoshimjonov, Dongsuk Kum, Junsoo Kim, Jeong Soo Kim, Jun Won Choi

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 13 pages, 2 figures, 7 tables, CVPR 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.14251 2021-03-29 eess.SY cs.LG cs.SY physics.soc-ph 57%

Embedding Power Flow into Machine Learning for Parameter and State Estimation

Laurent Pagnier, Michael Chertkov

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

Comments 7 pages, 3 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.01834 2021-03-29 cs.SE 57%

Reinforcement Learning for Test Case Prioritization

Mojtaba Bagherzadeh, Nafiseh Kahani, Lionel Briand

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.SE

Journal ref IEEE Transactions on Software Engineering (TSE). (2021) 1-21

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.04678 2021-03-18 cs.LG stat.ML 57%

Primal Wasserstein Imitation Learning

Robert Dadashi, Léonard Hussenot, Matthieu Geist, Olivier Pietquin

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Published in International Conference on Learning Representations (ICLR 2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.10178 2021-03-16 stat.ML cs.CV cs.LG 57%

Variational State-Space Models for Localisation and Dense 3D Mapping in 6 DoF

Atanas Mirchev, Baris Kayalibay, Patrick van der Smagt, Justin Bayer

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

Comments Update for ICLR2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.08107 2021-03-16 cs.LG 57%

Mutual Information State Intrinsic Control

Rui Zhao, Yang Gao, Pieter Abbeel, Volker Tresp, Wei Xu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Published in International Conference on Learning Representations (ICLR 2021) as Spotlight (top 5%), Link: https://openreview.net/forum?id=OthEq8I5v1. arXiv admin note: text overlap with arXiv:2002.01963

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.08079 2021-03-16 cs.HC cs.AI cs.RO 57%

Crossing the Tepper Line: An Emerging Ontology for Describing the Dynamic Sociality of Embodied AI

Katie Seaborn, Peter Pennefather, Norihisa P. Miyake, Mihoko Otake-Matsuura

专题命中 记忆与上下文管理 :agentic(abstract);分类 cs.AI

Comments Accepted at CHI EA '21

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.06371 2021-03-12 cs.AI 57%

Hard Attention Control By Mutual Information Maximization

Himanshu Sahni, Charles Isbell

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.13319 2021-03-11 cs.LG stat.ML 57%

Efficient Reinforcement Learning in Factored MDPs with Application to Constrained RL

Xiaoyu Chen, Jiachen Hu, Lihong Li, Liwei Wang

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.04152 2021-03-09 eess.SY cs.LG cs.SY 57%

Correlated Deep Q-learning based Microgrid Energy Management

Hao Zhou, Melike Erol-Kantarci

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Accepted by 2020 IEEE 25th International Workshop on CAMAD, 978-1-7281-6339-0/20/$31.00 ©2020 IEEE

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.06036 2021-03-08 cs.LG stat.ML 57%

Reinforcement Learning with Trajectory Feedback

Yonathan Efroni, Nadav Merlis, Shie Mannor

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments AAAI2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.07717 2021-03-02 stat.ML cs.LG 57%

Reinforcement Learning for Molecular Design Guided by Quantum Mechanics

Gregor N. C. Simm, Robert Pinsler, José Miguel Hernández-Lobato

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Journal ref Proceedings of the 37th International Conference on Machine Learning, Vienna, Austria, PMLR 119, 2020

详情

展开后加载摘要…

URL PDF HTML 收藏