arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 4753 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 记忆与上下文管理 4753 篇

2110.08704 2021-10-19 cs.IT cs.AI cs.LG math.IT 62%

A Q-Learning-based Approach for Distributed Beam Scheduling in mmWave Networks

Xiang Zhang, Shamik Sarkar, Arupjyoti Bhuyan, Sneha Kumar Kasera, Mingyue Ji

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.08888 2021-09-21 cs.LG cs.AI 62%

Multimodal Reward Shaping for Efficient Exploration in Reinforcement Learning

Mingqi Yuan, Mon-on Pun, Dong Wang, Yi Chen, Haojun Li

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 17 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.08236 2021-09-20 cs.LG cs.AI 62%

Reinforcement Learning on Encrypted Data

Alberto Jesu, Victor-Alexandru Darvariu, Alessandro Staffolani, Rebecca Montanari, Mirco Musolesi

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.05714 2021-09-06 cs.LG cs.AI 62%

Domain Adaptation In Reinforcement Learning Via Latent Unified State Representation

Jinwei Xing, Takashi Nagata, Kexin Chen, Xinyun Zou, Emre Neftci, Jeffrey L. Krichmar

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted by AAAI 2021; Fixed a typo in equation 3

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.08911 2021-08-23 cs.LG cs.AI 62%

Explainable Deep Reinforcement Learning Using Introspection in a Non-episodic Task

Angel Ayala, Francisco Cruz, Bruno Fernandes, Richard Dazeley

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.04444 2021-08-23 cs.LG cs.AI 62%

Jointly-Learned State-Action Embedding for Efficient Reinforcement Learning

Paul J. Pritz, Liang Ma, Kin K. Leung

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.01869 2021-08-05 cs.AI cs.LG 62%

Learning Task Agnostic Skills with Data-driven Guidance

Even Klemsdal, Sverre Herland, Abdulmajid Murad

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments ICML 2021 Workshop on Unsupervised Reinforcement Learning (https://openreview.net/forum?id=CPh9DeHv08U)

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.09394 2021-08-03 cs.DB cs.AI cs.LG 62%

Knowledge Graph-based Question Answering with Electronic Health Records

Junwoo Park, Youngwoo Cho, Haneol Lee, Jaegul Choo, Edward Choi

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted at Machine Learning in Health Care (MLHC) 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.03003 2021-07-08 cs.LG cs.AI stat.ML 62%

Harnessing Heterogeneity: Learning from Decomposed Feedback in Bayesian Modeling

Kai Wang, Bryan Wilder, Sze-chuan Suen, Bistra Dilkina, Milind Tambe

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.15714 2021-07-06 cs.LG cs.AI 62%

Active Finite Reward Automaton Inference and Reinforcement Learning Using Queries and Counterexamples

Zhe Xu, Bo Wu, Aditya Ojha, Daniel Neider, Ufuk Topcu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.00703 2021-07-05 cs.LG cs.AI 62%

Distilling Reinforcement Learning Tricks for Video Games

Anssi Kanervisto, Christian Scheller, Yanick Schraner, Ville Hautamäki

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments To appear in IEEE Conference on Games 2021. Experiment code is available at https://github.com/Miffyli/rl-human-prior-tricks

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.09578 2021-06-18 cs.CL cs.AI 62%

Modeling Worlds in Text

Prithviraj Ammanabrolu, Mark O. Riedl

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.CL

Comments Preprint. Under review. Benchmark can be found at https://github.com/JerichoWorld/JerichoWorld

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.02383 2021-06-18 cs.LG cs.AI stat.ML 62%

Randomized Value Functions via Posterior State-Abstraction Sampling

Dilip Arumugam, Benjamin Van Roy

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted to the Workshop on Biological and Artificial Reinforcement Learning (NeurIPS 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.07814 2021-06-16 cs.LG cs.AI stat.ML 62%

Sample Efficient Reinforcement Learning In Continuous State Spaces: A Perspective Beyond Linearity

Dhruv Malik, Aldo Pacchiano, Vishwak Srinivasan, Yuanzhi Li

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI、cs.LG

Comments ICML 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.06469 2021-06-14 cs.LG cs.AI 62%

Generalizable Episodic Memory for Deep Reinforcement Learning

Hao Hu, Jianing Ye, Guangxiang Zhu, Zhizhou Ren, Chongjie Zhang

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.01062 2021-06-11 cs.LG cs.AI stat.ML 62%

Exploration in Approximate Hyper-State Space for Meta Reinforcement Learning

Luisa Zintgraf, Leo Feng, Cong Lu, Maximilian Igl, Kristian Hartikainen, Katja Hofmann, Shimon Whiteson

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Published at the International Conference on Machine Learning (ICML) 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.03775 2021-06-08 cs.AI cs.LG 62%

Explainable Artificial Intelligence (XAI) for Increasing User Trust in Deep Reinforcement Learning Driven Autonomous Systems

Jeff Druce, Michael Harradon, James Tittle

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments NeurIPS Deep RL workshop, Conference on Neural Information Processing Systems (NeurIPS 2019), Vancouver, Canada

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.00123 2021-06-04 cs.LG cs.CL stat.ML 62%

Unsupervised Learning of KB Queries in Task-Oriented Dialogs

Dinesh Raghu, Nikhil Gupta, Mausam

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL、cs.LG

Comments Presented at ACL 2021

Journal ref Transactions of the Association for Computational Linguistics (2021) 9: 374-390

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.03752 2021-05-11 physics.flu-dyn cs.AI cs.LG physics.comp-ph 62%

Improving Deep Learning Performance for Predicting Large-Scale Porous-Media Flow through Feature Coarsening

Bicheng Yan, Dylan Robert Harp, Bailian Chen, Rajesh J. Pawar

专题命中 记忆与上下文管理 :workflow(abstract);分类 cs.AI、cs.LG

Comments 12 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2105.00499 2021-05-04 cs.RO cs.AI cs.LG cs.MA 62%

Curious Exploration and Return-based Memory Restoration for Deep Reinforcement Learning

Saeed Tafazzol, Erfan Fathi, Mahdi Rezaei, Ehsan Asali

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.11634 2021-05-03 cs.AI cs.LG cs.NE cs.SY eess.SY 62%

Assessment of Reward Functions for Reinforcement Learning Traffic Signal Control under Real-World Limitations

Alvaro Cabrejas-Egea, Shaun Howell, Maksis Knutins, Colm Connaughton

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Conference paper, 13 pages, 7 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.11917 2021-04-27 cs.RO cs.AI cs.LG 62%

Batch Exploration with Examples for Scalable Robotic Reinforcement Learning

Annie S. Chen, HyunJi Nam, Suraj Nair, Chelsea Finn

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 11 Pages, 11 Figures

Journal ref IEEE Robotics and Automation Letters ( Volume: 6, Issue: 3, July 2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.12088 2021-04-15 cs.AI cs.LG 62%

Fast Approximate Solutions using Reinforcement Learning for Dynamic Capacitated Vehicle Routing with Time Windows

Nazneen N Sultana, Vinita Baniwal, Ansuma Basumatary, Piyush Mittal, Supratim Ghosh, Harshad Khadilkar

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.02959 2021-04-08 cs.LG cs.AI 62%

The Emergence of Abstract and Episodic Neurons in Episodic Meta-RL

Badr AlKhamissi, Muhammad ElNokrashy, Michael Spranger

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments This work was accepted at the Learning to Learn Workshop (ICLR 2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2012.10140 2021-04-02 cs.LG cs.AI cs.RO cs.SY eess.SY 62%

Voronoi Progressive Widening: Efficient Online Solvers for Continuous State, Action, and Observation POMDPs

Michael H. Lim, Claire J. Tomlin, Zachary N. Sunberg

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.15073 2021-03-30 cs.LG cs.AI cs.SY eess.SY 62%

IUP: An Intelligent Utility Prediction Scheme for Solid-State Fermentation in 5G IoT

Min Wang, Shanchen Pang, Tong Ding, Sibo Qiao, Xue Zhai, Shuo Wang, Neal N. Xiong, Zhengwen Huang

专题命中 记忆与上下文管理 :workflow(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.03855 2021-03-18 cs.AI cs.LG 62%

Induction and Exploitation of Subgoal Automata for Reinforcement Learning

Daniel Furelos-Blanco, Mark Law, Anders Jonsson, Krysia Broda, Alessandra Russo

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Published in the Journal of Artificial Intelligence Research (JAIR)

Journal ref Journal of Artificial Intelligence Research, 70, 1031-1116 (2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.01883 2021-03-16 cs.LG cs.AI stat.ML 62%

Deep Radial-Basis Value Functions for Continuous Control

Kavosh Asadi, Neev Parikh, Ronald E. Parr, George D. Konidaris, Michael L. Littman

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments In Proceedings of the 35th AAAI Conference on Artificial Intelligence (AAAI)

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.09324 2021-03-09 cs.LG cs.AI stat.ML 62%

The Sample Complexity of Teaching-by-Reinforcement on Q-Learning

Xuezhou Zhang, Shubham Kumar Bharti, Yuzhe Ma, Adish Singla, Xiaojin Zhu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2102.11051 2021-02-23 cs.RO cs.AI cs.LG 62%

Improved Learning of Robot Manipulation Tasks via Tactile Intrinsic Motivation

Nikola Vulin, Sammy Christen, Stefan Stevsic, Otmar Hilliges

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏