arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 4753 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 记忆与上下文管理 4753 篇

2004.12846 2020-04-29 cs.NE cs.AI cs.LG 62%

Evolving Inborn Knowledge For Fast Adaptation in Dynamic POMDP Problems

Eseoghene Ben-Iwhiwhu, Pawel Ladosz, Jeffery Dick, Wen-Hua Chen, Praveen Pilly, Andrea Soltoggio

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 9 pages. Accepted as a full paper in the Genetic and Evolutionary Computation Conference (GECCO 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.05367 2020-04-29 cs.CL cs.LG stat.ML 62%

Learning in Text Streams: Discovery and Disambiguation of Entity and Relation Instances

Marco Maggini, Giuseppe Marra, Stefano Melacci, Andrea Zugarini

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.08648 2020-04-21 cs.LG cs.AI cs.RO stat.ML 62%

Modeling Survival in model-based Reinforcement Learning

Saeed Moazami, Peggy Doerschuk

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.03168 2020-04-08 cs.LG cs.AI stat.ML 62%

Trying AGAIN instead of Trying Longer: Prior Learning for Automatic Curriculum Learning

Rémy Portelas, Katja Hofmann, Pierre-Yves Oudeyer

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted to the ICLR 2020 workshop Beyond tabula rasa in RL (BeTR-RL)

详情

展开后加载摘要…

URL PDF HTML 收藏
2004.02919 2020-04-08 cs.LG cs.AI stat.ML 62%

Uniform State Abstraction For Reinforcement Learning

John Burden, Daniel Kudenko

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 8 Pages, 2 figures, Accepted for publication in the European Conference of Artificial Intelligence (ECAI 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.03456 2020-03-26 cs.LG cs.AI stat.ML 62%

A Farewell to Arms: Sequential Reward Maximization on a Budget with a Giving Up Option

P Sharoff, Nishant A. Mehta, Ravi Ganti

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 16 pages, AISTATS 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.10084 2020-03-10 cs.AI cs.LG 62%

SensAI+Expanse Adaptation on Human Behaviour Towards Emotional Valence Prediction

Nuno A. C. Henriques, Helder Coelho, Leonel Garcia-Marques

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted as regular paper in ADAPTIVE 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.12292 2020-03-03 cs.LG cs.AI 62%

RIDE: Rewarding Impact-Driven Exploration for Procedurally-Generated Environments

Roberta Raileanu, Tim Rocktäschel

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.05274 2020-03-02 cs.LG cs.AI cs.RO stat.ML 62%

Efficient Exploration via State Marginal Matching

Lisa Lee, Benjamin Eysenbach, Emilio Parisotto, Eric Xing, Sergey Levine, Ruslan Salakhutdinov

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Videos and code: https://sites.google.com/view/state-marginal-matching

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.12174 2020-02-28 cs.LG cs.AI stat.ML 62%

Optimistic Exploration even with a Pessimistic Initialisation

Tabish Rashid, Bei Peng, Wendelin Böhmer, Shimon Whiteson

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Published as a conference paper at ICLR 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.13406 2020-02-20 cs.LG cs.AI stat.ML 62%

Generalization of Reinforcement Learners with Working and Episodic Memory

Meire Fortunato, Melissa Tan, Ryan Faulkner, Steven Hansen, Adrià Puigdomènech Badia, Gavin Buttimore, Charlie Deck, Joel Z Leibo, Charles Blundell

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments NeurIPS 2019. Equal contribution of first 4 authors

Journal ref 33rd Conference on Neural Information Processing Systems (Neurips 2019)

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.06238 2020-02-18 cs.LG cs.AI stat.ML 62%

On State Variables, Bandit Problems and POMDPs

Warren B Powell

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.11274 2020-02-18 cs.LG cs.AI cs.HC cs.IR 62%

Scalable Psychological Momentum Forecasting in Esports

Alfonso White, Daniela M. Romano

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 8 pages, 8 figures

Journal ref Proceedings of Workshop SUM '20: State-based User Modelling, The 13th ACM International Conference on Web Search and Data Mining (WSDM '20), 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.06137 2020-02-17 cs.AI cs.LG 62%

RL agents Implicitly Learning Human Preferences

Nevan Wichers

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.08904 2020-01-27 cs.CL cs.LG stat.ML 62%

MT-BioNER: Multi-task Learning for Biomedical Named Entity Recognition using Deep Bidirectional Transformers

Muhammad Raza Khan, Morteza Ziyadi, Mohamed AbdelHady

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.12394 2020-01-17 cs.CL cs.CV cs.LG 62%

All-in-One Image-Grounded Conversational Agents

Da Ju, Kurt Shuster, Y-Lan Boureau, Jason Weston

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.08578 2019-12-20 cs.LG cs.AI cs.RO 62%

Taming an autonomous surface vehicle for path following and collision avoidance using deep reinforcement learning

Eivind Meyer, Haakon Robinson, Adil Rasheed, Omer San

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 16 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.00617 2019-12-03 cs.LG cs.AI stat.ML 62%

Explicit Explore-Exploit Algorithms in Continuous State Spaces

Mikael Henaff

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.13152 2019-12-02 cs.LG cs.AI cs.LO stat.ML 62%

Induction of Subgoal Automata for Reinforcement Learning

Daniel Furelos-Blanco, Mark Law, Alessandra Russo, Krysia Broda, Anders Jonsson

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Preprint accepted for publication to the 34th AAAI Conference on Artificial Intelligence (AAAI-20)

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.12511 2019-12-02 cs.AI cs.LG 62%

Algorithmic Improvements for Deep Reinforcement Learning applied to Interactive Fiction

Vishal Jain, William Fedus, Hugo Larochelle, Doina Precup, Marc G. Bellemare

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments To appear in Proceedings of the Thirty-Fourth AAAI Conference on Artificial Intelligence (AAAI-20). Accepted for Oral presentation

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.10304 2019-12-02 cs.CV cs.AI cs.LG cs.RO 62%

Where to Look Next: Unsupervised Active Visual Exploration on 360° Input

Soroush Seifi, Tinne Tuytelaars

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Oral Presentation and best Paper Award at 360 Perception and Interaction Workshop at ICCV 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.02022 2019-11-27 cs.CV cs.AI cs.CL cs.RO 62%

Chasing Ghosts: Instruction Following as Bayesian State Tracking

Peter Anderson, Ayush Shrivastava, Devi Parikh, Dhruv Batra, Stefan Lee

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.CL

Comments NeurIPS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.02583 2019-11-20 cs.LG cs.AI cs.CR stat.ML 62%

Spatiotemporally Constrained Action Space Attacks on Deep Reinforcement Learning Agents

Xian Yeow Lee, Sambit Ghadai, Kai Liang Tan, Chinmay Hegde, Soumik Sarkar

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Version 2 with supplementary materials

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.02047 2019-11-18 cs.AI cs.LG 62%

Age of Information-Aware Radio Resource Management in Vehicular Networks: A Proactive Deep Reinforcement Learning Perspective

Xianfu Chen, Celimuge Wu, Tao Chen, Honggang Zhang, Zhi Liu, Yan Zhang, Mehdi Bennis

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.14361 2019-11-01 cs.LG cs.AI stat.ML 62%

Object-oriented state editing for HRL

Victor Bapst, Alvaro Sanchez-Gonzalez, Omar Shams, Kimberly Stachenfeld, Peter W. Battaglia, Satinder Singh, Jessica B. Hamrick

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 8 pages; accepted to the Perception as Generative Reasoning workshop of the 33rd Conference on Neural InformationProcessing Systems (NeurIPS 2019)

详情

展开后加载摘要…

URL PDF HTML 收藏
1711.05726 2019-10-24 stat.ML cs.AI cs.LG 62%

Markov Decision Processes with Continuous Side Information

Aditya Modi, Nan Jiang, Satinder Singh, Ambuj Tewari

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Journal ref PMLR Volume 83: Algorithmic Learning Theory, 7-9 April 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.06424 2019-10-23 cs.LG cs.AI stat.ML 62%

Meta reinforcement learning as task inference

Jan Humplik, Alexandre Galashov, Leonard Hasenclever, Pedro A. Ortega, Yee Whye Teh, Nicolas Heess

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.06315 2019-10-15 cs.CV cs.CL cs.LG 62%

Dynamic Attention Networks for Task Oriented Grounding

Soumik Dasgupta, Badri N. Patro, Vinay P. Namboodiri

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL、cs.LG

Comments Accepted ICCV 2019 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.05389 2019-10-15 cs.CL cs.AI 62%

Model-based Interactive Semantic Parsing: A Unified Framework and A Text-to-SQL Case Study

Ziyu Yao, Yu Su, Huan Sun, Wen-tau Yih

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.CL

Comments 14 pages, 4 figures, accepted to EMNLP 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.13357 2019-10-10 cs.LG cs.AI stat.ML 62%

Reinforcement Learning for Mean Field Game

Mridul Agarwal, Vaneet Aggarwal, Arnob Ghosh, Nilay Tiwari

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏