arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 4753 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 记忆与上下文管理 4753 篇

2212.00187 2022-12-02 cs.AI cs.LG 62%

Five Properties of Specific Curiosity You Didn't Know Curious Machines Should Have

Nadia M. Ady, Roshan Shariff, Johannes Günther, Patrick M. Pilarski

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Submitted to the Journal of Artificial Intelligence Research (JAIR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.16002 2022-11-30 cs.CL cs.AI 62%

DiffG-RL: Leveraging Difference between State and Common Sense

Tsunehiko Tanaka, Daiki Kimura, Michiaki Tatsubori

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.CL

Comments Findings of EMNLP 2022. Code available at: https://github.com/ibm/diffg-rl

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.08965 2022-11-29 cs.CV cs.AI cs.CL cs.MM 62%

GSRFormer: Grounded Situation Recognition Transformer with Alternate Semantic Attention Refinement

Zhi-Qi Cheng, Qi Dai, Siyao Li, Teruko Mitamura, Alexander G. Hauptmann

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.CL

Comments ACM Multimedia 2022 (Oral), Code: https://github.com/zhiqic/GSRFormer

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.03235 2022-11-29 cs.AI cs.CE cs.LG cs.MS 62%

Simulation Intelligence: Towards a New Generation of Scientific Methods

Alexander Lavin, David Krakauer, Hector Zenil, Justin Gottschlich, Tim Mattson, Johann Brehmer, Anima Anandkumar, Sanjay Choudry, Kamil Rocki, Atılım Güneş Baydin, Carina Prunkl, Brooks Paige, Olexandr Isayev, Erik Peterson, Peter L. McMahon, Jakob Macke, Kyle Cranmer, Jiaxin Zhang, Haruko Wainwright, Adi Hanuka, Manuela Veloso, Samuel Assefa, Stephan Zheng, Avi Pfeffer

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.10419 2022-11-21 q-bio.NC cs.AI cs.LG 62%

A Neural Active Inference Model of Perceptual-Motor Learning

Zhizhuo Yang, Gabriel J. Diaz, Brett R. Fajen, Reynold Bailey, Alexander Ororbia

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 16 pages including references, 6 figures. Submitted to Frontiers in Computational Neuroscience

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.01078 2022-11-11 cs.LG cs.AI 62%

Deep Transformer Q-Networks for Partially Observable Reinforcement Learning

Kevin Esslinger, Robert Platt, Christopher Amato

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.04251 2022-11-09 cs.LG cs.AI 62%

State Advantage Weighting for Offline RL

Jiafei Lyu, Aicheng Gong, Le Wan, Zongqing Lu, Xiu Li

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 3rd Offline RL workshop at NeurIPS 2022. arXiv admin note: text overlap with arXiv:2206.07989

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.03281 2022-11-08 cs.LG cs.AI 62%

Reward-Predictive Clustering

Lucas Lehnert, Michael J. Frank, Michael L. Littman

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.02100 2022-11-07 cs.LG cs.AI 62%

Contrastive Value Learning: Implicit Models for Simple Offline RL

Bogdan Mazoure, Benjamin Eysenbach, Ofir Nachum, Jonathan Tompson

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Deep Reinforcement Learning Workshop, NeurIPS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.10638 2022-11-07 cs.IR cs.AI cs.LG 62%

Digital Human Interactive Recommendation Decision-Making Based on Reinforcement Learning

Xiong Junwu, Xiaoyun Feng, YunZhou Shi, James Zhang, Zhongzhou Zhao, Wei Zhou

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 9 pages, 1 figure, 1 table, the paper has been accepted and this is the final camera-ready for NeurIPS 2022 Workshop on Human in the Loop Learning, https://neurips-hill.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.15670 2022-10-31 cs.LG cs.AI 62%

Knowledge-Guided Exploration in Deep Reinforcement Learning

Sahisnu Mazumder, Bing Liu, Shuai Wang, Yingxuan Zhu, Xiaotian Yin, Lifeng Liu, Jian Li

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments This paper is an extended and revised version of the work: "Action permissibility in deep reinforcement learning and application to autonomous driving", KDD'18 Deep Learning Day (2018)

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.13383 2022-10-25 cs.AI cs.LG 62%

Evaluating Long-Term Memory in 3D Mazes

Jurgis Pasukonis, Timothy Lillicrap, Danijar Hafner

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Project website: https://github.com/jurgisp/memory-maze

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.09496 2022-10-25 cs.LG cs.AI 62%

CEIP: Combining Explicit and Implicit Priors for Reinforcement Learning with Demonstrations

Kai Yan, Alexander G. Schwing, Yu-Xiong Wang

专题命中 记忆与上下文管理 :autonomous agent(abstract);分类 cs.AI、cs.LG

Comments 27 pages; published as NeurIPS 2022 poster paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.11348 2022-10-21 cs.LG cs.AI cs.RO 62%

Hypernetworks in Meta-Reinforcement Learning

Jacob Beck, Matthew Thomas Jackson, Risto Vuorio, Shimon Whiteson

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Published at CoRL 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.08384 2022-10-18 cs.CL cs.LG 62%

Revisiting the Roles of "Text" in Text Games

Yi Gu, Shunyu Yao, Chuang Gan, Joshua B. Tenenbaum, Mo Yu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.02363 2022-10-13 cs.LG cs.AI cs.NE math.OC 62%

Meta-Reinforcement Learning with Self-Modifying Networks

Mathieu Chalvidal, Thomas Serre, Rufin VanRullen

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Published at Neurips 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.04157 2022-10-11 cs.LG cs.AI math.OC stat.ML 62%

The Role of Coverage in Online Reinforcement Learning

Tengyang Xie, Dylan J. Foster, Yu Bai, Nan Jiang, Sham M. Kakade

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.01231 2022-10-05 cs.LG cs.AI 62%

Interpretable Option Discovery using Deep Q-Learning and Variational Autoencoders

Per-Arne Andersen, Ole-Christoffer Granmo, Morten Goodwin

专题命中 记忆与上下文管理 :autonomous agent(abstract);分类 cs.AI、cs.LG

Comments 12 pages, 5 figures, Proceedings of the 3rd International Conference on Intelligent Technologies and Applications

Journal ref 2021 Springer Nature Switzerland AG

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.00859 2022-10-04 cs.SE cs.LG 62%

Requirements Engineering for Machine Learning: A Review and Reflection

Zhongyi Pei, Lin Liu, Chen Wang, Jianmin Wang

专题命中 记忆与上下文管理 :workflow(abstract);分类 cs.LG、cs.SE

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.08746 2022-09-26 cs.LG cs.AI cs.CR 62%

Real-time Adversarial Perturbations against Deep Reinforcement Learning Policies: Attacks and Defenses

Buse G. A. Tekgul, Shelly Wang, Samuel Marchal, N. Asokan

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Will appear in the proceedings of ESORICS 2022; 13 pages, 6 figures, 6 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.07928 2022-09-23 cs.AI cs.CL cs.SY eess.SY 62%

The BLue Amazon Brain (BLAB): A Modular Architecture of Services about the Brazilian Maritime Territory

Paulo Pirozelli, Ais B. R. Castro, Ana Luiza C. de Oliveira, André S. Oliveira, Flávio N. Cação, Igor C. Silveira, João G. M. Campos, Laura C. Motheo, Leticia F. Figueiredo, Lucas F. A. O. Pellicer, Marcelo A. José, Marcos M. José, Pedro de M. Ligabue, Ricardo S. Grava, Rodrigo M. Tavares, Vinícius B. Matos, Yan V. Sym, Anna H. R. Costa, Anarosa A. F. Brandão, Denis D. Mauá, Fabio G. Cozman, Sarajane M. Peres

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.CL

Journal ref AI: Modeling Oceans and Climate Change (IJCAI-ECAI), 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.09097 2022-09-20 cs.CV cs.AI cs.LG cs.RO 62%

Disentangling Shape and Pose for Object-Centric Deep Active Inference Models

Stefano Ferraro, Toon Van de Maele, Pietro Mazzaglia, Tim Verbelen, Bart Dhoedt

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.13451 2022-09-20 cs.LG cs.AI stat.ML 62%

Follow-the-Perturbed-Leader for Adversarial Markov Decision Processes with Bandit Feedback

Yan Dai, Haipeng Luo, Liyu Chen

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI、cs.LG

Comments Accepted to NeurIPS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.05840 2022-09-14 cs.CL cs.AI 62%

Visual Recipe Flow: A Dataset for Learning Visual State Changes of Objects with Recipe Flows

Keisuke Shirai, Atsushi Hashimoto, Taichi Nishimura, Hirotaka Kameko, Shuhei Kurita, Yoshitaka Ushiku, Shinsuke Mori

专题命中 记忆与上下文管理 :workflow(abstract);分类 cs.AI、cs.CL

Comments COLING 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.05302 2022-09-13 cs.LG cs.AI 62%

Unified State Representation Learning under Data Augmentation

Taylor Hearn, Sravan Jayanthi, Sehoon Ha

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 5 pages, 3 figures, 1 table, Georgia Tech CS 8803: Deep Reinforcement Learning for Intelligent Control

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.01876 2022-09-07 cs.LG cs.AI cs.IR cs.NI cs.SI 62%

SlateFree: a Model-Free Decomposition for Reinforcement Learning with Slate Actions

Anastasios Giovanidis

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 12 pages, 9 sub-figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2209.00459 2022-09-02 cs.AI cs.HC cs.LG 62%

Generative Personas That Behave and Experience Like Humans

Matthew Barthet, Ahmed Khalifa, Antonios Liapis, Georgios N. Yannakakis

专题命中 记忆与上下文管理 :AI agent(abstract);分类 cs.AI、cs.LG

Comments 10 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.04822 2022-08-31 cs.LG cs.AI 62%

Generalized Reinforcement Learning: Experience Particles, Action Operator, Reinforcement Field, Memory Association, and Decision Concepts

Po-Hsiang Chiu, Manfred Huber

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments 37 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.05056 2022-08-17 cs.LG cs.AI 62%

Model-Free Generative Replay for Lifelong Reinforcement Learning: Application to Starcraft-2

Zachary Daniels, Aswin Raghavan, Jesse Hostetler, Abrar Rahman, Indranil Sur, Michael Piacentino, Ajay Divakaran

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

Comments Accepted to the First Conference on Lifelong Learning Agents (CoLLAs 2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.08926 2022-07-26 cs.AI cs.LG q-bio.NC 62%

Generating Explanations from Deep Reinforcement Learning Using Episodic Memory

Sam Blakeman, Denis Mareschal

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏