arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 4758 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 记忆与上下文管理 4758 篇

1707.08553 2020-01-28 cs.LG 57%

Direct Load Control of Thermostatically Controlled Loads Based on Sparse Observations Using Deep Reinforcement Learning

Frederik Ruelens, Bert J. Claessens, Peter Vrancx, Fred Spiessens, Geert Deconinck

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments submitted and waiting review in IEEE transactions on smart grid 2017

Journal ref CSEE Journal of Power and Energy Systems, Vol. 5, Iss. 4, Dec. 2019, pp. 423-432

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.05636 2020-01-17 cs.LG stat.ML 57%

MIME: Mutual Information Minimisation Exploration

Haitao Xu, Brendan McCane, Lech Szymanski, Craig Atkinson

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.11461 2020-01-17 cs.MA cs.AI 57%

Action Semantics Network: Considering the Effects of Actions in Multiagent Systems

Weixun Wang, Tianpei Yang, Yong Liu, Jianye Hao, Xiaotian Hao, Yujing Hu, Yingfeng Chen, Changjie Fan, Yang Gao

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments accepted by ICLR2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.11832 2020-01-17 cs.LG cs.CV stat.ML 57%

Snooping Attacks on Deep Reinforcement Learning

Matthew Inkawhich, Yiran Chen, Hai Li

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 13 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.10515 2019-12-24 cs.LO cs.AI 57%

Bringing Belief Base Change into Dynamic Epistemic Logic

Marlo Souza, Álvaro Moreira

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Published at DaLI 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.04279 2019-12-19 cs.LG stat.ML 57%

Exploration via Hindsight Goal Generation

Zhizhou Ren, Kefan Dong, Yuan Zhou, Qiang Liu, Jian Peng

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Thirty-third Conference on Neural Information Processing Systems (NeurIPS 2019)

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.09682 2019-11-25 quant-ph cs.LG 57%

Quantum Observables for continuous control of the Quantum Approximate Optimization Algorithm via Reinforcement Learning

Artur Garcia-Saez, Jordi Riu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 6 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.08701 2019-11-21 cs.LG stat.ML 57%

Bayesian Curiosity for Efficient Exploration in Reinforcement Learning

Tom Blau, Lionel Ott, Fabio Ramos

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.07224 2019-11-19 cs.LG cs.RO stat.ML 57%

Learning from Trajectories via Subgoal Discovery

Sujoy Paul, Jeroen van Baar, Amit K. Roy-Chowdhury

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments NeurIPS 2019 Accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
1805.12375 2019-11-13 cs.LG stat.ML 57%

Sample-Efficient Deep Reinforcement Learning via Episodic Backward Update

Su Young Lee, Sungik Choi, Sae-Young Chung

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments NeurIPS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.05572 2019-11-11 cs.CL 57%

Proactive Human-Machine Conversation with Explicit Conversation Goals

Wenquan Wu, Zhen Guo, Xiangyang Zhou, Hua Wu, Xiyuan Zhang, Rongzhong Lian, Haifeng Wang

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments Accepted by ACL 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.00459 2019-11-04 cs.LG stat.ML 57%

Positive-Unlabeled Reward Learning

Danfei Xu, Misha Denil

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.04742 2019-10-31 cs.LG stat.ML 57%

Online Continual Learning with Maximally Interfered Retrieval

Rahaf Aljundi, Lucas Caccia, Eugene Belilovsky, Massimo Caccia, Min Lin, Laurent Charlin, Tinne Tuytelaars

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Journal ref NeurIPS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.12980 2019-10-30 cs.LG stat.ML 57%

Learning Transferable Graph Exploration

Hanjun Dai, Yujia Li, Chenglong Wang, Rishabh Singh, Po-Sen Huang, Pushmeet Kohli

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments To appear in NeurIPS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.12807 2019-10-29 stat.ML cs.LG 57%

Better Exploration with Optimistic Actor-Critic

Kamil Ciosek, Quan Vuong, Robert Loftin, Katja Hofmann

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 20 pages (including supplement)

Journal ref NeurIPS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.09314 2019-10-29 cs.LG stat.ML 57%

Meta-Inverse Reinforcement Learning with Probabilistic Context Variables

Lantao Yu, Tianhe Yu, Chelsea Finn, Stefano Ermon

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments NeurIPS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.01108 2019-10-28 cs.LG cs.CV stat.ML 57%

Injective State-Image Mapping facilitates Visual Adversarial Imitation Learning

Subhajit Chaudhury, Daiki Kimura, Asim Munawar, Ryuki Tachibana

专题命中 记忆与上下文管理 :autonomous agent(abstract);分类 cs.LG

Comments Updated the paper to match with version accepted at IEEE MMSP 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.01806 2019-10-15 cs.AI 57%

Deep Q-Network for Angry Birds

Ekaterina Nikonova, Jakub Gemrot

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.09454 2019-09-23 cs.AI cs.LO 57%

Memory Management in Resource-Bounded Agents

Valentina Pitoni

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments In Proceedings ICLP 2019, arXiv:1909.07646. arXiv admin note: substantial text overlap with arXiv:1909.08256

Journal ref EPTCS 306, 2019, pp. 452-460

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.12227 2019-09-20 math.PR cs.LG 57%

Reinforcement with Fading Memories

Kuang Xu, Se-Young Yun

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Forthcoming in Mathematics of Operations Research; An extended abstract appeared in the proceedings of ACM SIGMETRICS 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.08161 2019-09-19 cs.HC cs.AI cs.RO 57%

Multimodal Continuation-style Architectures for Human-Robot Interaction

Nikhil Krishnaswamy, James Pustejovsky

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Advances in Cognitive Systems Cognitive Vision Workshop (2019), 8 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.02265 2019-08-30 cs.CL 57%

Comprehensible Context-driven Text Game Playing

Xusen Yin, Jonathan May

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments IEEE Conference on Games 2019 Long Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.10714 2019-08-29 cs.LG cs.NE stat.ML 57%

Automated Architecture Design for Deep Neural Networks

Steven Abreu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Undergraduate Thesis

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.08835 2019-08-26 cs.CL 57%

Deep Learning Based Chatbot Models

Richard Csaky

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments 67 pages. Written in October of 2017 for a university conference. In April of 2019, it won first place at the Hungarian Scientific Students' Associations Report, which is a national competition-like conference for students

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.05451 2019-08-16 cs.LG stat.ML 57%

Mapping State Space using Landmarks for Universal Goal Reaching

Zhiao Huang, Fangchen Liu, Hao Su

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.04660 2019-08-14 cs.CL 57%

Playing log(N)-Questions over Sentences

Peter Potash, Kaheer Suleman

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments 5 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.07252 2019-07-23 cs.RO cs.CV cs.LG 57%

Sim-to-Real via Sim-to-Sim: Data-efficient Robotic Grasping via Randomized-to-Canonical Adaptation Networks

Stephen James, Paul Wohlhart, Mrinal Kalakrishnan, Dmitry Kalashnikov, Alex Irpan, Julian Ibarz, Sergey Levine, Raia Hadsell, Konstantinos Bousmalis

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition (CVPR 2019)

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.06396 2019-07-16 cs.LG stat.ML 57%

A Dual Memory Structure for Efficient Use of Replay Memory in Deep Reinforcement Learning

Wonshick Ko, Dong Eui Chang

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 4 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.03588 2019-07-09 eess.SY cs.IT cs.LG cs.SY math.IT 57%

A New Approach to Distributed Hypothesis Testing and Non-Bayesian Learning: Improved Learning Rate and Byzantine-Resilience

Aritra Mitra, John A. Richards, Shreyas Sundaram

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments arXiv admin note: text overlap with arXiv:1903.05817

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.03116 2019-07-09 cs.LG cs.RO stat.ML 57%

Intrinsic Motivation Driven Intuitive Physics Learning using Deep Reinforcement Learning with Intrinsic Reward Normalization

JaeWon Choi, Sung-eui Yoon

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏