arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 4758 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 记忆与上下文管理 4758 篇

2010.02316 2020-10-07 cs.CL 57%

Sentiment Analysis for Reinforcement Learning

Ameet Deshpande, Eve Fleisig

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments Work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.06283 2020-10-07 cs.RO cs.CV cs.LG 57%

Self-Supervised Learning of State Estimation for Manipulating Deformable Linear Objects

Mengyuan Yan, Yilin Zhu, Ning Jin, Jeannette Bohg

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

Comments v3: update acknowledgements

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.12330 2020-09-28 cs.SE 57%

Synthesis of Infinite-State Systems with Random Behavior

Andreas Katis, Grigory Fedyukovich, Jeffrey Chen, David Greve, Sanjai Rayadurgam, Michael W. Whalen

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.SE

Comments 12 pages, 7 figures, The 35th IEEE/ACM International Conference on Automated Software Engineering (ASE 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.10827 2020-09-25 cs.RO cs.LG 57%

Data-Driven Distributed State Estimation and Behavior Modeling in Sensor Networks

Rui Yu, Zhenyuan Yuan, Minghui Zhu, Zihan Zhou

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

Comments 8 pages, 5 figures. To appear at the IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.10855 2020-09-24 cs.CL 57%

Controlling Style in Generated Dialogue

Eric Michael Smith, Diana Gonzalez-Rico, Emily Dinan, Y-Lan Boureau

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.10430 2020-09-23 cs.CL 57%

Dual Learning for Dialogue State Tracking

Zhi Chen, Lu Chen, Yanbin Zhao, Su Zhu, Kai Yu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments 7 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.09543 2020-09-22 eess.SP cs.LG cs.SY eess.SY 57%

State-of-Charge Estimation of a Li-Ion Battery using Deep Forward Neural Networks

Alexandre Barbosa de Lima, Maurício B. C. Salles, José Roberto Cardoso

专题命中 记忆与上下文管理 :workflow(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.08278 2020-09-18 cs.LG physics.bio-ph 57%

Accelerated solving of coupled, non-linear ODEs through LSTM-AI

Camila Faccini de Lima, Juliano Ferrari Gianlupi, John Metzcar, Juliette Zerick

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 7 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2001.10283 2020-09-02 quant-ph cs.LG 57%

Real-time calibration of coherent-state receivers: learning by trial and error

M. Bilkis, M. Rosati, R. Morral Yepes, J. Calsamiglia

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 14+3 pages, 11 figures

Journal ref Phys. Rev. Research 2, 033295 (2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.11491 2020-08-31 q-bio.NC cs.CV cs.LG cs.NE 57%

Selective Particle Attention: Visual Feature-Based Attention in Deep Reinforcement Learning

Sam Blakeman, Denis Mareschal

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1909.09228 2020-08-26 cs.AI cs.MA cs.SI 57%

Non-Bayesian Social Learning with Uncertain Models

James Z. Hare, Cesar A. Uribe, Lance Kaplan, Ali Jadbabaie

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1908.09008 2020-08-19 cs.CV cs.LG stat.ML 57%

Conditional Flow Variational Autoencoders for Structured Sequence Prediction

Apratim Bhattacharyya, Michael Hanselmann, Mario Fritz, Bernt Schiele, Christoph-Nikolas Straehle

专题命中 记忆与上下文管理 :autonomous agent(abstract);分类 cs.LG

Comments To appear at Bayesian Deep Learning and Machine Learning for Autonomous Driving @NeurIPS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.06973 2020-08-18 cs.LG stat.ML 57%

An adaptive synchronization approach for weights of deep reinforcement learning

S. Amirreza Badran, Mansoor Rezghi

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.03517 2020-08-18 stat.ML cs.LG 57%

No-Regret Exploration in Goal-Oriented Reinforcement Learning

Jean Tarbouriech, Evrard Garcelon, Michal Valko, Matteo Pirotta, Alessandro Lazaric

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Journal ref International Conference on Machine Learning (ICML 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.06439 2020-08-17 cs.CV cs.LG 57%

RODEO: Replay for Online Object Detection

Manoj Acharya, Tyler L. Hayes, Christopher Kanan

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Accepted for poster presentation at BMVC2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.05124 2020-08-13 cs.LG stat.ML 57%

Leveraging Automated Mixed-Low-Precision Quantization for tiny edge microcontrollers

Manuele Rusci, Marco Fariselli, Alessandro Capotondi, Luca Benini

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.04500 2020-08-12 cs.LG cs.CR stat.ML 57%

Towards Plausible Differentially Private ADMM Based Distributed Machine Learning

Jiahao Ding, Jingyi Wang, Guannan Liang, Jinbo Bi, Miao Pan

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Comments: Accepted for publication in CIKM'20

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.02356 2020-08-07 cs.CV cs.AI 57%

A Neural-Symbolic Framework for Mental Simulation

Michael Kissner

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Dissertation

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.02624 2020-08-04 cs.LG cs.RO stat.ML 57%

Learning Efficient Representation for Intrinsic Motivation

Ruihan Zhao, Stas Tiomkin, Pieter Abbeel

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.11740 2020-07-24 cs.AI cs.HC cs.RO 57%

Improving Competence for Reliable Autonomy

Connor Basich, Justin Svegliato, Kyle Hollins Wray, Stefan J. Witwicki, Shlomo Zilberstein

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments In Proceedings AREA 2020, arXiv:2007.11260

Journal ref EPTCS 319, 2020, pp. 37-53

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.10960 2020-07-23 cs.LG stat.ML 57%

Adaptive Traffic Control with Deep Reinforcement Learning: Towards State-of-the-art and Beyond

Siavash Alemzadeh, Ramin Moslemi, Ratnesh Sharma, Mehran Mesbahi

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.08017 2020-07-16 cs.LG stat.ML 57%

Implicit Generative Modeling for Efficient Exploration

Neale Ratzlaff, Qinxun Bai, Li Fuxin, Wei Xu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 14 pages, 9 figures, Accepted to ICML 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.04862 2020-07-14 cs.AI 57%

Attention or memory? Neurointerpretable agents in space and time

Lennart Bramlage, Aurelio Cortese

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 8 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.03310 2020-07-08 cs.CV cs.CL 57%

DAM: Deliberation, Abandon and Memory Networks for Generating Detailed and Non-repetitive Responses in Visual Dialogue

Xiaoze Jiang, Jing Yu, Yajing Sun, Zengchang Qin, Zihao Zhu, Yue Hu, Qi Wu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments Accepted by IJCAI 2020. SOLE copyright holder is IJCAI (International Joint Conferences on Artificial Intelligence)

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.02823 2020-07-07 cs.AI cs.LO econ.TH 57%

Dynamic Awareness

Joseph Y. Halpern, Evan Piermont

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments To appear in the 17th International Conference on Principles of Knowledge Representation and Reasoning

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.00970 2020-07-06 cs.LG cs.CV cs.NE stat.ML 57%

MPLP: Learning a Message Passing Learning Protocol

Ettore Randazzo, Eyvind Niklasson, Alexander Mordvintsev

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Code at https://github.com/google-research/self-organising-systems/tree/master/mplp; code base link fixed

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.00722 2020-07-03 cs.LG stat.ML 57%

Sequential Transfer in Reinforcement Learning with a Generative Model

Andrea Tirinzoni, Riccardo Poiani, Marcello Restelli

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments ICML 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.13160 2020-06-24 cs.LG stat.ML 57%

Environment Shaping in Reinforcement Learning using State Abstraction

Parameswaran Kamalaruban, Rati Devidze, Volkan Cevher, Adish Singla

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1907.13266 2020-06-22 eess.SP cs.LG 57%

PrecoderNet: Hybrid Beamforming for Millimeter Wave Systems with Deep Reinforcement Learning

Qisheng Wang, Keming Feng, Xiao Li, Shi Jin

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 13 pages, 6 figures

Journal ref IEEE Wireless Communication Letters, 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.01963 2020-06-16 cs.LG stat.ML 57%

Mutual Information-based State-Control for Intrinsically Motivated Reinforcement Learning

Rui Zhao, Yang Gao, Pieter Abbeel, Volker Tresp, Wei Xu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 15 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏