arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 4758 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 记忆与上下文管理 4758 篇

2107.14702 2021-11-02 cs.GT cs.LG stat.ML 57%

Towards General Function Approximation in Zero-Sum Markov Games

Baihe Huang, Jason D. Lee, Zhaoran Wang, Zhuoran Yang

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.00511 2021-11-02 cs.LG stat.ML 57%

Curriculum Learning with a Progression Function

Andrea Bassich, Francesco Foglino, Matteo Leonetti, Daniel Kudenko

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.11660 2021-11-01 cs.LG 57%

Detecting Rewards Deterioration in Episodic Reinforcement Learning

Ido Greenberg, Shie Mannor

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments ICML 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.15097 2021-10-29 cs.LG cs.IR 57%

Choosing the Best of Both Worlds: Diverse and Novel Recommendations through Multi-Objective Reinforcement Learning

Dusan Stamenkovic, Alexandros Karatzoglou, Ioannis Arapakis, Xin Xin, Kleomenis Katevas

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 9 pages, 4 figures, Proc. ACM WSDM, 2022 In Proceedings of the 15th ACM International Conference on Web Search and Data Mining (WSDM '22), February 21-25, 2022, Phoenix, Arizona

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.11737 2021-10-25 cs.AI cs.MA 57%

Measuring the Non-Transitivity in Chess

Ricky Sanjaya, Jun Wang, Yaodong Yang

专题命中 记忆与上下文管理 :AI agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.11106 2021-10-22 cs.CV cs.LG 57%

Reinforcement Learning Based Optimal Camera Placement for Depth Observation of Indoor Scenes

Yichuan Chen, Manabu Tsukada, Hiroshi Esaki

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Accepted to IEEE International Conference on Networking, Sensing and Control (ICNSC) 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.10905 2021-10-22 cs.RO cs.AI 57%

Efficient Robotic Manipulation Through Offline-to-Online Reinforcement Learning and Goal-Aware State Information

Jin Li, Xianyuan Zhan, Zixu Xiao, Guyue Zhou

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.00919 2021-10-20 cs.CV cs.AI 57%

Continual Prototype Evolution: Learning Online from Non-Stationary Data Streams

Matthias De Lange, Tinne Tuytelaars

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 10 pages, code publicly available

Journal ref Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2021, pp. 8250-8259

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.08032 2021-10-18 cs.CL 57%

UniDS: A Unified Dialogue System for Chit-Chat and Task-oriented Dialogues

Xinyan Zhao, Bin He, Yasheng Wang, Yitong Li, Fei Mi, Yajiao Liu, Xin Jiang, Qun Liu, Huanhuan Chen

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.03655 2021-10-05 quant-ph cond-mat.quant-gas cs.LG physics.comp-ph 57%

Reinforcement Learning for Many-Body Ground-State Preparation Inspired by Counterdiabatic Driving

Jiahao Yao, Lin Lin, Marin Bukov

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Journal ref Phys. Rev. X 11, 031070 (2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.12691 2021-09-28 cs.AI 57%

Applying supervised and reinforcement learning methods to create neural-network-based agents for playing StarCraft II

Michał Opanowicz

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.06506 2021-09-28 cs.RO cs.LG 57%

SAINT-ACC: Safety-Aware Intelligent Adaptive Cruise Control for Autonomous Vehicles Using Deep Reinforcement Learning

Lokesh Das, Myounggyu Won

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Accepted for publication in ICML, 2021

Journal ref Proceedings of the 38th International Conference on Machine Learning (2021) 2445-2455

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.01738 2021-09-28 cs.LG cs.RO stat.ML 57%

State Representation Learning from Demonstration

Astrid Merckling, Alexandre Coninx, Loic Cressot, Stéphane Doncieux, Nicolas Perrin-Gilbert

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Published as a conference paper at LOD 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.09504 2021-09-21 eess.SP cs.LG 57%

The Devil Is in the Details: An Efficient Convolutional Neural Network for Transport Mode Detection

Hugues Moreau, Andréa Vassilev, Liming Chen

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

Comments 13 pages, 5 figures, 6 tables. Published in IEEE Transactions on Intelligent Transportation Systems

Journal ref IEEE Transactions on Intelligent Transportation Systems, 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.05042 2021-09-14 cs.CL 57%

Reference-Centric Models for Grounded Collaborative Dialogue

Daniel Fried, Justin T. Chiu, Dan Klein

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments EMNLP 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.03202 2021-09-08 cs.AI 57%

On the impact of MDP design for Reinforcement Learning agents in Resource Management

Renato Luiz de Freitas Cunha, Luiz Chaimowicz

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 15 pages, 6 figures. Accepted for publication at BRACIS 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.02178 2021-09-07 q-bio.QM cs.LG 57%

A Critical Review of the state-of-the-art on Deep Neural Networks for Blood Glucose Prediction in Patients with Diabetes

Felix Tena, Oscar Garnica, Juan Lanchares, J. Ignacio Hidalgo

专题命中 记忆与上下文管理 :workflow(abstract);分类 cs.LG

Comments 17 pages, 20 figures and 16 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.00077 2021-09-02 cs.CL 57%

Interactive Machine Comprehension with Dynamic Knowledge Graphs

Xingdi Yuan

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments Accepted at EMNLP 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.12198 2021-08-30 cs.NI cs.LG cs.SY eess.SY 57%

Deep Reinforcement Learning for Wireless Resource Allocation Using Buffer State Information

Eike-Manuel Bansbach, Victor Eliachevitch, Laurent Schmalen

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments accepted for publication at GLOBECOM 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.09701 2021-08-20 cs.CV cs.LG 57%

Always Be Dreaming: A New Approach for Data-Free Class-Incremental Learning

James Smith, Yen-Chang Hsu, Jonathan Balloch, Yilin Shen, Hongxia Jin, Zsolt Kira

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Accepted by the 2021 International Conference on Computer Vision (ICCV 2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.04523 2021-08-11 cs.FL cs.LO cs.MA cs.SE cs.SY eess.SY 57%

Decentralized Observation of Discrete-Event Systems: At Least One Can Tell

Stavros Tripakis, Karen Rudie

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.SE

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.02628 2021-08-06 cs.LG 57%

A New State-of-the-Art Transformers-Based Load Forecaster on the Smart Grid Domain

Andre Luiz Farias Novaes, Rui Alexandre de Matos Araujo, Jose Figueiredo, Lucas Aguiar Pavanelli

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.08052 2021-08-03 cs.RO cs.AI 57%

A Learning Approach to Robot-Agnostic Force-Guided High Precision Assembly

Jieliang Luo, Hui Li

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 7 pages, 6 figures; accepted to IROS 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.14345 2021-08-02 cs.RO cs.HC cs.LG 57%

Modeling User Empathy Elicited by a Robot Storyteller

Leena Mathur, Micol Spitale, Hao Xi, Jieyun Li, Maja J Matarić

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Accepted for publication and oral presentation at the International Conference on Affective Computing and Intelligent Interaction (ACII 2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.11587 2021-07-27 cs.LG 57%

Model-based micro-data reinforcement learning: what are the crucial model properties and which model to choose?

Balázs Kégl, Gabriel Hurtado, Albert Thomas

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Published at International Conference on Learning Representations, 2021: https://openreview.net/forum?id=p5uylG94S68

详情

展开后加载摘要…

URL PDF HTML 收藏
2104.01066 2021-07-21 cs.SI cs.AI cs.MA cs.SY eess.SY 57%

An active inference model of collective intelligence

Rafael Kaufmann, Pranav Gupta, Jacob Taylor

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 32 pages, 10 figures, manuscript under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.07031 2021-07-16 cs.AI 57%

Experimental Evidence that Empowerment May Drive Exploration in Sparse-Reward Environments

Francesco Massari, Martin Biehl, Lisa Meeden, Ryota Kanai

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 6 pages, 3 figures, to be published in proceedings of the International Conference on Development and Learning 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2003.08938 2021-07-15 cs.LG stat.ML 57%

Robust Deep Reinforcement Learning against Adversarial Perturbations on State Observations

Huan Zhang, Hongge Chen, Chaowei Xiao, Bo Li, Mingyan Liu, Duane Boning, Cho-Jui Hsieh

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Huan Zhang and Hongge Chen contributed equally

Journal ref Advances in Neural Information Processing Systems 33 (2020): 21024-21037

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.06117 2021-07-14 cs.GT cs.CC cs.LG cs.MA econ.TH 57%

The Platform Design Problem

Christos Papadimitriou, Kiran Vodrahalli, Mihalis Yannakakis

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments updated with more results

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.12194 2021-07-06 cs.RO cs.LG 57%

Uncertainty-Aware Model-Based Reinforcement Learning with Application to Autonomous Driving

Jingda Wu, Zhiyu Huang, Chen Lv

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏