arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 4758 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 记忆与上下文管理 4758 篇

2411.14085 2024-11-22 cs.LG 57%

Exploration by Running Away from the Past

Paul-Antoine Le Tolguenec, Yann Besse, Florent Teichteil-Koenigsbuch, Dennis G. Wilson, Emmanuel Rachelson

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.11188 2024-11-19 cs.LG 57%

AMAGO-2: Breaking the Multi-Task Barrier in Meta-Reinforcement Learning with Transformers

Jake Grigsby, Justin Sasek, Samyak Parajuli, Daniel Adebi, Amy Zhang, Yuke Zhu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09169 2024-11-15 cs.MA cs.AI cs.CY cs.HC nlin.AO 57%

Artificial Theory of Mind and Self-Guided Social Organisation

Michael S. Harré, Jaime Ruiz-Serra, Catherine Drysdale

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 4 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.08392 2024-11-14 cs.AI 57%

RLInspect: An Interactive Visual Approach to Assess Reinforcement Learning Algorithm

Geetansh Kalra, Divye Singh, Justin Jose

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.07954 2024-11-14 cs.LG cs.RO 57%

Learning Memory Mechanisms for Decision Making through Demonstrations

William Yue, Bo Liu, Peter Stone

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.05937 2024-11-13 cs.DC cs.AI cs.SY eess.SY 57%

A Deep Recurrent-Reinforcement Learning Method for Intelligent AutoScaling of Serverless Functions

Siddharth Agarwal, Maria A. Rodriguez, Rajkumar Buyya

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 12 pages, 15 figures, 4 tables

Journal ref in IEEE Transactions on Services Computing, vol. 17, no. 5, pp. 1899-1910, Sept.-Oct. 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.06034 2024-11-12 cs.AI 57%

CROPS: A Deployable Crop Management System Over All Possible State Availabilities

Jing Wu, Zhixin Lai, Shengjie Liu, Suiyao Chen, Ran Tao, Pan Zhao, Chuyuan Tao, Yikun Cheng, Naira Hovakimyan

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.10108 2024-11-11 cs.IR cs.AI 57%

On Generative Agents in Recommendation

An Zhang, Yuxin Chen, Leheng Sheng, Xiang Wang, Tat-Seng Chua

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments SIGIR 2024 perspective paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.04976 2024-11-08 cs.LG 57%

Noisy Zero-Shot Coordination: Breaking The Common Knowledge Assumption In Zero-Shot Coordination Games

Usman Anwar, Ashish Pandian, Jia Wan, David Krueger, Jakob Foerster

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.05269 2024-11-07 cs.CV cs.LG 57%

LifelongMemory: Leveraging LLMs for Answering Queries in Long-form Egocentric Videos

Ying Wang, Yanlai Yang, Mengye Ren

专题命中 记忆与上下文管理 :agentic(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.14751 2024-11-06 cs.LG 57%

AGILE: A Novel Reinforcement Learning Framework of LLM Agents

Peiyuan Feng, Yichen He, Guanhua Huang, Yuan Lin, Hanchong Zhang, Yuchen Zhang, Hang Li

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments accepted by NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2109.09013 2024-11-05 cs.LG eess.SP 57%

Hydroelectric Generation Forecasting with Long Short Term Memory (LSTM) Based Deep Learning Model for Turkey

Mehmet Bulut

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.17515 2024-10-31 cs.AI 57%

From News to Forecast: Integrating Event Analysis in LLM-Based Time Series Forecasting with Reflection

Xinlei Wang, Maike Feng, Jing Qiu, Jinjin Gu, Junhua Zhao

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments This paper has been accepted for NeurIPS 2024. Code and data are available at https://github.com/ameliawong1996/From_News_to_Forecast

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.20990 2024-10-29 cs.RO cs.LG cs.SY eess.SY 57%

Reference-Free Formula Drift with Reinforcement Learning: From Driving Data to Tire Energy-Inspired, Real-World Policies

Franck Djeumou, Michael Thompson, Makoto Suminaka, John Subosits

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Initial submission to ICRA 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2107.03263 2024-10-29 cs.LG 57%

Bandits with Stochastic Experts: Constant Regret, Empirical Experts and Episodes

Nihal Sharma, Rajat Sen, Soumya Basu, Karthikeyan Shanmugam, Sanjay Shakkottai

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Journal ref ACM Transactions on Modeling and Performance Evaluation of Computing Systems 9.3 (2024): 1-33

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.20669 2024-10-23 quant-ph cs.AI 57%

A Tutorial on the Use of Physics-Informed Neural Networks to Compute the Spectrum of Quantum Systems

Lorenzo Brevi, Antonio Mandarino, Enrico Prati

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 18 pages, 4 figures

Journal ref Technologies 12, 174 (2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.16029 2024-10-22 cs.LG cs.MA 57%

Natural GaLore: Accelerating GaLore for memory-efficient LLM Training and Fine-tuning

Arijit Das

专题命中 记忆与上下文管理 :function calling(abstract);分类 cs.LG

Comments 10 pages, 3 tables, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.15396 2024-10-22 cs.CR cs.AI 57%

The Best Defense is a Good Offense: Countering LLM-Powered Cyberattacks

Daniel Ayzenshteyn, Roy Weiss, Yisroel Mirsky

专题命中 记忆与上下文管理 :autonomous agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.15019 2024-10-22 cs.CL 57%

A Survey of Ontology Expansion for Conversational Understanding

Jinggui Liang, Yuxia Wu, Yuan Fang, Hao Fei, Lizi Liao

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments Accepted by EMNLP 2024, code and data are available at this https URL: https://github.com/liangjinggui/Ontology-Expansion

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.05822 2024-10-21 cs.CV cs.AI cs.HC 57%

Encode-Store-Retrieve: Augmenting Human Memory through Language-Encoded Egocentric Perception

Junxiao Shen, John Dudley, Per Ola Kristensson

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.02912 2024-10-18 cs.RO cs.AI 57%

KOI: Accelerating Online Imitation Learning via Hybrid Key-state Guidance

Jingxian Lu, Wenke Xia, Dong Wang, Zhigang Wang, Bin Zhao, Di Hu, Xuelong Li

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Accepted by CoRL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.13430 2024-10-17 physics.chem-ph cs.LG 57%

React-OT: Optimal Transport for Generating Transition State in Chemical Reactions

Chenru Duan, Guan-Horng Liu, Yuanqi Du, Tianrong Chen, Qiyuan Zhao, Haojun Jia, Carla P. Gomes, Evangelos A. Theodorou, Heather J. Kulik

专题命中 记忆与上下文管理 :workflow(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02874 2024-10-08 cs.RO cs.AI 57%

Real-World Cooking Robot System from Recipes Based on Food State Recognition Using Foundation Models and PDDL

Naoaki Kanazawa, Kento Kawaharazuka, Yoshiki Obinata, Kei Okada, Masayuki Inaba

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI

Comments Accepted at Advanced Robotics, website - https://kanazawanaoaki.github.io/cook-from-recipe-pddl/

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.10041 2024-10-08 cs.RO cs.AI 57%

Towards Embedding Dynamic Personas in Interactive Robots: Masquerading Animated Social Kinematics (MASK)

Jeongeun Park, Taemoon Jeong, Hyeonseong Kim, Taehyun Byun, Seungyoon Shin, Keunjun Choi, Jaewoon Kwon, Taeyoon Lee, Matthew Pan, Sungjoon Choi

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Accepted at Robotics and Automation Letters

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.01011 2024-10-07 cs.LG 57%

Back to Bayesics: Uncovering Human Mobility Distributions and Anomalies with an Integrated Statistical and Neural Framework

Minxuan Duan, Yinlong Qian, Lingyi Zhao, Zihao Zhou, Zeeshan Rasheed, Rose Yu, Khurram Shafique

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.03069 2024-10-04 cs.AI 57%

"Give Me an Example Like This": Episodic Active Reinforcement Learning from Demonstrations

Muhan Hou, Koen Hindriks, A. E. Eiben, Kim Baraka

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1706.00933 2024-10-03 cs.SE 57%

Evolution of statistical analysis in empirical software engineering research: Current state and steps forward

Francisco Gomes de Oliveira Neto, Richard Torkar, Robert Feldt, Lucas Gren, Carlo A. Furia, Ziwei Huang

专题命中 记忆与上下文管理 :workflow(abstract);分类 cs.SE

Comments journal submission, 34 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.19262 2024-10-02 eess.SP cs.LG 57%

Removing the need for ground truth UWB data collection: self-supervised ranging error correction using deep reinforcement learning

Dieter Coppens, Ben Van Herbruggen, Adnan Shahid, Eli De Poorter

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 13 pages, 9 figures and 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.04332 2024-10-02 cs.LG 57%

Imitation Learning by State-Only Distribution Matching

Damian Boborzi, Christoph-Nikolas Straehle, Jens S. Buchner, Lars Mikelsons

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.13392 2024-10-01 cs.LG math.OC stat.ML 57%

Combinatorial Causal Bandits without Graph Skeleton

Shi Feng, Nuoya Xiong, Wei Chen

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 56 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏