arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 4758 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 记忆与上下文管理 4758 篇

2401.12046 2024-03-19 cs.RO cs.LG 57%

Fourier Transporter: Bi-Equivariant Robotic Manipulation in 3D

Haojie Huang, Owen Howell, Dian Wang, Xupeng Zhu, Robin Walters, Robert Platt

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments ICLR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.06086 2024-03-12 cs.AI cs.RO 57%

Towards Generalizable and Interpretable Motion Prediction: A Deep Variational Bayes Approach

Juanwu Lu, Wei Zhan, Masayoshi Tomizuka, Yeping Hu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Accepted at AISTATS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.04050 2024-03-08 cs.LG 57%

Belief-Enriched Pessimistic Q-Learning against Adversarial State Perturbations

Xiaolin Sun, Zizhan Zheng

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments ICLR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.03994 2024-03-08 cs.CV cs.LG 57%

Video Relationship Detection Using Mixture of Experts

Ala Shaabana, Zahra Gharaee, Paul Fieguth

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.07791 2024-03-06 cs.AR cs.ET cs.LG 57%

Associative Memory Based Experience Replay for Deep Reinforcement Learning

Mengyuan Li, Arman Kazemi, Ann Franchesca Laguna, X. Sharon Hu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 9 pages, 9 figures. The work was accepted by the 41st International Conference on Computer-Aided Design (ICCAD), 2022, San Diego

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.09752 2024-02-27 cs.LG 57%

Contrastive Initial State Buffer for Reinforcement Learning

Nico Messikommer, Yunlong Song, Davide Scaramuzza

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Journal ref IEEE Conference on Robotics and Automation (ICRA 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.15932 2024-02-27 cs.LG cs.SY eess.SY 57%

Scalable Volt-VAR Optimization using RLlib-IMPALA Framework: A Reinforcement Learning Approach

Alaa Selim, Yanzhu Ye, Junbo Zhao, Bo Yang

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.16292 2024-02-23 cs.RO cs.CL 57%

DiLu: A Knowledge-Driven Approach to Autonomous Driving with Large Language Models

Licheng Wen, Daocheng Fu, Xin Li, Xinyu Cai, Tao Ma, Pinlong Cai, Min Dou, Botian Shi, Liang He, Yu Qiao

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments Published as a conference paper at ICLR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.09384 2024-02-22 econ.TH cs.AI cs.CY cs.GT cs.HC 57%

Persuasion, Delegation, and Private Information in Algorithm-Assisted Decisions

Ruqing Xu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.12465 2024-02-21 cs.LG cs.NE 57%

Neuro-mimetic Task-free Unsupervised Online Learning with Continual Self-Organizing Maps

Hitesh Vaidya, Travis Desell, Ankur Mali, Alexander Ororbia

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.11364 2024-02-20 cs.AI 57%

Counting Reward Automata: Sample Efficient Reinforcement Learning Through the Exploitation of Reward Function Structure

Tristan Bester, Benjamin Rosman, Steven James, Geraud Nangue Tasse

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 14 pages, 11 Figures, Published in AAAI W25: Neuro-Symbolic Learning and Reasoning in the era of Large Language Models (NuCLeaR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.03329 2024-02-07 cs.CV cs.AI 57%

Unsupervised Salient Patch Selection for Data-Efficient Reinforcement Learning

Zhaohui Jiang, Paul Weng

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.06014 2024-02-07 cs.LG cs.GT cs.IR 57%

Online Recommendations for Agents with Discounted Adaptive Preferences

Arpit Agarwal, William Brown

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Updates for camera-ready version (ALT 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.00736 2024-01-31 cs.LG cs.GT 57%

Approximating the Shapley Value without Marginal Contributions

Patrick Kolpaczki, Viktor Bengs, Maximilian Muschalik, Eyke Hüllermeier

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.14991 2024-01-31 cs.RO cs.AI 57%

Reachability Verification Based Reliability Assessment for Deep Reinforcement Learning Controlled Robotics and Autonomous Systems

Yi Dong, Xingyu Zhao, Sen Wang, Xiaowei Huang

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.08859 2024-01-26 cs.DC cs.LG 57%

Shabari: Delayed Decision-Making for Faster and Efficient Serverless Functions

Prasoon Sinha, Kostis Kaffes, Neeraja J. Yadwadkar

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 17 pages, 14 figures, update typo in manually entered arxiv title

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.11839 2024-01-23 cs.CL cs.CY 57%

AI for social science and social science of AI: A Survey

Ruoxi Xu, Yingfei Sun, Mengjie Ren, Shiguang Guo, Ruotong Pan, Hongyu Lin, Le Sun, Xianpei Han

专题命中 记忆与上下文管理 :AI agent(abstract);分类 cs.CL

Comments Accepted by Information Processing and Management (IP&M)

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.07863 2024-01-22 cs.AI 57%

Synapse: Trajectory-as-Exemplar Prompting with Memory for Computer Control

Longtao Zheng, Rundong Wang, Xinrun Wang, Bo An

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments ICLR 2024. Project page: https://ltzheng.github.io/Synapse

详情

展开后加载摘要…

URL PDF HTML 收藏
2208.11344 2024-01-22 cs.LG 57%

Time-to-Green predictions for fully-actuated signal control systems with supervised learning

Alexander Genser, Michail A. Makridis, Kaidi Yang, Lukas Ambühl, Monica Menendez, Anastasios Kouvelas

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.08081 2024-01-17 cs.LG cs.SI 57%

Predicting Next Useful Location With Context-Awareness: The State-Of-The-Art

Alireza Nezhadettehad, Arkady Zaslavsky, Rakib Abdur, Siraj Ahmed Shaikh, Seng W. Loke, Guang-Li Huang, Alireza Hassani

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.07298 2024-01-17 stat.ML cs.LG 57%

Efficient Frameworks for Generalized Low-Rank Matrix Bandit Problems

Yue Kang, Cho-Jui Hsieh, Thomas C. M. Lee

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments Revision of the paper accepted by NeurIPS 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.03581 2024-01-09 cs.HC cs.AI cs.RO 57%

Evaluating and Personalizing User-Perceived Quality of Text-to-Speech Voices for Delivering Mindfulness Meditation with Different Physical Embodiments

Zhonghao Shi, Han Chen, Anna-Maria Velentza, Siqi Liu, Nathaniel Dennler, Allison O'Connell, Maja Matarić

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Published in Proceedings of the 2023 ACM/IEEE International Conference on Human-Robot Interaction, pp. 516-524. 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.13863 2024-01-04 cs.LG cs.CR cs.RO 57%

Manipulating Trajectory Prediction with Backdoors

Kaouther Messaoud, Kathrin Grosse, Mickael Chen, Matthieu Cord, Patrick Pérez, Alexandre Alahi

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 9 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.09595 2024-01-04 cs.AI 57%

Minimum Coverage Sets for Training Robust Ad Hoc Teamwork Agents

Arrasy Rahman, Jiaxun Cui, Peter Stone

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments Accepted at AAAI-24 conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.01165 2024-01-03 cs.LG eess.SP 57%

Reinforcement Learning for SAR View Angle Inversion with Differentiable SAR Renderer

Yanni Wang, Hecheng Jia, Shilei Fu, Huiping Lin, Feng Xu

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.15523 2023-12-27 cs.CY cs.CL cs.HC physics.soc-ph 57%

The Persuasive Power of Large Language Models

Simon Martin Breum, Daniel Vædele Egdal, Victor Gram Mortensen, Anders Giovanni Møller, Luca Maria Aiello

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments 9 pages, 6 figures, 3 tables, 1 page in appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.08638 2023-12-27 cs.SE 57%

Building Domain-Specific Machine Learning Workflows: A Conceptual Framework for the State-of-the-Practice

Bentley James Oakes, Michalis Famelis, Houari Sahraoui

专题命中 记忆与上下文管理 :workflow(abstract);分类 cs.SE

Comments 33 pages 14 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.12461 2023-12-21 cs.LG 57%

Bird Movement Prediction Using Long Short-Term Memory Networks to Prevent Bird Strikes with Low Altitude Aircraft

Elaheh Sabziyan Varnousfaderani, Syed A. M. Shihab

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

Comments 84 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.11663 2023-12-20 cs.LG stat.ML 57%

Eliciting Kemeny Rankings

Anne-Marie George, Christos Dimitrakakis

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments This is a long version of the AAAI'24 publication under the same title

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.08602 2023-12-15 cs.LO cs.LG 57%

Omega-Regular Decision Processes

Ernst Moritz Hahn, Mateo Perez, Sven Schewe, Fabio Somenzi, Ashutosh Trivedi, Dominik Wojtczak

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏