arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 4758 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 记忆与上下文管理 4758 篇

2507.06221 2025-07-09 cs.AI cs.GT 57%

Aligned Textual Scoring Rules

Yuxuan Lu, Yifan Wu, Jason Hartline, Michael J. Curry

机构 * Peking University(北京大学) Northwestern University(西北大学) University of Illinois Chicago(伊利诺伊大学香槟分校)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.11150 2025-07-09 cs.CV cs.LG 57%

GC-GAT: Multimodal Vehicular Trajectory Prediction using Graph Goal Conditioning and Cross-context Attention

Mahir Gulzar, Yar Muhammad, Naveed Muhammad

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03411 2025-07-08 cs.LG cs.GT eess.SP 57%

A Hybrid Game-Theory and Deep Learning Framework for Predicting Tourist Arrivals via Big Data Analytics and Opinion Leader Detection

Ali Nikseresht

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04682 2025-07-08 cs.CE cs.LG 57%

Operator-based machine learning framework for generalizable prediction of unsteady treatment dynamics in stormwater infrastructure

Mohamed Shatarah, Kai Liu, Haochen Li

机构 * Water Infrastructure Laboratory, Department of Civil and Environmental Engineering, University of Tennessee, Knoxville, Tennessee 37996, USA(水基础设施实验室,土木与环境工程系,田纳西大学, Knoxville,田纳西州 37996,美国)

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

Comments 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.18186 2025-07-08 stat.ML cs.LG 57%

Model-free Posterior Sampling via Learning Rate Randomization

Daniil Tiapkin, Denis Belomestny, Daniele Calandriello, Eric Moulines, Remi Munos, Alexey Naumov, Pierre Perrault, Michal Valko, Pierre Menard

机构 * CMAP, École Polytechnique(École Polytechnique 的 CMAP) HSE University(俄罗斯高等经济大学) Duisburg-Essen University(杜伊斯堡- Essen 大学) Google DeepMind(谷歌DeepMind) Mohamed Bin Zayed University of AI, UAE(阿联酋人工智能大学) IDEMIA ENS Lyon(里昂高等师范学院)

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.LG

Comments This revision fixed an error connected to an incorrect use of Proposition 7 inside of Lemma 4, and a misprint in Lemma 12. In the current version, we modified the martingale construction and applied the same argument as before; no results need to be modified as a result of these fixes

Journal ref Advances in Neural Information Processing Systems 36 (NeurIPS 2023)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06495 2025-07-08 cs.LG 57%

Boosting Robustness in Preference-Based Reinforcement Learning with Dynamic Sparsity

Calarina Muslimani, Bram Grooten, Deepak Ranganatha Sastry Mamillapalli, Mykola Pechenizkiy, Decebal Constantin Mocanu, Matthew E. Taylor

机构 * University of Alberta(阿尔伯塔大学) Eindhoven University of Technology(埃因霍温理工大学) University of Luxembourg(卢森堡大学) Alberta Machine Intelligence Institute(阿尔伯塔人工智能研究所)

专题命中 记忆与上下文管理 :autonomous agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02712 2025-07-04 cs.LG 57%

A Forget-and-Grow Strategy for Deep Reinforcement Learning Scaling in Continuous Control

Zilin Kang, Chenyuan Hu, Yu Luo, Zhecheng Yuan, Ruijie Zheng, Huazhe Xu

机构 * Computer Science, University of Maryland(大学计算机科学系,马里兰大学) Shanghai Qi Zhi Institute(上海启智研究院) Department of Computer Science(计算机科学系) Technology, Tsinghua University(技术,清华大学) Institute for Interdisciplinary Information Sciences, Tsinghua University(交叉信息科学研究所,清华大学) Huawei Noah's Ark Lab(华为诺亚实验室) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.24019 2025-07-01 cs.CV cs.CL 57%

Ella: Embodied Social Agents with Lifelong Memory

Hongxin Zhang, Zheyuan Zhang, Zeyuan Wang, Zunzhe Zhang, Lixing Fang, Qinhong Zhou, Chuang Gan

机构 * University of Massachusetts Amherst(马萨诸塞大学阿姆赫斯特分校) Johns Hopkins University(约翰霍普金斯大学) Tsinghua University(清华大学)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21727 2025-06-30 cs.GT cs.AI 57%

Simultaneously Fair Allocation of Indivisible Items Across Multiple Dimensions

Yasushi Kawase, Bodhayan Roy, Mohammad Azharuddin Sanpui

机构 * The University of Tokyo(东京大学) Indian Institute of Technology Kharagpur(印度理工学院哈里科格)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19785 2025-06-25 cs.AI 57%

Learning Task Belief Similarity with Latent Dynamics for Meta-Reinforcement Learning

Menglong Zhang, Fuyuan Qian

机构 * Southern University of Science and Technology(南方科技大学)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments ICLR2025 https://openreview.net/forum?id=5YbuOTUFQ4

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.01336 2025-06-25 cs.LG 57%

Enhancing Diversity in Parallel Agents: A Maximum State Entropy Exploration Story

Vincenzo De Paola, Riccardo Zamboni, Mirco Mutti, Marcello Restelli

机构 * AIRLAB, Politecnico di Milano, Milan, Italy(AIRLAB,米兰理工学院,米兰,意大利)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.01702 2025-06-23 cs.LG 57%

Al-Khwarizmi: Discovering Physical Laws with Foundation Models

Christopher E. Mower, Haitham Bou-Ammar

机构 * Huawei's Noahs Ark Lab, London, UK.(华为诺亚实验室,伦敦,英国)

专题命中 记忆与上下文管理 :agentic(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13366 2025-06-19 cs.CL 57%

Enhancing Goal-oriented Proactive Dialogue Systems via Consistency Reflection and Correction

Didi Zhang, Yaxin Fan, Peifeng Li, Qiaoming Zhu

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.CL

Comments Accepted by ACL'25 (main conference)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.02043 2025-06-19 cs.LG 57%

Constrained Linear Thompson Sampling

Aditya Gangrade, Venkatesh Saligrama

机构 * Boston University(波士顿大学)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13825 2025-06-18 cs.AI 57%

The Reflexive Integrated Information Unit: A Differentiable Primitive for Artificial Consciousness

Gnankan Landry Regis N'guessan, Issa Karambal

机构 * Research Group(研究组) Department of Applied Mathematics and Computational Science(应用数学与计算科学系) The Nelson Mandela African Institution of Science and Technology (NM-AIST)(纳尔逊·曼德拉非洲科学技术研究所) African Institute for Mathematical Sciences (AIMS)(非洲数学科学研究所)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.01876 2025-06-18 cs.LG 57%

Reinforcement Learning with Segment Feedback

Yihan Du, Anna Winnicki, Gal Dalal, Shie Mannor, R. Srikant

机构 * Stanford University(斯坦福大学) NVIDIA Research(NVIDIA研究)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13246 2025-06-17 cs.CR cs.AI cs.DC 57%

On Immutable Memory Systems for Artificial Agents: A Blockchain-Indexed Automata-Theoretic Framework Using ECDH-Keyed Merkle Chains

Craig Steven Wright

机构 * Dr Craig S. Wright University of Exeter Business School(埃克塞特大学商学院)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments 47 pages, includes formal automata specifications, cryptographic constructions, and epistemic architecture schema

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.17657 2025-06-17 cs.CL 57%

ReflecTool: Towards Reflection-Aware Tool-Augmented Clinical Agents

Yusheng Liao, Shuyang Jiang, Yanfeng Wang, Yu Wang

机构 * Shanghai Jiao Tong University(上海交通大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Fudan University(复旦大学)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments ACL 2025 Main Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.10934 2025-06-13 cs.CL 57%

Dynamic Epistemic Friction in Dialogue

Timothy Obiso, Kenneth Lai, Abhijnan Nath, Nikhil Krishnaswamy, James Pustejovsky

机构 * Brandeis University(布兰迪斯大学) Colorado State University(科罗拉多州立大学)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments 11 pages, 2 figures, 2 tables, CoNLL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.08633 2025-06-11 eess.AS cs.CL 57%

Approaching Dialogue State Tracking via Aligning Speech Encoders and LLMs

Šimon Sedláček, Bolaji Yusuf, Ján Švec, Pradyoth Hegde, Santosh Kesiraju, Oldřich Plchot, Jan Černocký

机构 * Brno University of Technology(布拉格技术大学) Indian Institute of Information Technology Dharwad(德瓦达信息科技学院)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.CL

Comments Accepted to Interspeech 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.16270 2025-06-05 cs.AI 57%

Reflection-Bench: Evaluating Epistemic Agency in Large Language Models

Lingyu Li, Yixu Wang, Haiquan Zhao, Shuqi Kong, Yan Teng, Chunbo Li, Yingchun Wang

专题命中 记忆与上下文管理 :AI agent(abstract);分类 cs.AI

Comments 29 pages, 19 figures, 9 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02158 2025-06-04 cs.AI 57%

Reflection-Based Memory For Web navigation Agents

Ruhana Azam, Aditya Vempaty, Ashish Jagmohan

机构 * UIUC(伊利诺伊大学) Emergence AI

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.19318 2025-06-04 cs.AI 57%

MINDSTORES: Memory-Informed Neural Decision Synthesis for Task-Oriented Reinforcement in Embodied Systems

Anirudh Chari, Suraj Reddy, Aditya Tiwari, Richard Lian, Brian Zhou

机构 * Massachusetts Institute of Technology(麻省理工学院) Illinois Mathematics and Science Academy(伊利诺伊数学与科学学院) Harvard University(哈佛大学)

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.13995 2025-06-04 cs.LG cs.CR 57%

Adversarial Inception Backdoor Attacks against Reinforcement Learning

Ethan Rathbun, Alina Oprea, Christopher Amato

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments 9 pages, 6 figures, ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.17034 2025-06-04 cs.LG 57%

Learning Actionable Counterfactual Explanations in Large State Spaces

Keziah Naggita, Matthew R. Walter, Avrim Blum

机构 * Toyota Technological Institute at Chicago(丰田技术研究所(芝加哥))

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.01312 2025-06-03 cs.CL 57%

Growing Through Experience: Scaling Episodic Grounding in Language Models

Chunhui Zhang, Sirui, Wang, Zhongyu Ouyang, Xiangchi Yuan, Soroush Vosoughi

机构 * Department of Computer Science, Dartmouth College(达特茅斯学院计算机科学系) School of Computer Science, Georgia Institute of Technology(佐治亚理工学院计算机科学学院)

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.CL

Comments Accepted at The 63rd Annual Meeting of the Association for Computational Linguistics (ACL 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.01174 2025-06-03 cs.AI 57%

GraphPad: Inference-Time 3D Scene Graph Updates for Embodied Question Answering

Muhammad Qasim Ali, Saeejith Nair, Alexander Wong, Yuchen Cui, Yuhao Chen

机构 * University of Waterloo(滑铁卢大学) University of California, Los Angeles(加州大学洛杉矶分校)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.AI

Comments CVPR 2025 Workshop on 3D-LLM/VLA: Bridging Language, Vision and Action in 3D Environments

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.00425 2025-06-02 cs.RO cs.AI 57%

ManiSkill3: GPU Parallelized Robotics Simulation and Rendering for Generalizable Embodied AI

Stone Tao, Fanbo Xiang, Arth Shukla, Yuzhe Qin, Xander Hinrichsen, Xiaodi Yuan, Chen Bao, Xinsong Lin, Yulin Liu, Tse-kai Chan, Yuan Gao, Xuanlin Li, Tongzhou Mu, Nan Xiao, Arnav Gurha, Viswesh Nagaswamy Rajesh, Yong Woo Choi, Yen-Ru Chen, Zhiao Huang, Roberto Calandra, Rui Chen, Shan Luo, Hao Su

机构 * University of California San Diego(加州大学圣地亚哥分校) Carnegie Mellon University(卡内基梅隆大学) Hillbot(Hillbot公司) TU Dresden(德累斯顿技术大学) Tsinghua University(清华大学) King’s College London(伦敦国王学院)

专题命中 记忆与上下文管理 :planning(abstract);分类 cs.AI

Comments Project website: http://maniskill.ai/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23150 2025-05-30 cs.LG 57%

Bigger, Regularized, Categorical: High-Capacity Value Functions are Efficient Multi-Task Learners

Michal Nauman, Marek Cygan, Carmelo Sferrazza, Aviral Kumar, Pieter Abbeel

机构 * UC, Berkeley(伯克利大学) University of Warsaw(华沙大学) Nomagic(Nomagic公司) CMU(卡内基梅隆大学)

专题命中 记忆与上下文管理 :agent(abstract);分类 cs.LG

Comments preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.20701 2025-05-29 cs.SE cs.HC 57%

System-driven Cloud Architecture Design Support with Structured State Management and Guided Decision Assistance

Ryosuke Kohita, Akira Kasuga

专题命中 记忆与上下文管理 :workflow(abstract);分类 cs.SE

详情

展开后加载摘要…

URL PDF HTML 收藏