arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 5014 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 工具调用 5014 篇

2509.04642 2025-09-08 cs.AI cs.CL cs.LG cs.SE 83%

Maestro: Joint Graph & Config Optimization for Reliable AI Agents

Wenxiao Wang, Priyatham Kattakinda, Soheil Feizi

专题命中 工具调用 :AI agent(title);agent(abstract);分类 cs.AI、cs.CL、cs.LG

Comments Technical Report by RELAI.ai

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.11301 2025-08-18 cs.AI 83%

Learning to Be A Doctor: Searching for Effective Medical Agent Architectures

Yangyang Zhuang, Wenjia Jiang, Jiayu Zhang, Ze Yang, Joey Tianyi Zhou, Chi Zhang

机构 * AGI Lab, Westlake University(西拉雅大学AGI实验室) Henan University(河南大学) Affiliated Hospital of Xuzhou Medical University(徐州医科大学附属医院) Xuzhou Medical University(徐州医科大学) Nanyang Technological University(南洋理工大学) IHPC, Agency for Science, Technology and Research, Singapore(新加坡科技研究局IHPC) CFAR, Agency for Science, Technology and Research, Singapore(新加坡科技研究局CFAR)

专题命中 工具调用 :agent(title,abstract);workflow(abstract);分类 cs.AI

Comments Accepted at ACM MM 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02921 2025-08-06 cs.AI cs.CR 83%

PentestJudge: Judging Agent Behavior Against Operational Requirements

Shane Caldwell, Max Harley, Michael Kouremetis, Vincent Abruzzo, Will Pearce

机构 * dreadnode

专题命中 工具调用 :agent(title,abstract);tool-use(abstract);分类 cs.AI

Comments 18 pages, 5 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14246 2025-05-21 cs.CV cs.AI 83%

Visual Agentic Reinforcement Fine-Tuning

Ziyu Liu, Yuhang Zang, Yushan Zou, Zijian Liang, Xiaoyi Dong, Yuhang Cao, Haodong Duan, Dahua Lin, Jiaqi Wang

机构 * Shanghai Jiaotong University(上海交通大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) The Chinese University of Hong Kong(香港中文大学) Wuhan University(武汉大学)

专题命中 工具调用 :agentic(title,abstract);function calling(abstract);分类 cs.AI

Comments project url: https://github.com/Liuziyu77/Visual-RFT/tree/main/Visual-ARFT

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.02506 2025-05-21 cs.CL 83%

ToolHop: A Query-Driven Benchmark for Evaluating Large Language Models in Multi-Hop Tool Use

Junjie Ye, Zhengyin Du, Xuesong Yao, Weijian Lin, Yufei Xu, Zehui Chen, Zaiyuan Wang, Sining Zhu, Zhiheng Xi, Siyu Yuan, Tao Gui, Qi Zhang, Xuanjing Huang, Jiecao Chen

机构 * Fudan University(复旦大学) ByteDance(字节跳动) Shanghai Collaborative Innovation Center of Intelligent Visual Computing(上海智能视觉计算协同创新中心)

专题命中 工具调用 :tool use(title,abstract);tool-use(abstract);分类 cs.CL

Comments Accepted by ACL 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04354 2025-05-08 math.OC cs.AI 83%

Optimization Problem Solving Can Transition to Evolutionary Agentic Workflows

Wenhao Li, Bo Jin, Mingyi Hong, Changhong Lu, Xiangfeng Wang

机构 * Tongji University(同济大学) East China Normal University(华东师范大学) University of Minnesota(明尼苏达大学)

专题命中 工具调用 :agentic(title,abstract);workflow(abstract);分类 cs.AI

Comments 27 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.09713 2025-02-25 cs.IR cs.AI 83%

Agentic Information Retrieval

Weinan Zhang, Junwei Liao, Ning Li, Kounianhua Du, Jianghao Lin

专题命中 工具调用 :agentic(title,abstract);AI agent(abstract);分类 cs.AI

Comments 11 pages, perspective paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.17464 2025-01-09 cs.CL 83%

Efficient Tool Use with Chain-of-Abstraction Reasoning

Silin Gao, Jane Dwivedi-Yu, Ping Yu, Xiaoqing Ellen Tan, Ramakanth Pasunuru, Olga Golovneva, Koustuv Sinha, Asli Celikyilmaz, Antoine Bosselut, Tianlu Wang

专题命中 工具调用 :tool use(title,abstract);planning(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15399 2024-11-26 cs.PF cs.DC cs.LG 83%

Less is More: Optimizing Function Calling for LLM Execution on Edge Devices

Varatheepan Paramanayakam, Andreas Karatzas, Iraklis Anagnostopoulos, Dimitrios Stamoulis

专题命中 工具调用 :function calling(title,abstract);agentic(abstract);分类 cs.LG

Comments Accepted at DATE 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.18890 2024-10-25 cs.AI 83%

Improving Small-Scale Large Language Models Function Calling for Reasoning Tasks

Graziano A. Manduzio, Federico A. Galatolo, Mario G. C. A. Cimino, Enzo Pasquale Scilingo, Lorenzo Cominelli

专题命中 工具调用 :function calling(title,abstract);agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.17950 2024-10-24 cs.AI 83%

Benchmarking Floworks against OpenAI & Anthropic: A Novel Framework for Enhanced LLM Function Calling

Nirav Bhan, Shival Gupta, Sai Manaswini, Ritik Baba, Narun Yadav, Hillori Desai, Yash Choudhary, Aman Pawar, Sarthak Shrivastava, Sudipta Biswas

专题命中 工具调用 :function calling(title,abstract);tool use(abstract);分类 cs.AI

Comments 15 pages for main paper, 21 pages in total including references and appendix, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.13050 2024-09-24 cs.HC cs.AI 83%

Human-Centered LLM-Agent User Interface: A Position Paper

Daniel Chin, Yuxuan Wang, Gus Xia

专题命中 工具调用 :agent(title,abstract);workflow(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.09014 2023-03-17 cs.CL 83%

ART: Automatic multi-step reasoning and tool-use for large language models

Bhargavi Paranjape, Scott Lundberg, Sameer Singh, Hannaneh Hajishirzi, Luke Zettlemoyer, Marco Tulio Ribeiro

专题命中 工具调用 :tool-use(title,abstract);tool use(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.14998 2022-07-04 cs.RO cs.AI 83%

Understanding Physical Effects for Effective Tool-use

Zeyu Zhang, Ziyuan Jiao, Weiqi Wang, Yixin Zhu, Song-Chun Zhu, Hangxin Liu

专题命中 工具调用 :tool-use(title,abstract);planning(abstract);分类 cs.AI

Journal ref IEEE Robotics and Automation Letters, 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.03200 2020-02-04 cs.AI cs.RO 83%

Decentralized Cooperative Planning for Automated Vehicles with Continuous Monte Carlo Tree Search

Karl Kurzer, Florian Engelhorn, J. Marius Zöllner

专题命中 工具调用 :planning(title,abstract);agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1511.00787 2015-11-04 cs.AI 83%

A Pareto Optimal D* Search Algorithm for Multiobjective Path Planning

Alexander Lavin

专题命中 工具调用 :planning(title,abstract);agent(abstract);分类 cs.AI

Comments arXiv admin note: substantial text overlap with arXiv:1505.05947

详情

展开后加载摘要…

URL PDF HTML 收藏
1410.6519 2014-10-27 cs.AI 83%

Justifying and Improving Meta-Agent Conflict-Based Search

David Tolpin

专题命中 工具调用 :agent(title,abstract);multi-agent(abstract);分类 cs.AI

Comments 7 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.02464 2026-08-04 cs.AI cs.LG cs.SE 新提交 83%

Real-Time Detection and Repair of LLM Agent Failures

LLM智能体故障的实时检测与修复

Sunny Dubey

专题命中 工具调用 :agent(title,abstract);分类 cs.AI、cs.LG、cs.SE

AI总结 该研究提出基于LLM智能体步骤遥测的实时故障检测与修复系统,结合单类回声状态网络集成与确定性验证层,可高效检测并修复故障,提升任务成功率且成本极低。

Comments 16 pages, 5 figures. Code, data and demo: github.com/sunnydubey1111/agent-trajectory-sentinel Walkthrough: youtu.be/a05n_000klE

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.05915 2023-10-10 cs.CL cs.AI cs.LG 83%

FireAct: Toward Language Agent Fine-tuning

Baian Chen, Chang Shu, Ehsan Shareghi, Nigel Collier, Karthik Narasimhan, Shunyu Yao

专题命中 工具调用 :agent(title,abstract);分类 cs.AI、cs.CL、cs.LG

Comments Code, data, and models are available at https://fireact-agent.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.01236 2026-08-11 cs.CL cs.AI 版本更新 82%

Safeguarding LLM Agents from Misalignment through Provenance Analysis

通过溯源分析保护LLM智能体免受失调影响

Yining She, Yiliang Liang, Eunsuk Kang

机构 * Carnegie Mellon University(卡内基梅隆大学)

专题命中 工具调用 :agent(summary_cn,abstract);分类 cs.AI、cs.CL

AI总结 提出基于溯源分析的ProvenanceGuard框架,通过多阶段流水线检测工具调用前的三种失调类型,在Agent-SafetyBench和WorkBench上显著降低失调错误率并减少不必要的干预。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.05159 2026-08-05 cs.CR cs.AI cs.LG 版本更新 82%

Malice in Agentland: Down the Rabbit Hole of Backdoors in the AI Supply Chain

Agentland中的恶意行为:深入AI供应链后门问题

Léo Boisvert, Abhay Puri, Chandra Kiran Reddy Evuru, Nazanin Sepahvand, Nicolas Chapados, Quentin Cappart, Jason Stanley, Alexandre Lacoste, Krishnamurthy Dj Dvijotham, Alexandre Drouin

机构 * ServiceNow Research Mila - Qu\'ebec AI Institute

专题命中 工具调用 :agent(abstract);AI agent(abstract);tool use(abstract);agentic(abstract)

AI总结 研究探讨了在交互数据上微调AI代理时引入的安全漏洞,提出三种供应链威胁模型,证明通过污染少量演示即可使代理泄露用户信息。

Comments 27 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06716 2026-08-04 cs.AI cs.CL cs.CR 版本更新 82%

SIEVE: Selective Integrity Verification and Escalation for Defending LLM Agents against Indirect Prompt Injection

认知控制架构(CCA):一种用于鲁棒对齐AI代理的生命周期监督框架

Zhibo Liang, Tianze Hu, Zaiye Chen, Mingjie Tang

机构 * Sichuan University(四川大学)

专题命中 工具调用 :agent(abstract);tool-use(abstract);planning(abstract);agentic(abstract)

AI总结 本文提出认知控制架构(CCA),通过全生命周期认知监督框架,有效应对复杂IPI攻击,实现安全、功能与效率的平衡。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22431 2026-07-03 cs.AI cs.CL cs.FL 82%

Monadic Context Engineering

单子上下文工程

Yifan Zhang, Yang Yuan, Mengdi Wang, Andrew Chi-Chih Yao

机构 * IIIS, Tsinghua University(清华大学人工智能研究院)

专题命中 工具调用 :agent(abstract);AI agent(abstract);autonomous agent(abstract);tool use(abstract)

AI总结 本文提出Monadic Context Engineering,利用函子、应用函子和单子的代数结构,为智能体设计提供形式基础,通过分层方法构建高效可靠的AI智能体。

Comments We found some issues in the categorical foundations of this work, so we respectfully withdraw it

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.17946 2026-05-21 cs.AI cs.CV cs.LG 82%

SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain

SVFSearch: 一种面向游戏垂直领域的多模态知识密集型短视频帧搜索基准

Lingtao Mao, Huangyu Dai, Xinyu Sun, Zihan Liang, Ben Chen, Chenyi Lei, Wenwu Ou

机构 * Kuaishou Technology(快手科技)

专题命中 工具调用 :agent(abstract);tool-use(abstract);workflow(abstract);agentic(abstract)

AI总结 本文提出SVFSearch,首个针对中文游戏领域短视频帧搜索的多模态知识密集型基准,通过5000个四选一测试示例和4198个辅助训练示例,评估了从直接问答到计划-行动-重新计划代理等多种方法在短视频帧搜索中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24459 2025-10-29 cs.AI cs.MA cs.SE 82%

Affordance Representation and Recognition for Autonomous Agents

Habtom Kahsay Gidey, Niklas Huber, Alexander Lenz, Alois Knoll

机构 * Technische Universität München(慕尼黑技术大学) Jessy Works(杰西工作)

专题命中 工具调用 :autonomous agent(title);agent(abstract,journal_ref);分类 cs.AI、cs.SE;multi-agent(journal_ref)

Journal ref The Second International Workshop on Hypermedia Multi-Agent Systems (HyperAgents 2025), in conjunction with the 28th European Conference on Artificial Intelligence (ECAI 2025); October 26, 2025, Bologna, Italy

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13310 2025-09-17 cs.CL 82%

Scaling Agents via Continual Pre-training

Liangcai Su, Zhen Zhang, Guangyu Li, Zhuo Chen, Chenxi Wang, Maojia Song, Xinyu Wang, Kuan Li, Jialong Wu, Xuanzhong Chen, Zile Qiao, Zhongwang Zhang, Huifeng Yin, Shihao Cai, Runnan Fang, Zhengwei Tao, Wenbiao Yin, Chenxiong Qian, Yong Jiang, Pengjun Xie, Fei Huang, Jingren Zhou

专题命中 工具调用 :agent(abstract,comments);tool use(abstract);tool-use(abstract);agentic(abstract)

Comments https://tongyi-agent.github.io/blog/introducing-tongyi-deep-research/

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.05493 2026-08-07 cs.PL cs.AI cs.CL cs.SE 新提交 82%

Learning Context-Free Grammars for Grammar-Constrained Decoding via Declarative Agentic Programming with Guarantees

通过带保障的声明式智能体编程学习用于语法约束解码的上下文无关文法

Kevin Cheang, Geoff Hulette, Rahul Kumar, Felipe R. Monteiro, Federico Mora, Robin Salkeld, Lin Tan, Serdar Tasiran

专题命中 工具调用 :agentic(title);agent(abstract);分类 cs.AI、cs.CL、cs.SE

AI总结 本研究提出名为Autogrammar的声明式智能体,可从文档和执行数据自动学习上下文无关文法,用于语法约束解码,在三种DSLs上的实验显示其生成的文法性能优于现有基线,能显著提升端到端LM的真实任务表现。

Comments 9 pages, 3 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.02611 2026-08-05 cs.DC cs.AI cs.LG cs.SE 新提交 82%

KernelBrain: Coarse-to-Fine, Budget-Aware Search for Agentic GPU Kernel Optimization

KernelBrain:面向智能体的GPU内核优化的粗到细、预算感知搜索

Shuai Che, Gang Peng

专题命中 工具调用 :agentic(title);agent(abstract);分类 cs.AI、cs.LG、cs.SE

AI总结 KernelBrain是结合LLM引导变异等技术的GPU内核优化智能体,可提升内核质量与搜索效率,在Triton内核任务中获多倍加速且缩短优化时间。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.02217 2026-08-04 cs.CV 新提交 82%

VC-Tooler: Learning Compositional and Adaptive Visual Tool Use

VC-Tooler:学习组合式与自适应视觉工具使用

Yizheng Wu, Jiashen Hua, Bing Deng, Jieping Ye

专题命中 工具调用 :tool use(title,abstract);agentic(abstract)

AI总结 VC-Tooler是一种学习组合式与自适应视觉工具使用的模型,通过分层合成轨迹库与两阶段训练,在通用和具身基准上达到开源模型最优性能,且具良好迁移能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.23624 2026-07-30 cs.SE cs.AI cs.CL 版本更新 82%

Where Is the Cost of Third-Party API Routers in Agentic Software Development?

代理软件开发中第三方 API 路由器的成本在哪里?

Donghao Fu, Jingxin Li, Xue Jiang, Yihong Dong

专题命中 工具调用 :agentic(title);agent(abstract);分类 cs.AI、cs.CL、cs.SE

AI总结 研究第三方 API 路由器在代理软件开发中的影响,通过实证研究编码代理中路由器端注入的四个干预级别,开发 SIDEL 框架评估代理,发现路由器端干预难测,客户端缓解措施未完全恢复控制,强调需提供商端输出完整性保证。

详情

展开后加载摘要…

URL PDF HTML 收藏