arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 5047 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 工具调用 5047 篇

2603.05413 2026-03-18 cs.SD 67%

Building Enterprise Realtime Voice Agents from Scratch: A Technical Tutorial

从零构建企业级实时语音代理:技术教程

Jielin Qiu, Zixiang Chen, Liangwei Yang, Ming Zhu, Zhiwei Liu, Juntao Tan, Wenting Zhao, Rithesh Murthy, Roshan Ram, Akshara Prabhakar, Shelby Heinecke, Caiming Xiong, Silvio Savarese, Huan Wang

机构 * Salesforce AI Research(Salesforce AI研究院)

专题命中 工具调用 :agent(abstract);function calling(abstract)

AI总结 本文介绍如何从零构建企业级实时语音代理,通过Deepgram、vLLM和ElevenLabs实现端到端流程,展示最佳实践与代码实现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.15653 2026-03-18 cs.CL cs.AI cs.LG 67%

Recursive Language Models Meet Uncertainty: The Surprising Effectiveness of Self-Reflective Program Search for Long Context

递归语言模型与不确定性:自我反思程序搜索在长上下文中的意外效果

Keivan Alizadeh, Parshin Shojaee, Minsik Cho, Mehrdad Farajtabar

机构 * Apple(苹果公司)

专题命中 工具调用 :agentic(abstract);分类 cs.AI、cs.CL、cs.LG

AI总结 本文提出SRLM框架,通过引入不确定性感知的自我反思,提升长上下文处理能力,实验显示其在多种任务中优于现有基线,且无需递归机制。

Comments preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
1910.02239 2026-03-16 cs.DC 67%

The Role of A-priori Information in Networks of Rational Agents

在理性代理网络中先验信息的作用

Yehuda Afek, Yishay Mansour, Shaked Rafaeli, Moshe Sulamy

专题命中 工具调用 :agent(abstract);tool use(abstract)

AI总结 本文研究了在理性代理网络中,先验信息对均衡的影响,通过分析复制行为和先验分布,得出了在不同分布式计算问题中达到均衡所需的先验知识界限。

Comments This paper is the full version of the DISC 2018 paper. arXiv admin note: substantial text overlap with arXiv:1711.04728

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.23109 2026-03-02 cs.RO 67%

Towards Intelligible Human-Robot Interaction: An Active Inference Approach to Occluded Pedestrian Scenarios

迈向可解释的人机交互:一种主动推断方法用于遮挡行人场景

Kai Chen, Yuyao Huang, Guang Chen

机构 * Tongji University(同济大学)

专题命中 工具调用 :agent(abstract);planning(abstract)

AI总结 本文提出基于主动推断的方法,用于解决遮挡行人场景中的安全挑战,通过结合 RBPF 和 CEM 增强的 MPPI 控制器,实现可解释的人机交互。

Comments 14 pages, 6 figures, Proceedings of the 2026 ACM/IEEE International Conference on Human-Robot Interaction (HRI'26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.09496 2026-02-11 cs.HC 67%

Jokeasy: Exploring Human-AI Collaboration in Thematic Joke Generation

Jokeasy:探索主题笑话生成中的人机协作

Yate Ge, Lin Tian, Chiqian Xu, Luyao Xu, Meiying Li, Yuanda Hu, Weiwei Guo

专题命中 工具调用 :agent(abstract);workflow(abstract)

AI总结 Jokeasy通过人机协作提升主题笑话生成,结合搜索功能与双角色LLM代理,优化创意流程与素材整合。

Comments Accepted at IASDR 2025. This is the author-accepted version. Correspondence to first author: geyate@gmail.com

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01129 2026-02-03 cs.CR 67%

SMCP: Secure Model Context Protocol

SMCP: 安全模型上下文协议

Xinyi Hou, Shenao Wang, Yifan Zhang, Ziluo Xue, Yanjie Zhao, Cai Fu, Haoyu Wang

专题命中 工具调用 :workflow(abstract);agentic(abstract)

AI总结 SMCP通过统一身份管理、强认证和细粒度策略执行,提升智能体系统在工具调用中的安全性和可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01107 2026-02-03 cs.SE cs.AI cs.LG 67%

SPELL: Synthesis of Programmatic Edits using LLMs

通过LLM合成程序性编辑:SPELL

Daniel Ramos, Catarina Gamboa, Inês Lynce, Vasco Manquinho, Ruben Martins, Claire Le Goues

机构 * Carnegie Mellon University(卡内基梅隆大学) Carnegie Mellon University USA(卡内基梅隆大学(美国)) INESC-ID / IST - Universidade de Lisboa(INESC-ID / IST - 莱里斯本大学)

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG、cs.SE

AI总结 SPELL通过LLM提取迁移示例并泛化为可重用的转换脚本,实现自动化API迁移。

Comments pre-print

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12272 2026-01-21 cs.CV 67%

AgenticPruner: MAC-Constrained Neural Network Compression via LLM-Driven Strategy Search

AgenticPruner: 通过LLM驱动的策略搜索实现MAC约束的神经网络压缩

Shahrzad Esmat, Mahdi Banisharif, Ali Jannesari

机构 * Iowa State University(爱荷华州立大学)

专题命中 工具调用 :agent(abstract);workflow(abstract)

AI总结 AgenticPruner通过LLM驱动策略搜索实现MAC约束的神经网络压缩,通过三个专门代理协调优化,提升收敛成功率并实现精确的MAC预算控制。

Comments 38 pages, 2 figures, 14 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09937 2026-01-16 cs.HC cs.IR 67%

From SERPs to Agents: A Platform for Comparative Studies of Information Interaction

从搜索结果页面到代理:一个用于信息交互比较研究的平台

Saber Zerhoudi, Michael Granitzer

专题命中 工具调用 :agent(abstract);autonomous agent(abstract)

AI总结 UXLab是一个开源平台,用于比较信息交互系统,通过无代码实验设计支持多模态交互研究。

Journal ref Proceedings of the 2026 ACM SIGIR Conference on Human Information Interaction and Retrieval (CHIIR '26)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.05887 2026-01-12 cs.CR 67%

Cybersecurity AI: A Game-Theoretic AI for Guiding Attack and Defense

网络空间AI:一种基于博弈论的AI用于引导攻击与防御

Víctor Mayoral-Vilches, María Sanz-Gómez, Francesco Balassone, Stefan Rass, Lidia Salas-Espejo, Benjamin Jablonski, Luis Javier Navarrete-Lozano, Maite del Mundo de Torres, Cristóbal R. J. Veas Chavez

专题命中 工具调用 :agent(abstract);agentic(abstract)

AI总结 本文提出G-CTR,一种基于博弈论的AI指导层,通过生成摘要引导攻击与防御行为,提升网络安全测试效率与成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03733 2026-01-08 cs.CV cs.AI cs.CL cs.CY cs.LG 67%

RadDiff: Describing Differences in Radiology Image Sets with Natural Language

RadDiff:用自然语言描述放射学图像集的差异

Xiaoxian Shen, Yuhui Zhang, Sahithi Ankireddy, Xiaohan Wang, Maya Varma, Henry Guo, Curtis Langlotz, Serena Yeung-Levy

机构 * Stanford University(斯坦福大学)

专题命中 工具调用 :agentic(abstract);分类 cs.AI、cs.CL、cs.LG

AI总结 RadDiff通过多模态代理系统实现放射学图像集差异的自然语言描述,结合医学知识和多模态推理,在放射学研究配对中取得高准确率,推动临床影像分析的发展。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10581 2025-12-23 cs.GR 67%

GraphTracer: Graph-Guided Failure Tracing in LLM Agents for Robust Multi-Turn Deep Search

GraphTracer: LLM代理中基于图的故障追踪以实现鲁棒多轮深度搜索

Heng Zhang, Yuling Shi, Xiaodong Gu, Haochen You, Zijian Zhang, Lubin Gan, Yilei Yuan, Jin Huang

专题命中 工具调用 :agent(abstract);multi-agent(abstract)

AI总结 GraphTracer通过信息流分析和依赖图构建,提升多代理系统在多轮深度搜索中的故障归因准确性与鲁棒性。

Comments This submission has been withdrawn by the authors due to a fundamental error in the methodology that affects the validity of the main results

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09481 2025-12-03 cs.SE cs.AI cs.CL 67%

Evaluating LLMs on Sequential API Call Through Automated Test Generation

通过自动化测试生成评估LLMs的顺序API调用

Yuheng Huang, Jiayang Song, Da Song, Zhenlan Ji, Wenhan Wang, Shuai Wang, Lei Ma

机构 * The University of Tokyo(东京大学) Macau University of Science and Technology(澳门科技大学) Shandong University(山东大学) Hong Kong University of Science and Technology(香港科技大学) Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所) University of Alberta(阿尔伯塔大学)

专题命中 工具调用 :tool use(abstract);分类 cs.AI、cs.CL、cs.SE

AI总结 本文提出StateGen框架,通过自动化测试生成评估LLMs在顺序API调用中的性能,构建了包含120个测试用例的StateEval基准测试,揭示了当前LLM在API整合方面的改进方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.18335 2025-11-25 cs.CL cs.AI cs.LG 67%

OmniStruct: Universal Text-to-Structure Generation across Diverse Schemas

OmniStruct: 跨多样的模式生成的通用文本到结构生成

James Y. Huang, Wenxuan Zhou, Nan Xu, Fei Wang, Qin Liu, Sheng Zhang, Hoifung Poon, Muhao Chen

机构 * University of Southern California(南加州大学) University of California, Davis(加州大学戴维斯分校) Microsoft Research(微软研究院)

专题命中 工具调用 :function calling(abstract);分类 cs.AI、cs.CL、cs.LG

AI总结 OmniStruct提出了一种跨多种模式的通用文本到结构生成方法,通过合成数据训练小型模型,实现与GPT-4o相当的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13428 2025-11-24 cs.RO 67%

VLM-SFD: VLM-Assisted Siamese Flow Diffusion Framework for Dual-Arm Cooperative Manipulation

VLM-SFD:基于视觉语言模型的双臂协作操作Siamese流扩散框架

Jiaming Chen, Yiyu Jiang, Aoshen Huang, Yang Li, Wei Pan

机构 * Department of Computer Science, The University of Manchester(计算机科学系,曼彻斯特大学) School of Control Science and Engineering, Shandong University(控制科学与工程学院,山东大学)

专题命中 工具调用 :tool use(abstract);planning(abstract)

AI总结 VLM-SFD通过双编码器-解码器架构和视觉语言模型,提升双臂协作操作的模仿学习效率与泛化能力。

Comments Accepted by IEEE RA-L

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09148 2025-11-19 cs.CL cs.AI cs.LG 67%

LoopTool: Closing the Data-Training Loop for Robust LLM Tool Calls

Kangning Zhang, Wenxiang Jiao, Kounianhua Du, Yuan Lu, Weiwen Liu, Weinan Zhang, Yong Yu

机构 * Shanghai Jiao Tong University(上海交通大学) Xiaohongshu Inc.(小红书公司)

专题命中 工具调用 :tool-use(abstract);分类 cs.AI、cs.CL、cs.LG

Comments The code is accessible at https://github.com/Rednote-DeepExperience/LoopTool. The LoopTool-8B is accessible at https://huggingface.co/zhuiguang-ning/LoopTool-8B

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11235 2025-11-04 cs.RO math.DS math.OC 67%

Ergodic exploration of dynamic distribution

Luka Lanča, Karlo Jakac, Sylvain Calinon, Stefan Ivić

机构 * Faculty of Engineering, University of Rijeka(里耶卡大学工程学院) Idiap Research Institute(伊迪亚普研究 institute)

专题命中 工具调用 :agent(abstract);multi-agent(abstract)

Comments Initial version

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00086 2025-11-04 cs.LG cs.AI cs.CL 67%

Generalizing Test-time Compute-optimal Scaling as an Optimizable Graph

Fali Wang, Jihai Chen, Shuhua Yang, Runxue Bao, Tianxiang Zhao, Zhiwei Zhang, Xianfeng Tang, Hui Liu, Qi He, Suhang Wang

机构 * The Pennsylvania State University(宾夕法尼亚州立大学) University of Pittsburgh(匹兹堡大学) Amazon(亚马逊) Microsoft(微软)

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.CL、cs.LG

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.02878 2025-10-31 cs.LG cs.AI cs.CL 67%

Language Models can Self-Improve at State-Value Estimation for Better Search

Ethan Mendes, Alan Ritter

机构 * Georgia Institute of Technology(佐治亚理工学院)

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.19405 2025-10-23 cs.CY 67%

Designing Knowledge Tools: How Students Transition from Using to Creating Generative AI in STEAM classroom

Qian Huang, Nachamma Sockalingam, Thijs Willems, King Wang Poon

专题命中 工具调用 :tool use(abstract);planning(abstract)

Comments to be published in IEEE TALE 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16944 2025-10-21 cs.CY cs.SC 67%

Learning Ecology with VERA Using Conceptual Models and Simulations

Spencer Rugaber, Scott Bunin, Andrew Hornback, Sungeun An, Ashok Goel

专题命中 工具调用 :agent(abstract);tool use(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.00784 2025-10-20 cs.IR cs.AI cs.CL cs.LG 67%

FIRE: Fact-checking with Iterative Retrieval and Verification

Zhuohan Xie, Rui Xing, Yuxia Wang, Jiahui Geng, Hasan Iqbal, Dhruv Sahnan, Iryna Gurevych, Preslav Nakov

机构 * MBZUAI The University of Melbourne(墨尔本大学)

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.CL、cs.LG

Comments 4 figures, 8 tables, accepted to Findings of NAACL

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14688 2025-10-01 cs.CL cs.AI cs.LG 67%

Mind the Gap: A Review of Arabic Post-Training Datasets and Their Limitations

Mohammed Alkhowaiter, Norah Alshahrani, Saied Alshahrani, Reem I. Masoud, Alaa Alzahrani, Deema Alnuhait, Emad A. Alghamdi, Khalid Almubarak

机构 * Refine AI ASAS AI University of Bisha(比沙大学) University College London(伦敦大学学院) King Salman Global Academy for Arabic(萨勒曼全球阿拉伯学院) University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) King Abdulaziz University(阿卜杜勒阿齐兹大学) HUMAIN

专题命中 工具调用 :function calling(abstract);分类 cs.AI、cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.09864 2025-09-15 cs.LG cs.AI cs.CL 67%

Latency and Token-Aware Test-Time Compute

Jenny Y. Huang, Mehul Damani, Yousef El-Kurdi, Ramon Astudillo, Wei Sun

机构 * Department of Electrical Engineering and Computer Science, MIT(麻省理工学院电气工程与计算机科学系) IBM Research(IBM研究院) MIT-IBM Watson AI Lab(麻省理工-IBM沃森人工智能实验室)

专题命中 工具调用 :agentic(abstract);分类 cs.AI、cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.07744 2025-08-29 cs.MA 67%

CBS with Continuous-Time Revisit

Andy Li, Zhe Chen, Danial Harabor, Mor Vered

专题命中 工具调用 :agent(abstract);multi-agent(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15595 2025-08-22 cs.NI 67%

Interface on demand: Towards AI native Control interfaces for 6G

Abhishek Dandekar, Prashiddha D. Thapa, Ashrafur Rahman, Julius Schulz-Zander

专题命中 工具调用 :agent(abstract);multi-agent(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03487 2025-08-06 cs.SE cs.AI cs.LG 67%

BitsAI-Fix: LLM-Driven Approach for Automated Lint Error Resolution in Practice

Yuanpeng Li, Qi Long, Zhiyuan Yao, Jian Xu, Lintao Xie, Xu He, Lu Geng, Xin Han, Yueyan Chen, Wenbo Duan

机构 * ByteDance(字节跳动) Carnegie Mellon University(卡内基梅隆大学) Zhejiang University(浙江大学)

专题命中 工具调用 :workflow(abstract);分类 cs.AI、cs.LG、cs.SE

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12621 2025-07-21 cs.HC cs.GR cs.MA 67%

NLI4VolVis: Natural Language Interaction for Volume Visualization via LLM Multi-Agents and Editable 3D Gaussian Splatting

Kuangshi Ai, Kaiyuan Tang, Chaoli Wang

专题命中 工具调用 :agent(abstract);multi-agent(abstract)

Comments IEEE VIS 2025. Project Page: https://nli4volvis.github.io/

Journal ref IEEE Transactions on Visualization and Computer Graphics (TVCG), vol. 32, no. 1, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06346 2025-07-10 cs.RO stat.CO 67%

Solving the Constrained Random Disambiguation Path Problem via Lagrangian Relaxation and Graph Reduction

Li Zhou, Elvan Ceyhan

机构 * Department of Mathematics and Statistics, Auburn University(数学与统计学系,阿伯丁大学)

专题命中 工具调用 :agent(abstract);planning(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14763 2025-06-18 cs.RO 67%

RobotSmith: Generative Robotic Tool Design for Acquisition of Complex Manipulation Skills

Chunru Lin, Haotian Yuan, Yian Wang, Xiaowen Qiu, Tsun-Hsuan Wang, Minghao Guo, Bohan Wang, Yashraj Narang, Dieter Fox, Chuang Gan

机构 * University of Massachusetts Amherst(马萨诸塞大学阿默斯特分校) Massachusetts Institute of Technology(麻省理工学院) National University of Singapore(新加坡国立大学) NVIDIA(英伟达) MIT-IBM Watson AI Lab(麻省理工-IBM沃森人工智能实验室)

专题命中 工具调用 :tool use(abstract);tool-use(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏