arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 3200 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 软件智能体 3200 篇

2110.10973 2021-10-22 cs.AI cs.CL cs.LG cs.RO 67%

LOA: Logical Optimal Actions for Text-based Interaction Games

Daiki Kimura, Subhajit Chaudhury, Masaki Ono, Michiaki Tatsubori, Don Joven Agravante, Asim Munawar, Akifumi Wachi, Ryosuke Kohita, Alexander Gray

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.CL、cs.LG

Comments ACL-IJCNLP 2021 (demo paper)

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.08858 2021-10-12 cs.AI cs.CL cs.LG 67%

Grounding Spatio-Temporal Language with Transformers

Tristan Karch, Laetitia Teodorescu, Katja Hofmann, Clément Moulin-Frier, Pierre-Yves Oudeyer

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.CL、cs.LG

Comments Contains main article and supplementaries

Journal ref Neurips 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2006.10720 2021-09-24 cs.LG cs.AI cs.PL cs.SE stat.ML 67%

IReEn: Reverse-Engineering of Black-Box Functions via Iterative Neural Program Synthesis

Hossein Hajipour, Mateusz Malinowski, Mario Fritz

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.LG、cs.SE

Comments 15 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.05793 2021-08-16 cs.CV 67%

Progressive Coordinate Transforms for Monocular 3D Object Detection

Li Wang, Li Zhang, Yi Zhu, Zhi Zhang, Tong He, Mu Li, Xiangyang Xue

专题命中 软件智能体 :agent(abstract);AI agent(abstract)

Comments Code is available at: https://github.com/amazon-research/progressive-coordinate-transforms

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.00472 2021-06-02 cs.CV 67%

Detecting Anomalies in Semantic Segmentation with Prototypes

Dario Fontanel, Fabio Cermelli, Massimiliano Mancini, Barbara Caputo

专题命中 软件智能体 :agent(abstract);autonomous agent(abstract)

Comments SAIAD CVPR21 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2103.12134 2021-03-24 cs.IT math.IT 67%

A Joint Reinforcement-Learning Enabled Caching and Cross-Layer Network Code for Sum-Rate Maximization in F-RAN with D2D Communications

Mohammed S. Al-Abiad, Md. Zoheb Hassan, Md. Jahangir Hossain

专题命中 软件智能体 :agent(abstract);multi-agent(abstract)

Comments 15 pages, 9 figures, journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2010.12435 2020-10-26 cs.CV 67%

Pathological Visual Question Answering

Xuehai He, Zhuo Cai, Wenlan Wei, Yichen Zhang, Luntian Mou, Eric Xing, Pengtao Xie

专题命中 软件智能体 :agent(abstract);AI agent(abstract)

Comments arXiv admin note: text overlap with arXiv:2003.10286

详情

展开后加载摘要…

URL PDF HTML 收藏
2007.06708 2020-07-15 cs.CV cs.RO 67%

CPL-SLAM: Efficient and Certifiably Correct Planar Graph-Based SLAM Using the Complex Number Representation

Taosha Fan, Hanlin Wang, Michael Rubenstein, Todd Murphey

专题命中 软件智能体 :agent(abstract);autonomous agent(abstract)

Journal ref IEEE Transactions on Robotics, 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.08454 2018-12-27 cs.CL cs.AI cs.CV cs.LG 67%

Attention Based Natural Language Grounding by Navigating Virtual Environment

Akilesh B, Abhishek Sinha, Mausoom Sarkar, Balaji Krishnamurthy

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.CL、cs.LG

Comments Accepted at WACV 2019. Also at NeurIPS 2017 workshop on Visually-Grounded Interaction and Language (ViGIL)

详情

展开后加载摘要…

URL PDF HTML 收藏
1711.06351 2017-11-20 cs.CL cs.AI cs.LG 67%

Question Asking as Program Generation

Anselm Rothe, Brenden M. Lake, Todd M. Gureckis

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.CL、cs.LG

Comments Published in Advances in Neural Information Processing Systems (NIPS) 30, December 2017

Journal ref Rothe, A., Lake, B. M., and Gureckis, T. M. (2017). Question asking as program generation. Advances in Neural Information Processing Systems 30

详情

展开后加载摘要…

URL PDF HTML 收藏
1412.6958 2016-05-23 math.DS 67%

Global Stabilization of Triangulated Formations

Xudong Chen, M. -A. Belabbas, Tamer Başar

专题命中 软件智能体 :agent(abstract);autonomous agent(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
1403.5734 2015-12-08 cs.MA cs.CY 67%

Software Agents Interaction Algorithms in Virtual Learning Environment

Zahi A. M. Abu Sarhan

专题命中 软件智能体 :agent(abstract);multi-agent(abstract)

Journal ref The World of Computer Science and Information Technology Journal (WSCIT). 2014, Volume 4, Issue 2. pp. 18.25

详情

展开后加载摘要…

URL PDF HTML 收藏
1308.0315 2013-08-02 cs.MM cs.CV 67%

MAS for video objects segmentation and tracking based on active contours and SURF descriptor

Mohamed Chakroun, Ali Wali, Adel M. Alimi

专题命中 软件智能体 :agent(abstract);multi-agent(abstract)

Comments 6 pages

Journal ref IJCSI International Journal of Computer Science Issues, Vol. 10, Issue 2, No 3, March 2013

详情

展开后加载摘要…

URL PDF HTML 收藏
0806.0189 2009-12-01 cs.HC 67%

Investigating the use of Software Agents to Reduce The Risk of Undetected Errors in Strategic Spreadsheet Applications

Pat Cleary, Dr David Ball, Mukul Madahar, Simon Thorne, Christopher Gosling, Karen Fernandez

专题命中 软件智能体 :agent(abstract);planning(abstract)

Comments 12 Pages, 3 Tables, 3 Figures

Journal ref Proc. European Spreadsheet Risks Int. Grp. (EuSpRIG) 2003 147-159 ISBN 1 86166 199 1

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.21656 2026-07-27 cs.SE cs.AI 新提交 66%

Cross-Model LLM Code Review: Should you use Claude to review Codex or vice versa?

跨模型大语言模型代码审查:应该用Claude审查Codex还是相反?

Zuodong Xiang, Yike Zhang, YueMing Zhang, Hailu Xu

机构 * University of California, Davis(加州大学戴维斯分校) Johns Hopkins University(约翰霍普金斯大学) California State University, Long Beach(长滩加州州立大学)

专题命中 软件智能体 :workflow(abstract);分类 cs.AI、cs.SE;agentic(comments)

AI总结 研究开发者同时使用Claude和Codex进行代码审查的成本、时间及配对顺序问题,通过对116个任务的六种条件实验发现,Claude审查Codex草稿效果好,反向则不佳,有用的配对是不对称的,应Claude审查Codex。

Comments This paper had been accepted by Agentic SE @ KDD'26

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.18449 2026-05-19 cs.LG cs.AI 66%

Modelling Customer Trajectories with Reinforcement Learning for Practical Retail Insights

用强化学习建模客户轨迹以获得实际零售洞察

Ken Ming Lee, Paul Barde, Maxime C. Cohen, Derek Nowrouzezahrai

机构 * McGill University(麦吉尔大学) Mila - Quebec AI Institute(魁北克人工智能研究所)

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.LG;autonomous agent(comments)

AI总结 本文提出了一种基于智能体的建模框架,将客户轨迹预测转化为最大熵强化学习问题,以更准确地反映具有有限理性的客户行为,从而提供更精确的冲动购买率和货架交通密度估计。

Comments Proceeding of the 25th International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.03400 2026-04-07 cs.CV cs.AI cs.LG 66%

Banana100: Breaking NR-IQA Metrics by 100 Iterative Image Replications with Nano Banana Pro

Banana100: 通过100次迭代图像复制打破NR-IQA度量标准

Kenan Tang, Praveen Arunshankar, Andong Hua, Anthony Yang, Yao Qin

机构 * University of California, Santa Barbara(加州大学圣塔芭芭拉分校)

专题命中 软件智能体 :agentic(abstract,comments);分类 cs.AI、cs.LG

AI总结 Banana100通过100次迭代编辑生成28000张退化图像,揭示多轮编辑中图像质量退化问题,发现现有NR-IQA度量标准无法检测严重退化图像,威胁未来模型训练稳定性。

Comments Accepted to CVPR 2026 Workshop on Agentic AI for Visual Media

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.21354 2025-12-29 cs.CR cs.AI cs.SE 66%

Reflection-Driven Control for Trustworthy Code Agents

基于反射的可信代码代理控制

Bin Wang, Jiazheng Quan, Xingrui Yu, Hansen Hu, Yuhao, Ivor Tsang

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.SE;agentic(comments)

AI总结 本文提出反射驱动控制方法,通过内部反思循环提升代码生成的安全性和合规性,实现自主、安全且可审计的AI代码代理。

Comments Accepted to AAAI 2026 Workshop on Trust and Control in Agentic AI (TrustAgent)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14372 2025-07-22 cs.CL cs.AI cs.DB cs.HC 66%

Text-to-SQL for Enterprise Data Analytics

Albert Chen, Manas Bundele, Gaurav Ahlawat, Patrick Stetz, Zhitao Wang, Qiang Fei, Donghoon Jung, Audrey Chu, Bharadwaj Jayaraman, Ayushi Panth, Yatin Arora, Sourav Jain, Renjith Varma, Alexey Ilin, Iuliia Melnychuk, Chelsea Chueh, Joyan Sil, Xiaofeng Wang

机构 * LinkedIn(领英)

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.CL;agentic(comments)

Comments 11 pages, 8 figures, Workshop on Agentic AI for Enterprise at KDD '25

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08669 2025-06-18 cs.CL cs.AI 66%

SOPBench: Evaluating Language Agents at Following Standard Operating Procedures and Constraints

Zekun Li, Shinda Huang, Jiangtian Wang, Nathan Zhang, Antonis Antoniades, Wenyue Hua, Kaijie Zhu, Sirui Zeng, Chi Wang, William Yang Wang, Xifeng Yan

专题命中 软件智能体 :agent(abstract,comments);分类 cs.AI、cs.CL

Comments Code, data, and over 24k agent trajectories are released at https://github.com/Leezekun/SOPBench

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.02409 2019-03-07 cs.AI cs.LG 66%

A Grounded Interaction Protocol for Explainable Artificial Intelligence

Prashan Madumal, Tim Miller, Liz Sonenberg, Frank Vetere

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.LG;autonomous agent(comments)

Comments To appear in 18th International Conference on Autonomous Agents and Multiagent Systems (AAMAS 2019) as a full paper. arXiv admin note: substantial text overlap with arXiv:1806.08055

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06591 2025-12-09 cs.HC cs.AI 65%

Beyond Satisfaction: From Placebic to Actionable Explanations For Enhanced Understandability

超越满意:从安慰性到可操作性解释以提升可理解性

Joe Shymanski, Jacob Brue, Sandip Sen

机构 * The University of Tulsa(图兰大学)

专题命中 软件智能体 :agent(abstract,journal_ref);分类 cs.AI;multi-agent(journal_ref)

AI总结 本文探讨了可解释性在提升系统可理解性中的作用,通过实验发现可操作性解释在任务表现上优于安慰性解释,但用户满意度评分相同,强调需结合客观指标与主观评估来衡量解释质量。

Comments 21 pages, 7 figures, 6 tables. EXTRAAMAS 2025 submission. Preprint version

Journal ref In: Calvaresi, D., et al. Explainable, Trustworthy, and Responsible AI and Multi-Agent Systems. EXTRAAMAS 2025. Lecture Notes in Computer Science. Springer, Cham

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.10681 2025-05-19 cs.CY cs.AI cs.HC 65%

Towards an LLM-powered Social Digital Twinning Platform

Önder Gürcan, Vanja Falck, Markus G. Rousseau, Larissa L. Lima

机构 * Center for Modeling Social Systems(社会科学建模中心) NORCE Norwegian Research Center AS(挪威NORCE研究机构) Kristiansand, Norway(挪威克里斯蒂安桑)

专题命中 软件智能体 :agent(abstract,comments);分类 cs.AI;multi-agent(comments)

Comments 13 pages, 3 figures, 23rd International Conference on Practical applications of Agents and Multi-Agent Systems (PAAMS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.10319 2026-08-18 cs.SE cs.AI 版本更新 62%

Do Personalized Skills Help Coding Agents? An Empirical Study of Developer Interaction Histories

个性化技能对编码智能体有帮助吗?开发者交互历史的实证研究

Shuyan Huang, Kai Du, Andrew Lan

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.SE

AI总结 本研究通过对13位开发者的206个真实会话实验,发现从交互历史提炼的开发者个性化技能改进有限,而汇集自所有开发者的通用技能增益最大且最一致,为编码智能体的个性化策略提供了实证依据。

Comments 15 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.09018 2026-08-18 cs.NE cs.AI cs.LG 版本更新 62%

Evolving Ensemble of Agents

进化代理群体

Zongmin Yu, Liu Yang

机构 * National University of Singapore(新加坡国立大学)

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出EvE框架,通过进化编码代理群体实现算法发现,解决了传统方法的局限,展示了自修正群体在复杂代码库中的优势。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.13867 2026-08-17 cs.SE cs.AI 新提交 62%

Engineering Reliable Coding Agents: Evaluating and Operating the System Around the Model

构建可靠的编码智能体:评估与运行模型周边系统

Stephanie Jarmak

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.SE

AI总结 本研究针对AI编码智能体的可靠性问题,整合多源证据构建系统级评估与运行框架,区分模型与基础设施效应,提出可靠性记录目录及相关方法以提升智能体系统可靠性。

Comments Technical review and engineering monograph, 314 pages, 30 figures. Includes an evidence audit, a companion research artifact with 206 reliability records, and runnable protocols for evaluating and operating AI coding agents. August 2026. Source, companion, and reusable protocols: https://github.com/sjarmak/engineering-reliable-coding-agents

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.13681 2026-08-17 cs.SE cs.AI cs.ET 新提交 62%

Fine-Tuning Qwen3-27B for C-to-Rust Code Translation: A Three-Stage Curriculum of Pretraining, Debugging-Aware SFT, and Task-Specific SFT

微调Qwen3-27B实现C到Rust代码翻译:预训练、调试感知监督微调与任务特定监督微调的三阶段课程

Pu Zhao, Changdi Yang, Yixiao Chen, Yi Gao, Yifan Cao, Haochen Zeng, Yanzhi Wang

专题命中 软件智能体 :agentic(abstract);分类 cs.AI、cs.SE

AI总结 本研究针对C到Rust代码翻译任务,对Qwen3-27B采用三阶段微调课程,结合SACTOR框架评估后,其性能优于基线模型及其他LLM。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.13568 2026-08-17 cs.CL cs.AI 新提交 62%

Does a Language Server Save Tokens for Coding Agents? A Measurement Methodology and Preliminary Study

语言服务器能为编码智能体节省 token 吗?一种测量方法与初步研究

Pengcheng Xu

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.CL

AI总结 本文通过五组消融实验等方法,对比 LSP 与 grep 检索的 token 效率,发现 LSP 通常不节省 token,仅对最弱模型有 token 节省效果,需根据任务、模型等选择检索工具。

Comments 13 pages, 6 figures. Code and data: https://github.com/Poytr1/lsp-vs-grep-token-study

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.29516 2026-08-17 cs.SE cs.AI 版本更新 62%

From Code Review to Code Critique: Intent, Drift, and Spotlight for AI-Generated Diffs at Scale

从代码审查到代码批判:大规模AI生成代码差异的意图、漂移与焦点

Chandra Maddila, Mashrur Rashik, Euna Mehnaz Khan, Smriti Jha, James Saindon, Nachi Nagappan, Peter C. Rigby

专题命中 软件智能体 :agent(abstract);分类 cs.AI、cs.SE

AI总结 针对AI生成代码超出传统审查能力且现有工具忽视核心问题的缺陷,提出ARCTIC系统,经实验验证其在意图预测、漂移检测等方面表现优异,可降低代码不一致性并获高认可。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25692 2026-08-17 cs.SE cs.LG 版本更新 62%

A Configuration-First Framework for Reproducible, Low-Code Localization

一种面向可重复、低代码定位的配置优先框架

Tim Strnad, Blaž Bertalanič, Carolina Fortuna

机构 * Jožef Stefan Institute(乔塞夫·斯蒂芬研究所)

专题命中 软件智能体 :workflow(abstract);分类 cs.LG、cs.SE

AI总结 本文提出了一种低代码、配置优先的框架,通过版本化配置和自动化流程提升定位任务的可重复性和效率。

Comments 16 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏