arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

2026-03-12 至 2026-03-12 共收录 9 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 软件智能体 9 篇

2603.10808 2026-03-12 cs.AI cs.HC cs.SE 88%

Nurture-First Agent Development: Building Domain-Expert AI Agents Through Conversational Knowledge Crystallization

先培养型代理开发:通过对话知识结晶构建领域专家AI代理

Linghao Zhang

机构 * Jiangsu Key Laboratory of Wireless Communications(江苏无线通信重点实验室)

专题命中 软件智能体 :agent(title,abstract);AI agent(title,abstract);分类 cs.AI、cs.SE

AI总结 本文提出先培养开发范式,通过结构化对话与领域专家交互逐步构建领域专家AI代理,强调知识结晶循环和三层认知架构。

Comments 24 pages, 8 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05621 2026-03-12 cs.RO cs.AI cs.CL cs.LG cs.MA 82%

RACAS: Controlling Diverse Robots With a Single Agentic System

RACAS:通过单一代理系统控制多样化机器人

Dylan R. Ashley, Jan Przepióra, Yimeng Chen, Ali Abualsaud, Nurzhan Yesmagambet, Shinkyu Park, Eric Feron, Jürgen Schmidhuber

机构 * Center of Excellence in Generative AI, King Abdullah University of Science and Technology (KAUST), Saudi Arabia(沙特王国科学与技术大学生成人工智能卓越中心) Dalle Molle Institute for Artificial Intelligence Research (IDSIA), Switzerland(人工智能研究达勒莫利 institute) Università della Svizzera italiana (USI), Switzerland(瑞士意大利大学) Scuola universitaria professionale della Svizzera italiana (SUPSI), Switzerland(瑞士意大利专业大学) Robotics, Intelligent Systems, and Control Lab, King Abdullah University of Science and Technology (KAUST), Saudi Arabia(机器人、智能系统与控制实验室,沙特王国科学与技术大学(KAUST)) Department of Process Control, AGH University of Krakow, Poland(波兰克拉科夫AGH大学过程控制系)

专题命中 软件智能体 :agentic(title,abstract);分类 cs.AI、cs.CL、cs.LG

AI总结 RACAS通过单一代理系统实现跨平台机器人控制,无需修改代码或模型,有效降低机器人原型开发难度。

Comments 7 pages in main text + 1 page of appendices + 1 page of references, 5 figures in main text + 1 figure in appendices, 2 tables in main text; source code available at https://github.com/janprz11/robot-agnostic-control

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.10387 2026-03-12 cs.CR 67%

Don't Let the Claw Grip Your Hand: A Security Analysis and Defense Framework for OpenClaw

不要让爪子抓住你的手:OpenClaw平台的安全分析与防御框架

Zhengyang Shan, Jiayun Xin, Yue Zhang, Minghui Xu

专题命中 软件智能体 :agent(abstract);AI agent(abstract)

AI总结 本文分析了OpenClaw平台的安全问题,提出HITL防御机制,显著提升系统防御率至19%-92%。

Comments 12 pages, 2 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06739 2026-03-12 cs.SE cs.AI 62%

ResearchEnvBench: Benchmarking Agents on Environment Synthesis for Research Code Execution

ResearchEnvBench: 为研究代码执行环境合成的基准测试

Yubang Wang, Chenxi Zhang, Bowen Chen, Zezheng Huai, Zihao Dai, Xinchi Chen, Yuxin Wang, Yining Zheng, Jingjing Gong, Xipeng Qiu

机构 * Institute of Trustworthy Embodied AI(可信具身AI研究院) Shanghai Innovation Institution(上海创新机构) Shanghai Key Laboratory of Multimodal Embodied AI(上海多模态具身AI重点实验室) College of Computer Science and Artificial Intelligence(计算机科学与人工智能学院) OpenMOSS Wuhan University(武汉大学) Nanjing University(南京大学) Jilin University(吉林大学)

专题命中 软件智能体 :autonomous agent(abstract);分类 cs.AI、cs.SE

AI总结 ResearchEnvBench 是一个用于评估自主代理在研究代码执行中环境合成能力的基准测试,旨在解决当前代理在依赖解析和版本耦合方面的不足。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.10597 2026-03-12 cs.RO cs.AI 57%

Recover to Predict: Progressive Retrospective Learning for Variable-Length Trajectory Prediction

恢复以预测:渐进性回顾学习用于可变长度轨迹预测

Hao Zhou, Lu Qi, Jason Li, Jie Zhang, Yi Liu, Xu Yang, Mingyu Fan, Fei Luo

机构 * Great Bay University(大湾大学) Tsinghua SIGS(清华大学SIGS) Wuhan University(武汉大学) NTU(国立科技大学) Donghua University(东华大学) CASIA(中国科学院自动化研究所)

专题命中 软件智能体 :planning(abstract);分类 cs.AI

AI总结 本文提出渐进回顾框架和滚动起始训练策略,用于解决可变长度轨迹预测中的信息缺口问题,提升预测准确性。

Comments Paper is accepted by CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.10194 2026-03-12 cs.CR cs.AI 57%

MCP-in-SoS: Risk assessment framework for open-source MCP servers

MCP-in-SoS:开源MCP服务器的风险评估框架

Pratyay Kumar, Miguel Antonio Guirao Aguilera, Srikathyayani Srikanteswara, Satyajayant Misra, Abu Saleh Md Tayeen

机构 * New Mexico State University(新墨西哥州立大学) University of Hartford(哈特福德大学)

专题命中 软件智能体 :agent(abstract);分类 cs.AI

AI总结 本文提出了一种针对开源MCP服务器的风险评估框架,通过静态代码分析识别弱点并评估其安全风险,揭示了现有系统存在的安全隐患。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.10790 2026-03-12 astro-ph.IM 50%

Identifying and Measuring Satellite Streaks in DECam Images

识别和测量DECam图像中的卫星轨迹

Alexandra Serrano Mendoza, Meredith L. Rawls, Andrés Alejandro Plazas Malagón

专题命中 软件智能体 :workflow(abstract)

AI总结 本研究通过DECam图像开发工作流程,识别和测量卫星轨迹,揭示了不同轨道物体在亮度上的显著差异,为未来卫星对天文观测影响的研究奠定基础。

Comments Published in Revista eSpectra (Observatorio Astronómico Nacional de Colombia; https://drive.google.com/file/d/197nayTaqmTiJN0onE_tskd_5oKVXJGbA/view). Research conducted as part of the RECA Internship Program 2025 (https://www.astroreca.org/en/2025)

Journal ref Revista eSpectra, Vol. 4, Num. 1, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.19806 2026-03-12 physics.soc-ph cond-mat.stat-mech cs.SI 50%

Nonequilibrium phase transitions in a racism-spreading model with interaction-driven dynamics

非平衡相变在具有交互驱动动态的种族主义传播模型中

Nuno Crokidakis, Lucas Sigaud

专题命中 软件智能体 :agent(abstract)

AI总结 本文提出一个三状态模型研究种族主义内容在不同网络结构中的传播与抑制,揭示了网络拓扑对相变和吸收状态的影响。

Comments 22 pages, 7 figures, to appear in EPJB

Journal ref Eur. Phys. J. B 99, 34 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03824 2026-03-12 cs.GT cs.MA 50%

What Do Agents Think One Another Want? Level-2 Inverse Games for Inferring Agents' Estimates of Others' Objectives

智能体彼此想要什么?用于推断智能体对他人目标估计的二级逆向游戏

Hamzah I. Khan, Jingqi Li, David Fridovich-Keil

专题命中 软件智能体 :agent(abstract)

AI总结 本文提出二级逆向游戏框架,用于推断智能体对他人目标的估计,解决了传统一级推断在现实场景中的局限性。

Comments 6 pages + appendix with supplements

详情

展开后加载摘要…

URL PDF HTML 收藏