arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

2026-03-06 至 2026-03-06 共收录 138 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 工具调用 5 篇

2510.00425 2026-03-06 cs.MA cs.RO 91%

Conflict-Based Search as a Protocol: A Multi-Agent Motion Planning Protocol for Heterogeneous Agents, Solvers, and Independent Tasks

基于冲突的搜索作为协议:一种多智能体运动规划协议用于异构智能体、求解器和独立任务

Rishi Veerapaneni, Alvin Tang, Haodong He, Sophia Zhao, Viraj Shah, Yidai Cen, Ziteng Ji, Gabriel Olin, Jon Arrizabalaga, Yorai Shaoul, Jiaoyang Li, Maxim Likhachev

机构 * Carnegie Mellon University(卡内基梅隆大学) Tongji University(同济大学) UC Berkeley(伯克利大学)

专题命中 工具调用 :agent(title,abstract);planning(title,abstract);multi-agent(title,abstract)

AI总结 本文提出了一种基于冲突的搜索协议,用于异构智能体的多智能体运动规划,利用多种单智能体规划算法实现无碰撞路径规划。

Comments Published at ICRA 2026, Project webpage: https://rishi-v.github.io/CBS-Protocol/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19255 2026-03-06 cs.LG cs.AI 81%

VTool-R1: VLMs Learn to Think with Images via Reinforcement Learning on Multimodal Tool Use

VTool-R1: 通过在多模态工具使用上的强化学习使VLMs学会通过图像思考

Mingyuan Wu, Jingcheng Yang, Jize Jiang, Meitang Li, Kaizhuo Yan, Hanchao Yu, Minjia Zhang, Chengxiang Zhai, Klara Nahrstedt

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) University of Michigan Ann Arbor(密歇根大学安娜堡分校) Independent Researcher(独立研究者)

专题命中 工具调用 :tool use(title,abstract);分类 cs.AI、cs.LG

AI总结 VTool-R1通过强化学习训练视觉语言模型生成多模态思考链,提升其通过图像进行推理的能力。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18764 2026-03-06 cs.AI cs.CL 66%

The Convergence of Schema-Guided Dialogue Systems and the Model Context Protocol

模式引导对话系统与模型上下文协议的收敛

Andreas Schlapbach

机构 * SBB-IT

专题命中 工具调用 :agent(abstract,comments);分类 cs.AI、cs.CL

AI总结 本文通过分析模式引导对话系统与模型上下文协议的收敛,提出五条模式设计原则,揭示了SGD与MCP的内在联系及软件3.0中的关键监督机制。

Comments 18 sections, 4 figures, 7 tables, 40 references. Original research presenting: (1) formal framework mapping Schema-Guided Dialogue principles to Model Context Protocol concepts, (2) five foundational design principles for LLM-native schema authoring, (3) architectural patterns for secure, scalable agent orchestration. Research supported by SBB (Swiss Federal Railways)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04735 2026-03-06 cs.AI cs.CL 62%

Solving an Open Problem in Theoretical Physics using AI-Assisted Discovery

用人工智能辅助发现解决理论物理中的一个开放问题

Michael P. Brenner, Vincent Cohen-Addad, David Woodruff

机构 * Google Research(谷歌研究) School of Engineering and Applied Sciences, Harvard University(工程与应用科学学院,哈佛大学) School of Computer Science, Carnegie Mellon University(计算机科学学院,卡内基梅隆大学)

专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.CL

AI总结 本文提出了一种结合大型语言模型和树搜索框架的混合系统,通过AI辅助发现解决了理论物理中宇宙弦引力辐射功率谱的开放问题,并推导出新的解析解。

Comments 22 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04705 2026-03-06 cs.RO cs.HC 50%

LEGS-POMDP: Language and Gesture-Guided Object Search in Partially Observable Environments

LEGS-POMDP:语言和手势引导的在部分可观测环境中的对象搜索

Ivy Xiao He, Stefanie Tellex, Jason Xinyu Liu

机构 * Brown University(布朗大学)

专题命中 工具调用 :planning(abstract)

AI总结 LEGS-POMDP通过整合语言、手势和视觉信息,在部分可观测环境中实现高效的开放世界对象搜索,显著提升了多模态感知和不确定性处理能力。

Comments 10 pages, 8 figures, accepted at ACM/IEEE International Conference on Human-Robot Interaction (HRI 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 规划决策 54 篇

2603.04750 2026-03-06 cs.AI cs.CL 92%

HiMAP-Travel: Hierarchical Multi-Agent Planning for Long-Horizon Constrained Travel

HiMAP-Travel: 基于长时域约束旅行的分层多智能体规划

The Viet Bui, Wenjun Li, Yong Liu

机构 * Department of XXX, University of YYY, Location, Country(YYY大学XXX系) School of ZZZ, Institute of WWW, Location, Country(WWW研究所ZZZ学院) Singapore Management University(新加坡管理大学)

专题命中 规划决策 :agent(title,abstract);planning(title,abstract);multi-agent(title,abstract);分类 cs.AI、cs.CL

AI总结 HiMAP-Travel通过分层多智能体框架提升长时域约束旅行规划效率,实现更高通过率和更低延迟。

Comments 33 pages, v1

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04659 2026-03-06 cs.RO cs.AI 90%

GIANT - Global Path Integration and Attentive Graph Networks for Multi-Agent Trajectory Planning

GIANT - 全局路径整合与注意力图网络用于多智能体轨迹规划

Jonas le Fevre Sejersen, Toyotaro Suzumura, Erdal Kayacan

机构 * Artificial Intelligence in Robotics Laboratory (AiR Lab), Department of Electrical and Computer Engineering, Aarhus University(人工智能机器人实验室(AiR实验室),电气与计算机工程系,奥胡斯大学) Foundation Models for Artificial Intelligence group, Department of Information and Communication Engineering, Tokyo University(人工智能基础模型组,信息与通信工程系,东京大学) Automatic Control Group, Department of Electrical Engineering and Information Technology, Paderborn University(自动控制组,电气工程与信息科技系,波德恩大学)

专题命中 规划决策 :planning(title,abstract);agent(title);multi-agent(title);分类 cs.AI

AI总结 GIANT通过结合全局路径规划与注意力图网络,提升多智能体在复杂动态环境中的避障与导航性能。

Comments Published in: 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05120 2026-03-06 cs.AI 88%

Bidirectional Curriculum Generation: A Multi-Agent Framework for Data-Efficient Mathematical Reasoning

双向课程生成:一种用于数据高效数学推理的多智能体框架

Boren Hu, Xiao Liu, Boci Peng, Xinping Zhao, Xiaoran Shang, Yun Zhu, Lijun Wu

机构 * Zhejiang University(浙江大学) University of Macau(澳门大学) Peking University(北京大学) Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)) Wuhan University(武汉大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)

专题命中 规划决策 :agent(title,abstract);multi-agent(title,abstract);分类 cs.AI

AI总结 本文提出双向课程生成框架,通过多智能体系统动态生成数据以提升数学推理效率,优化学习轨迹并减少训练样本需求。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.02663 2026-03-06 cs.CL cs.AI 84%

When Do Tools and Planning Help Large Language Models Think? A Cost- and Latency-Aware Benchmark

当工具和规划如何帮助大语言模型思考?一个成本和延迟感知的基准测试

Subha Ghoshal, Ali Al-Bustami

机构 * College of Engineering(工程学院) University of Michigan-Dearborn(密歇根大学-迪尔伯恩分校)

专题命中 规划决策 :planning(title,abstract);agent(abstract);分类 cs.AI、cs.CL

AI总结 研究通过对比不同配置评估工具和规划对大语言模型性能的影响,发现工具增强能显著提升准确性但增加延迟,而复杂工具协调易导致模型退化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04900 2026-03-06 cs.AI 83%

EvoTool: Self-Evolving Tool-Use Policy Optimization in LLM Agents via Blame-Aware Mutation and Diversity-Aware Selection

EvoTool: 通过责备感知变异和多样性感知选择实现LLM代理中自演化工具使用策略优化

Shuo Yang, Soyeon Caren Han, Xueqi Ma, Yan Li, Mohammad Reza Ghasemi Madani, Eduard Hovy

机构 * School of Computing and Information Systems, The University of Melbourne(计算与信息系统学院,墨尔本大学)

专题命中 规划决策 :tool-use(title,abstract);agent(abstract);分类 cs.AI

AI总结 EvoTool通过责备感知变异和多样性感知选择,优化LLM代理的模块化工具使用策略,提升效率与可迁移性。

Comments Work under review, 9 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22571 2026-03-06 cs.AI 83%

PerfGuard: A Performance-Aware Agent for Visual Content Generation

PerfGuard: 一种面向视觉内容生成的性能感知代理

Zhipeng Chen, Zhongrui Zhang, Chao Zhang, Yifan Xu, Lan Yang, Jun Liu, Ke Li, Yi-Zhe Song

机构 * School of Artificial Intelligence, Beijing University of Posts and Telecommunications, China(北京邮电大学人工智能学院) Beijing Digital Native Digital City Research Center, China(北京数字原生数字城市研究中心) School of Computer Science and Engineering, Beihang University, China(北航计算机科学与工程学院) SketchX, CVSSP, University of Surrey, United Kingdom(SketchX、CVSSP、英国萨里大学) Chenxi Shuzhi (Beijing) Technology Co., Ltd, China(北京晨曦智数科技有限公司)

专题命中 规划决策 :agent(title,abstract);planning(abstract);分类 cs.AI

AI总结 PerfGuard通过性能感知机制提升视觉内容生成任务中工具选择的准确性与执行可靠性。

Comments This paper has been accepted by ICLR 2026. The original paper link is: https://openreview.net/pdf?id=tdN42GTv4S The code repository link is: https://github.com/FelixChan9527/PerfGuard

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.03540 2026-03-06 cs.LG cs.AI 81%

Path Planning for Masked Diffusion Model Sampling

基于掩码扩散模型采样的路径规划

Fred Zhangzhi Peng, Zachary Bezemek, Sawan Patel, Jarrid Rector-Brooks, Sherwood Yao, Avishek Joey Bose, Alexander Tong, Pranam Chatterjee

机构 * Duke University(杜克大学) Atom Bioworks(Atom生物工坊) Mila – Québec AI Institute(魁北克AI研究院) Université de Montréal(蒙特利尔大学) The University of Oxford(牛津大学) Aithyra(Aithyra公司) University of Pennsylvania(宾夕法尼亚大学)

专题命中 规划决策 :planning(title,abstract);分类 cs.AI、cs.LG

AI总结 本文提出路径规划策略,通过规划和去噪两个阶段提升掩码扩散模型的生成质量,实验表明在多个领域均取得显著改进。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23589 2026-03-06 cs.AI cs.CV cs.LG 81%

BridgeDrive: Diffusion Bridge Policy for Closed-Loop Trajectory Planning in Autonomous Driving

BridgeDrive: 基于扩散桥的闭环轨迹规划扩散策略

Shu Liu, Wenlin Chen, Weihao Li, Zheng Wang, Lijin Yang, Jianing Huang, Yipin Zhang, Zhongzhan Huang, Ze Cheng, Hao Yang

机构 * Bosch (China) Investment Ltd(博世(中国)投资有限公司)

专题命中 规划决策 :planning(title,abstract);分类 cs.AI、cs.LG

AI总结 BridgeDrive提出一种基于锚点的扩散桥策略,通过实时闭环轨迹规划提升自动驾驶的安全性和效率。

Comments Accepted for publication at ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05438 2026-03-06 cs.CV cs.AI cs.RO 79%

Planning in 8 Tokens: A Compact Discrete Tokenizer for Latent World Model

8个标记的规划:一种用于潜在世界模型的紧凑离散标记器

Dongwon Kim, Gawon Seo, Jinsung Lee, Minsu Cho, Suha Kwak

机构 * KAIST(韩国科学技术院) POSTECH(POSTECH大学) RLWRLD

专题命中 规划决策 :planning(title,abstract);分类 cs.AI

AI总结 本文提出CompACT,一种将观测压缩为8个标记的离散标记器,显著降低计算成本并提升规划效率,为潜在世界模型的实际应用提供可行方案。

Comments CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05229 2026-03-06 cs.HC cs.AI 79%

Not All Trust is the Same: Effects of Decision Workflow and Explanations in Human-AI Decision Making

并非所有信任都相同:决策流程和解释在人机决策中的影响

Laura Spillner, Rachel Ringe, Robert Porzel, Rainer Malaka

机构 * Digital Media Lab, University of Bremen(数字媒体实验室,不莱梅大学)

专题命中 规划决策 :workflow(title,abstract);分类 cs.AI

AI总结 研究探讨了决策流程和解释对人机决策信任的影响,发现信任和依赖行为是不同构念,需分别评估。

Comments Accepted at Conversations 2025 Symposium

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04921 2026-03-06 cs.CL 79%

AILS-NTUA at SemEval-2026 Task 10: Agentic LLMs for Psycholinguistic Marker Extraction and Conspiracy Endorsement Detection

AILS-NTUA 在 SemEval-2026 任务 10: 用于心理语言学标记提取和阴谋支持检测的代理 LLM

Panagiotis Alexios Spanakis, Maria Lymperaiou, Giorgos Filandrianos, Athanasios Voulodimos, Giorgos Stamou

机构 * School of Electrical and Computer Engineering, AILS Laboratory(电气与计算机工程学院,AILS实验室) National Technical University of Athens(希腊雅典国家技术大学)

专题命中 规划决策 :agentic(title,abstract);分类 cs.CL

AI总结 AILS-NTUA 提出了一种代理 LLM 管道,通过动态判别链式思维和对抗性架构,在 SemEval-2026 任务 10 中实现了高精度的心理语言学标记提取和阴谋支持检测。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04898 2026-03-06 cs.LG cs.NI 79%

U-Parking: Distributed UWB-Assisted Autonomous Parking System with Robust Localization and Intelligent Planning

U-Parking:基于UWB的分布式自主泊车系统,具备稳健定位与智能规划

Yiang Wu, Qiong Wu, Pingyi Fan, Kezhi Wang, Wen Chen, Guoqiang Mao, Khaled B. Letaief

机构 * Jiangnan University(江南大学) Zhuhai Fudan Innovation Institute(珠海复旦创新研究院) Tsinghua University(清华大学) Brunel University of London(布鲁内尔大学) Shanghai Jiao Tong University(上海交通大学) Southeast University(东南大学) The Hong Kong University of Science and Technology(香港科学与技术大学)

专题命中 规划决策 :planning(title,abstract);分类 cs.LG

AI总结 U-Parking通过整合大语言模型辅助规划、稳健定位与轨迹跟踪,实现具有挑战性的室内环境中的可靠自动化泊车。

Comments This paper has been accepted by infocom. The source code has been released at: https://github.com/qiongwu86/U-Parking

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22425 2026-03-06 cs.CV cs.AI 79%

FluenceFormer: Transformer-Driven Multi-Beam Fluence Map Regression for Radiotherapy Planning

FluenceFormer:基于Transformer的多束荧光图回归用于放疗规划

Ujunwa Mgboh, Rafi Ibn Sultan, Joshua Kim, Kundan Thind, Dongxiao Zhu

机构 * Wayne State University(韦恩州立大学) Henry Ford Health(亨利福特医疗)

专题命中 规划决策 :planning(title,abstract);分类 cs.AI

AI总结 FluenceFormer 是一种基于 Transformer 的多束荧光图回归方法,通过两阶段设计和物理约束损失提升放疗计划的结构保真度和能量精度。

Comments Accepted at Medical Imaging with Deep Learning (MIDL-2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.21039 2026-03-06 cs.IR cs.LG 79%

Agentic Multi-Persona Framework for Evidence-Aware Fake News Detection

代理多角色框架用于证据感知的虚假新闻检测

Roopa Bukke, Soumya Pandey, Suraj Kumar, Soumi Chattopadhyay, Chandranath Adak

机构 * Dept. of CSE, Indian Institute of Technology Indore(计算机科学与工程系,印度理工学院印度奥尔德分校) Dept. of CSE, Indian Institute of Technology Patna(计算机科学与工程系,印度理工学院帕纳分校)

专题命中 规划决策 :agentic(title,abstract);分类 cs.LG

AI总结 本文提出AMPEND-LS框架,通过LLM与SLM协同作用,实现多模态虚假新闻检测,提升准确率和鲁棒性,增强信息安全性。

Comments 10 pages, 3 tables, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05333 2026-03-06 cs.RO 78%

CT-Enabled Patient-Specific Simulation and Contact-Aware Robotic Planning for Cochlear Implantation

基于CT的患者特异性模拟与接触感知的耳蜗植入手术机器人规划

Lingxiao Xun, Gang Zheng, Alexandre Kruszewski, Renato Torres

专题命中 规划决策 :planning(title,abstract)

AI总结 本文提出了一种基于CT的患者特异性模拟与接触感知的耳蜗植入手术机器人规划方法,通过低维可微模型和伪动力学正则化,实现精确的接触力预测和手术安全。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26139 2026-03-06 cs.RO 78%

Kinodynamic Task and Motion Planning using VLM-guided and Interleaved Sampling

基于VLM引导和交错采样的运动动力学任务与动作规划

Minseo Kwon, Young J. Kim

机构 * Department of Computer Science and Engineering(计算机科学与工程系)

专题命中 规划决策 :planning(title,abstract)

AI总结 本文提出一种基于VLM引导和交错采样的运动动力学任务与动作规划方法,通过混合状态树联合决策任务和运动,提升规划效率和成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05017 2026-03-06 cs.RO 78%

Direct Contact-Tolerant Motion Planning With Vision Language Models

直接接触容忍的运动规划与视觉语言模型

He Li, Jian Sun, Chengyang Li, Guoliang Li, Qiyu Ruan, Shuai Wang, Chengzhong Xu

机构 * State Key Laboratory of Internet of Things for Smart City (SKL-IOTSC), University of Macau(物联网智能城市国家重点实验室,澳门大学) Shenzhen Institutes of Advanced Technology (SIAT), Chinese Academy of Sciences(深圳先进技术研究院,中国科学院) Department of Electrical and Computer Engineering, The University of Hong Kong(香港大学电子与计算机工程系)

专题命中 规划决策 :planning(title,abstract)

AI总结 本文提出直接接触容忍运动规划方法,利用视觉语言模型实现接触感知导航,提升机器人在拥挤环境中的鲁棒性和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04547 2026-03-06 cs.RO 78%

Many-RRT*: Robust Joint-Space Trajectory Planning for Serial Manipulators

多RRT*:串联机械臂的联合空间轨迹规划

Theodore M. Belmont, Benjamin A. Christie, Anton Netchaev

机构 * USACE ERDC, Information Technology Lab(美国陆军工程兵团工程研究中心信息科技实验室) Collaborative Robotics Lab(协作机器人实验室)

专题命中 规划决策 :planning(title,abstract)

AI总结 Many-RRT*通过同时规划多个目标解决串联机械臂在关节空间中规划的鲁棒性和最优性问题,显著提高了轨迹质量和成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04463 2026-03-06 cs.RO 78%

GAIDE: Graph-based Attention Masking for Spatial- and Embodiment-aware Motion Planning

基于图的注意力掩码的时空和具身感知运动规划

Davood Soleymanzadeh, Xiao Liang, Minghui Zheng

机构 * J. Mike Walker ’66 Department of Mechanical Engineering, Texas A&M University, College Station, TX 77843, USA(J. Mike Walker ’66 机械工程系,德克萨斯A&M大学,College Station, TX 77843, USA) Zachry Department of Civil and Environmental Engineering, Texas A&M University, College Station, TX 77843 USA(Zachry 市政与环境工程系,德克萨斯A&M大学,College Station, TX 77843 USA)

专题命中 规划决策 :planning(title,abstract)

AI总结 GAIDE是一种基于图的神经指导采样器,通过结合空间结构和机械臂的具身性提升运动规划的效率和成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.03137 2026-03-06 cs.RO 78%

RL-Based Coverage Path Planning for Deformable Objects on 3D Surfaces

基于强化学习的可变形物体3D表面覆盖路径规划

Yuhang Zhang, Jinming Ma, Feng Wu

机构 * School of Computer Science and Technology, University of Science and Technology of China(计算机科学与技术学院,中国科学技术大学) Xiaomi Robotics Lab(小米机器人实验室)

专题命中 规划决策 :planning(title);agent(abstract)

AI总结 本文提出基于强化学习的覆盖路径规划方法,利用谐波UV映射和SGCNN处理可变形物体的触觉反馈,通过模拟器训练机器人高效完成表面擦拭任务。

Comments 8 pages, 8 figures. Accepted to the 2026 IEEE International Conference on Robotics and Automation (ICRA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05218 2026-03-06 cs.AI cs.LG 73%

KARL: Knowledge Agents via Reinforcement Learning

通过强化学习的知识代理:KARL

Jonathan D. Chang, Andrew Drozdov, Shubham Toshniwal, Owen Oertell, Alexander Trott, Jacob Portes, Abhay Gupta, Pallavi Koppol, Ashutosh Baheti, Sean Kulinski, Ivan Zhou, Irene Dea, Krista Opsahl-Ong, Simon Favreau-Lessard, Sean Owen, Jose Javier Gonzalez Ortiz, Arnav Singhvi, Xabi Andrade, Cindy Wang, Kartik Sreenivasan, Sam Havens, Jialu Liu, Peyton DeNiro, Wen Sun, Michael Bendersky, Jonathan Frankle

机构 * Databricks AI Research(Databricks人工智能研究)

专题命中 规划决策 :tool use(abstract);agentic(abstract);分类 cs.AI、cs.LG

AI总结 KARL通过强化学习训练企业搜索代理,结合多任务学习和定制合成数据,实现高效且高性能的知识代理,在多个复杂任务中表现优异。

Comments 77 pages, 43 figures, 17 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04861 2026-03-06 cs.AI cs.LG cs.RO 62%

Causally Robust Reward Learning from Reason-Augmented Preference Feedback

基于理由增强的偏好反馈的因果鲁棒奖励学习

Minjune Hwang, Yigit Korkmaz, Daniel Seita, Erdem Bıyık

机构 * Thomas Lord Department of Computer Science, University of Southern California(美国南加州大学计算机科学系)

专题命中 规划决策 :agent(abstract);分类 cs.AI、cs.LG

AI总结 ReCouPLe通过自然语言理由提供因果信号,提升奖励学习在分布偏移和新任务中的性能。

Comments Published in International Conference on Learning Representations (ICLR) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04534 2026-03-06 cs.LG cs.AI cs.CY 62%

Invariant Causal Routing for Governing Social Norms in Online Market Economies

不变因果路由:用于在线市场经济中规范治理

Xiangning Yu, Qirui Mi, Xiao Xue, Haoxuan Li, Yiwei Shi, Xiaowei Liu, Mengyue Yang

机构 * Tianjin University(天津大学) Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) Peking University(北京大学) University of Bristol(布里斯托大学)

专题命中 规划决策 :agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出不变因果路由(ICR)框架,通过因果推理和不变因果发现,实现在线市场经济中稳定社会规范的治理,提升政策的可解释性和有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19948 2026-03-06 cs.CL cs.AI cs.CY cs.HC cs.MA 62%

Assessing Risks of Large Language Models in Mental Health Support: A Framework for Automated Clinical AI Red Teaming

评估大型语言模型在心理健康支持中的风险:一种用于自动化临床AI红队测试的框架

Ian Steenstra, Paola Pedrelli, Weiyan Shi, Stacy Marsella, Timothy W. Bickmore

机构 * Northeastern University(东北大学) Harvard Medical School(哈佛医学院)

专题命中 规划决策 :AI agent(abstract);分类 cs.AI、cs.CL

AI总结 本文提出了一种评估AI心理治疗师在心理健康支持中安全风险的框架,通过模拟测试发现AI在治疗中的潜在风险,并验证了交互式可视化工具的有效性。

Comments This paper is a condensed version of the first author's Ph.D. dissertation submitted to Northeastern University

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.11063 2026-03-06 cs.RO cs.AI cs.CV cs.LG cs.MA 62%

EmboTeam: Grounding LLM Reasoning into Reactive Behavior Trees via PDDL for Embodied Multi-Robot Collaboration

EmboTeam: 通过PDDL将LLM推理接地为反应式行为树以实现具身多机器人协作

Haishan Zeng, Mengna Wang, Peng Li

机构 * University of Chinese Academy of Sciences(中国科学院大学) Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所)

专题命中 规划决策 :planning(abstract);分类 cs.AI、cs.LG

AI总结 EmboTeam通过PDDL将LLM推理转化为反应式行为树,提升多机器人协作任务的成功率和目标回忆率

详情

展开后加载摘要…

URL PDF HTML 收藏