arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 93044 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 工具调用 5059 篇

1904.01241 2020-12-21 cs.CV cs.AI 70%

Centerline Depth World Reinforcement Learning-based Left Atrial Appendage Orifice Localization

Walid Abdullah Al, Il Dong Yun, Eun Ju Chun

专题命中 工具调用 :agent(abstract);planning(abstract);分类 cs.AI

Comments 10 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2009.05643 2020-09-15 cs.AI 70%

The Design Of "Stratega": A General Strategy Games Framework

Diego Perez-Liebana, Alexander Dockhorn, Jorge Hurtado Grueso, Dominik Jeurissen

专题命中 工具调用 :agent(abstract);planning(abstract);分类 cs.AI

Comments 7 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2008.12610 2020-08-31 cs.RO cs.LG cs.MA 70%

Collaborative Multi-Robot Systems for Search and Rescue: Coordination and Perception

Jorge Peña Queralta, Jussi Taipalmaa, Bilge Can Pullinen, Victor Kathan Sarker, Tuan Nguyen Gia, Hannu Tenhunen, Moncef Gabbouj, Jenni Raitoharju, Tomi Westerlund

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.00618 2020-08-18 cs.LG math.OC stat.ML 70%

What is Local Optimality in Nonconvex-Nonconcave Minimax Optimization?

Chi Jin, Praneeth Netrapalli, Michael I. Jordan

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.LG

Comments This paper has been published at ICML2020. This new version made a correction to Proposition 19, and added more related works

详情

展开后加载摘要…

URL PDF HTML 收藏
2005.04974 2020-05-12 eess.IV cs.CV cs.LG 70%

Deep Reinforcement Learning for Organ Localization in CT

Fernando Navarro, Anjany Sekuboyina, Diana Waldmannstetter, Jan C. Peeken, Stephanie E. Combs, Bjoern H. Menze

专题命中 工具调用 :agent(abstract);planning(abstract);分类 cs.LG

Comments Accepted paper in MIDL 2020

Journal ref https://openreview.net/forum?id=0vDeD2UD0S&referrer=%5BAuthor%20Console%5D(%2Fgroup%3Fid%3DMIDL.io%2F2020%2FConference%2FAuthors%23your-submissions)

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.04307 2020-04-15 cs.LG cond-mat.soft 70%

A cooperative game for automated learning of elasto-plasticity knowledge graphs and models with AI-guided experimentation

Kun Wang, WaiChing Sun, Qiang Du

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1911.03366 2020-03-09 cs.NI cs.LG stat.ML 70%

Deep Reinforcement Learning for Distributed Uncoordinated Cognitive Radios Resource Allocation

Ankita Tondwalkar, Dr Andres Kwasinski

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.LG

Comments This paper has been submitted in the 21st IEEE International Workshop On Signal Processing Advances In Wireless Communications (SPAWC 2020)

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.05751 2020-03-05 stat.ML cs.LG cs.RO 70%

Trajectory Optimization for Unknown Constrained Systems using Reinforcement Learning

Kei Ota, Devesh K. Jha, Tomoaki Oiki, Mamoru Miura, Takashi Nammoto, Daniel Nikovski, Toshisada Mariyama

专题命中 工具调用 :agent(abstract);planning(abstract);分类 cs.LG

Comments 8 pages, 6 figures, Accepted to IROS 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.10090 2020-01-16 cs.LG stat.ML 70%

Non-Stationary Markov Decision Processes, a Worst-Case Approach using Model-Based Reinforcement Learning, Extended version

Erwan Lecarpentier, Emmanuel Rachelson

专题命中 工具调用 :agent(abstract);planning(abstract);分类 cs.LG

Comments Published at NeurIPS 2019, 17 pages, 3 figures

Journal ref year: 2019; page range: 7214--7223

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.02807 2020-01-13 cs.LG stat.ML 70%

Combining Q-Learning and Search with Amortized Value Estimates

Jessica B. Hamrick, Victor Bapst, Alvaro Sanchez-Gonzalez, Tobias Pfaff, Theophane Weber, Lars Buesing, Peter W. Battaglia

专题命中 工具调用 :agent(abstract);planning(abstract);分类 cs.LG

Comments Published as a conference paper at ICLR 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1912.02318 2019-12-06 cs.AI cs.MA 70%

Improving Policies via Search in Cooperative Partially Observable Games

Adam Lerer, Hengyuan Hu, Jakob Foerster, Noam Brown

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.AI

Journal ref AAAI 2020

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.08792 2019-06-24 cs.LG cs.IT math.IT stat.ML 70%

When Multiple Agents Learn to Schedule: A Distributed Radio Resource Management Framework

Navid Naderializadeh, Jaroslaw Sydir, Meryem Simsek, Hosein Nikopour, Shilpa Talwar

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.LG

Comments Submitted to IEEE Wireless Communications Magazine

详情

展开后加载摘要…

URL PDF HTML 收藏
1901.01855 2019-02-19 cs.AI 70%

A* Tree Search for Portfolio Management

Xiaojie Gao, Shikui Tu, Lei Xu

专题命中 工具调用 :agent(abstract);planning(abstract);分类 cs.AI

Comments The paper needs a major revision including the title

详情

展开后加载摘要…

URL PDF HTML 收藏
1805.08263 2018-10-01 cs.AI cs.RO 70%

Learning What Information to Give in Partially Observed Domains

Rohan Chitnis, Leslie Pack Kaelbling, Tomás Lozano-Pérez

专题命中 工具调用 :agent(abstract);autonomous agent(abstract);分类 cs.AI

Comments CoRL 2018 final version

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.09487 2018-06-26 cs.AI 70%

Finding Optimal Solutions to Token Swapping by Conflict-based Search and Reduction to SAT

Pavel Surynek

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1803.00259 2018-03-02 cs.AI 70%

Deep Reinforcement Learning for Sponsored Search Real-time Bidding

Jun Zhao, Guang Qiu, Ziyu Guan, Wei Zhao, Xiaofei He

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1702.06915 2017-02-24 cs.AI 70%

Solving DCOPs with Distributed Large Neighborhood Search

Ferdinando Fioretto, Agostino Dovier, Enrico Pontelli, William Yeoh, Roie Zivan

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1612.04876 2016-12-16 cs.AI 70%

Collaborative creativity with Monte-Carlo Tree Search and Convolutional Neural Networks

Memo Akten, Mick Grierson

专题命中 工具调用 :agent(abstract);autonomous agent(abstract);分类 cs.AI

Comments Presented at the Constructive Machine Learning workshop at NIPS 2016 as a poster and spotlight talk. 8 pages including 2 page references, 2 page appendix, 3 figures. Blog post (including videos) at https://medium.com/@memoakten/collaborative-creativity-with-monte-carlo-tree-search-and-convolutional-neural-networks-and-other-69d7107385a0

详情

展开后加载摘要…

URL PDF HTML 收藏
1503.04193 2015-09-07 cs.LO cs.AI 70%

Non-normal modalities in variants of Linear Logic

Daniele Porello, Nicolas Troquard

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1505.01629 2015-05-08 cs.LO cs.AI cs.MA cs.MS 70%

LeoPARD --- A Generic Platform for the Implementation of Higher-Order Reasoners

Max Wisniewski, Alexander Steen, Christoph Benzmüller

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.AI

Comments 6 pages, to appear in the proceedings of CICM'2015 conference

详情

展开后加载摘要…

URL PDF HTML 收藏
1405.1734 2014-05-16 cs.MA cs.AI 70%

Logic and Constraint Logic Programming for Distributed Constraint Optimization

Tiep Le, Enrico Pontelli, Tran Cao Son, William Yeoh

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.AI

Comments To appear in Theory and Practice of Logic Programming (TPLP)

详情

展开后加载摘要…

URL PDF HTML 收藏
1402.0587 2014-02-05 cs.AI 70%

Asymmetric Distributed Constraint Optimization Problems

Tal Grinshpoun, Alon Grubshtein, Roie Zivan, Arnon Netzer, Amnon Meisels

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.AI

Journal ref Journal Of Artificial Intelligence Research, Volume 47, pages 613-647, 2013

详情

展开后加载摘要…

URL PDF HTML 收藏
1401.3857 2014-01-17 cs.AI 70%

Case-Based Subgoaling in Real-Time Heuristic Search for Video Game Pathfinding

Vadim Bulitko, Yngvi Björnsson, Ramon Lawrence

专题命中 工具调用 :agent(abstract);planning(abstract);分类 cs.AI

Journal ref Journal Of Artificial Intelligence Research, Volume 39, pages 269-300, 2010

详情

展开后加载摘要…

URL PDF HTML 收藏
1207.1832 2012-07-11 cs.AI cs.LO 70%

Minimal Proof Search for Modal Logic K Model Checking

Abdallah Saffidine

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.AI

Comments Extended version of the JELIA 2012 paper with the same title

详情

展开后加载摘要…

URL PDF HTML 收藏
1110.4076 2011-10-19 cs.AI 70%

Learning in Real-Time Search: A Unifying Framework

V. Bulitko, G. Lee

专题命中 工具调用 :agent(abstract);planning(abstract);分类 cs.AI

Journal ref Journal Of Artificial Intelligence Research, Volume 25, pages 119-157, 2006

详情

展开后加载摘要…

URL PDF HTML 收藏
1110.2209 2011-10-12 cs.AI 70%

Bin Completion Algorithms for Multicontainer Packing, Knapsack, and Covering Problems

A. S. Fukunaga, R. E. Korf

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.AI

Journal ref Journal Of Artificial Intelligence Research, Volume 28, pages 393-429, 2007

详情

展开后加载摘要…

URL PDF HTML 收藏
1004.2880 2010-04-19 cs.AI cs.MA 70%

GRASP for the Coalition Structure Formation Problem

Nicola Di Mauro, Teresa M. A. Basile, Stefano Ferilli, Floriana Esposito

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.AI

Comments 12 pages, Submitted to an International Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
cs/0509071 2009-12-01 cs.GT cs.AI 70%

CP-nets and Nash equilibria

Krzysztof R. Apt, Francesca Rossi, K. Brent Venable

专题命中 工具调用 :agent(abstract);multi-agent(abstract);分类 cs.AI

Comments 6 pages. in: roc. of the Third International Conference on Computational Intelligence, Robotics and Autonomous Systems (CIRAS '05). To appear

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.05519 2026-08-07 cs.AI cs.CL cs.LG 新提交 69%

EcoAgent-Bench: Evaluating Economic Decision-Making in Budget-Constrained LLM Agents

EcoAgent-Bench:评估预算约束下大语言模型智能体的经济决策能力

Jie Wu, Ming Gong, Feixiang Cheng, Qinqin Zhao

专题命中 工具调用 :agent(abstract,comments);分类 cs.AI、cs.CL、cs.LG

AI总结 本文提出EcoAgent-Bench基准,评估大语言模型智能体在预算约束下的经济决策能力,发现现有智能体在该任务上表现不佳,发布了相关研究资源。

Comments 8 pages, 3 figures, 4 tables. Benchmark, dataset (304 budget-conditioned agent tasks), and evaluation harness; artifacts to be released

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.18573 2026-08-20 cs.CV 新提交 67%

PATE-Forensics: Perception-as-Tool for Explainable Deepfake Forensics with General-Purpose MLLMs

PATE-Forensics:以通用多模态大语言模型(MLLM)为工具的可解释深度伪造取证方法

Yaqi Li, Jielun Peng, Yabin Wang, Jincheng Liu, Xiaopeng Hong

机构 * Harbin Institute of Technology(哈尔滨工业大学)

专题命中 工具调用 :agent(abstract);tool use(abstract)

AI总结 本研究提出PATE-Forensics,采用“感知即工具”范式,基于DINOv3构建取证感知工具,结合通用MLLM实现可解释深度伪造取证,在DDL-X Track 3数据集上取得0.89的最佳官方分数,较次席高出0.19分。

Comments 9 pages, 3 figures, 2 tables; DDL-X Track 3, IJCAI 2026 AI Safety Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏