arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 5062 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 工具调用 5062 篇

1905.13521 2019-06-03 cs.AI 57%

Multiple Policy Value Monte Carlo Tree Search

Li-Cheng Lan, Wei Li, Ting-Han Wei, I-Chen Wu

专题命中 工具调用 :agent(abstract);分类 cs.AI

Comments Proceedings of the 28th International Joint Conference on Artificial Intelligence (IJCAI-19)

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.07189 2019-05-31 cs.RO cs.AI 57%

Reinforcement Learning with Probabilistic Guarantees for Autonomous Driving

Maxime Bouton, Jesper Karlsson, Alireza Nakhaei, Kikuo Fujimura, Mykel J. Kochenderfer, Jana Tumova

专题命中 工具调用 :agent(abstract);分类 cs.AI

Journal ref Workshop on Safety Risk and Uncertainty in Reinforcement Learning, Conference on Uncertainty in Artificial Intelligence (UAI), 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.10974 2019-05-28 cs.CV cs.LG eess.IV 57%

Style transfer-based image synthesis as an efficient regularization technique in deep learning

Agnieszka Mikołajczyk, Michał Grochowski

专题命中 工具调用 :tool use(abstract);分类 cs.LG

Comments 6 pages, 4 figures, accepted to the 24th International Conference on Methods and Models in Automation and Robotics (MMAR 2019)

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.08332 2019-05-22 cs.RO cs.AI 57%

Behavior Identification and Prediction for a Probabilistic Risk Framework

Jasprit Singh Gill, Pierluigi Pisu, Venkat N. Krovi, Matthias J. Schmid

专题命中 工具调用 :planning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.06402 2019-05-17 cs.AI 57%

Improved Safe Real-time Heuristic Search

Bence Cserna, Kevin C. Gall, Wheeler Ruml

专题命中 工具调用 :planning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1905.03697 2019-05-16 cs.CV cs.LG 57%

On Applying Machine Learning/Object Detection Models for Analysing Digitally Captured Physical Prototypes from Engineering Design Projects

Jorgen F. Erichsen, Sampsa Kohtala, Martin Steinert, Torgeir Welo

专题命中 工具调用 :workflow(abstract);分类 cs.LG

Comments 13 pages, 4 tables, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.02282 2019-04-23 stat.ML cs.LG math.OC 57%

Finding the bandit in a graph: Sequential search-and-stop

Pierre Perrault, Vianney Perchet, Michal Valko

专题命中 工具调用 :agent(abstract);分类 cs.LG

Comments in International Conference on Artificial Intelligence and Statistics (AISTATS 2019), April 2019, Naha, Okinawa, Japan

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.07366 2019-04-17 cs.AI cs.DM 57%

Efficiently Exploring Ordering Problems through Conflict-directed Search

Jingkai Chen, Cheng Fang, David Wang, Andrew Wang, Brian Williams

专题命中 工具调用 :planning(abstract);分类 cs.AI

Comments Accepted at ICAPS2019. 9 pages, 4 figures, 2 tables,

详情

展开后加载摘要…

URL PDF HTML 收藏
1904.06879 2019-04-16 cs.AI cs.RO 57%

Improving interactive reinforcement learning: What makes a good teacher?

Francisco Cruz, Sven Magg, Yukie Nagai, Stefan Wermter

专题命中 工具调用 :agent(abstract);分类 cs.AI

Comments 21 pages, 12 figures

Journal ref Connection Science, Vol. 30, Nr. 3, 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.08922 2019-04-08 cs.IR cs.LG stat.ML 57%

Context-Aware Systems for Sequential Item Recommendation

Moin Nadeem, Dustin Stansbury, Shane Mooney

专题命中 工具调用 :tool use(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.10236 2019-04-05 cs.CL 57%

Learning When Not to Answer: A Ternary Reward Structure for Reinforcement Learning based Question Answering

Fréderic Godin, Anjishnu Kumar, Arpit Mittal

专题命中 工具调用 :agent(abstract);分类 cs.CL

Comments Accepted at NAACL 2019. Version 1 was presented at NIPS 2018 workshop on Relational Representation Learning

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.07260 2019-03-19 cs.AI 57%

Intelligent Solution System towards Parts Logistics Optimization

Yaoting Huang, Boyu Chen, Wenlian Lu, Zhong-Xiao Jin, Ren Zheng

专题命中 工具调用 :planning(abstract);分类 cs.AI

Comments WCGO 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.03756 2019-03-12 cs.LG stat.ML 57%

Two-Hop Walks Indicate PageRank Order

Ying Tang

专题命中 工具调用 :tool use(abstract);分类 cs.LG

Comments 29 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1903.01747 2019-03-11 cs.LG stat.ML 57%

Towards Understanding Chinese Checkers with Heuristics, Monte Carlo Tree Search, and Deep Reinforcement Learning

Ziyu Liu, Meng Zhou, Weiqing Cao, Qiang Qu, Henry Wing Fung Yeung, Vera Yuk Ying Chung

专题命中 工具调用 :agent(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.07828 2019-02-22 stat.ML cs.IT cs.LG math.IT 57%

Correspondence Analysis Using Neural Networks

Hsiang Hsu, Salman Salamatian, Flavio P. Calmon

专题命中 工具调用 :tool use(abstract);分类 cs.LG

Comments Accepted to AISTATS 2019. Overlaps with arXiv:1806.08449

详情

展开后加载摘要…

URL PDF HTML 收藏
1805.01367 2019-02-14 cs.LG stat.ML 57%

Open Loop Execution of Tree-Search Algorithms, extended version

Erwan Lecarpentier, Guillaume Infantes, Charles Lesire, Emmanuel Rachelson

专题命中 工具调用 :planning(abstract);分类 cs.LG

Comments 10 pages, 10 figures

Journal ref 27th International Joint Conference on Artificial Intelligence (IJCAI 2018)

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.03142 2019-02-11 cs.AI 57%

Novelty Search for Deep Reinforcement Learning Policy Network Weights by Action Sequence Edit Metric Distance

Ethan C. Jackson, Mark Daley

专题命中 工具调用 :agent(abstract);分类 cs.AI

Comments Submitted to GECCO 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.02870 2019-02-11 cs.AI cs.CV cs.RO 57%

Visual search and recognition for robot task execution and monitoring

Lorenzo Mauro, Francesco Puja, Simone Grazioso, Valsamis Ntouskos, Marta Sanzari, Edoardo Alati, Fiora Pirri

专题命中 工具调用 :planning(abstract);分类 cs.AI

Journal ref Frontiers in Artificial Intelligence and Applications 310 (2018) 94-109

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.02778 2019-02-08 cs.LG stat.ML 57%

KLUCB Approach to Copeland Bandits

Nischal Agrawal, Prasanna Chaporkar

专题命中 工具调用 :agent(abstract);分类 cs.LG

Comments 10 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1902.00843 2019-02-05 cs.LG stat.ML 57%

A Meta-MDP Approach to Exploration for Lifelong Reinforcement Learning

Francisco M. Garcia, Philip S. Thomas

专题命中 工具调用 :agent(abstract);分类 cs.LG

Comments Accepted as Extended Abstract, AAMAS, 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1901.09125 2019-01-29 cs.AI cs.LO 57%

The informal semantics of Answer Set Programming: A Tarskian perspective

Marc Denecker, Yuliya Lierler, Miroslaw truszczynski, Joost Vennekens

专题命中 工具调用 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1901.08013 2019-01-24 cs.NE cs.LG stat.ML 57%

DarwinML: A Graph-based Evolutionary Algorithm for Automated Machine Learning

Fei Qi, Zhaohui Xia, Gaoyang Tang, Hang Yang, Yu Song, Guangrui Qian, Xiong An, Chunhuan Lin, Guangming Shi

专题命中 工具调用 :workflow(abstract);分类 cs.LG

Comments 8 pages, 7 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
1812.03625 2018-12-11 cs.AI 57%

A context-based geoprocessing framework for optimizing meetup location of multiple moving objects along road networks

Shaohua Wang, Song Gao, Xin Feng, Alan T. Murray, Yuan Zeng

专题命中 工具调用 :planning(abstract);分类 cs.AI

Comments 34 pages, 8 figures

Journal ref International Journal of Geographical Information Science, 32(7), 1368-1390 (2018)

详情

展开后加载摘要…

URL PDF HTML 收藏
1811.12444 2018-12-03 cs.LG stat.ML 57%

Flow Shape Design for Microfluidic Devices Using Deep Reinforcement Learning

Xian Yeow Lee, Aditya Balu, Daniel Stoecklein, Baskar Ganapathysubramanian, Soumik Sarkar

专题命中 工具调用 :agent(abstract);分类 cs.LG

Comments Neurips 2018 Deep RL workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
1806.00553 2018-11-27 cs.AI 57%

Deep Curiosity Search: Intra-Life Exploration Can Improve Performance on Challenging Deep Reinforcement Learning Problems

Christopher Stanton, Jeff Clune

专题命中 工具调用 :agent(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.10325 2018-10-29 cs.CV cs.LG stat.ML 57%

Multi-Stage Reinforcement Learning For Object Detection

Jonas Koenig, Simon Malberg, Martin Martens, Sebastian Niehaus, Artus Krohn-Grimberghe, Arunselvan Ramaswamy

专题命中 工具调用 :agent(abstract);分类 cs.LG

Comments Accepted for the Computer Vision Conference (CVC) 2019

详情

展开后加载摘要…

URL PDF HTML 收藏
1810.10908 2018-10-26 cs.AI 57%

MGP: Un algorithme de planification temps réel prenant en compte l'évolution dynamique du but

Damien Pellier, Mickaël Vanneufville, Humbert Fiorino, Marc Métivier, Bruno Bouzy

专题命中 工具调用 :planning(abstract);分类 cs.AI

Comments in French

Journal ref Journées Francophones de Planification, Décision, Apprentissage pour la conduite de Systèmes, 2012

详情

展开后加载摘要…

URL PDF HTML 收藏
1809.08613 2018-09-25 cs.RO cs.LG stat.ML 57%

Detecting Features of Tools, Objects, and Actions from Effects in a Robot using Deep Learning

Namiko Saito, Kitae Kim, Shingo Murata, Tetsuya Ogata, Shigeki Sugano

专题命中 工具调用 :tool-use(abstract);分类 cs.LG

Comments 7 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
1807.06286 2018-09-20 cs.AI 57%

Preference-Based Monte Carlo Tree Search

Tobias Joppen, Christian Wirth, Johannes Fürnkranz

专题命中 工具调用 :agent(abstract);分类 cs.AI

Comments To be published

Journal ref Proceedings of the 41st German Conference on Artificial Intelligence (KI-18), 2018

详情

展开后加载摘要…

URL PDF HTML 收藏
1707.02553 2018-09-19 cs.SE cs.DL 57%

Guidelines for including grey literature and conducting multivocal literature reviews in software engineering

Vahid Garousi, Michael Felderer, Mika V. Mäntylä

专题命中 工具调用 :planning(abstract);分类 cs.SE

详情

展开后加载摘要…

URL PDF HTML 收藏