arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

AI Agent

智能体、工具调用、规划、工作流、多智能体和自主任务执行。

共收录 590 信号源:cs.AI, cs.CL, cs.LG, cs.SE

1. 其他Agent 590 篇

2604.25199 2026-04-29 cond-mat.mtrl-sci cond-mat.str-el cs.AI cs.LG physics.comp-ph 62%

Kohn-Sham Hamiltonian from Effective Field Theory: Quasiparticle Band Narrowing from Frozen Core Dynamics

从有效场论看Kohn-Sham哈密顿量:冻结核心动力学导致的准粒子带变窄

Xiansheng Cai, Han Wang, Kun Chen

机构 * Institute of Theoretical Physics, Chinese Academy of Sciences, Beijing 100190, China(中国科学院理论物理研究所,北京100190,中国)

专题命中 其他Agent :agentic(abstract);分类 cs.AI、cs.LG

AI总结 本文通过有效场论揭示Kohn-Sham带宽与ARPES测量的差异,提出冻结核心修正因子,解决准粒子带宽问题,验证了第一性原理智能科学方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.18349 2026-03-30 cs.AI cs.CL 62%

Large-Scale Analysis of Persuasive Content on Moltbook

对Moltbook上说服性内容的大规模分析

Julia Jose, Meghna Manoj Nair, Rachel Greenstadt

机构 * New York University(纽约大学)

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.CL

AI总结 研究通过LLM分类器分析Moltbook上政治宣传内容,发现政治宣传占所有帖子的1%和政治内容的42%,集中于少数社区,由少数代理生成,但评论未显著放大宣传内容。

Comments 9 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08139 2026-03-18 cs.CL cs.AI 62%

Can LLMs Detect Their Confabulations? Estimating Reliability in Uncertainty-Aware Language Models

LLMs能否检测其编造?在不确定性感知语言模型中估计可靠性

Tianyi Zhou, Johanne Medina, Sanjay Chawla

机构 * KTH Royal Institute of Technology(皇家理工学院) QCRI, HBKU(哈马德 bin 玉素菲大学量子计算与人工智能研究所)

专题命中 其他Agent :agentic(abstract);分类 cs.AI、cs.CL

AI总结 研究探讨了上下文信息如何影响模型行为,并提出利用token级不确定性来指导内部表示聚合的可靠性估计方法,通过实验发现正确上下文能提升回答准确性,而误导性上下文常导致自信错误响应。

Comments Published at AAAI'26

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08575 2026-03-10 cs.AI cs.LG 62%

Trust via Reputation of Conviction

通过可信度声誉建立信任

Aravind R. Iyengar

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出基于可信度的声誉框架,用于人工智能代理的信任建立,强调可信度而非正确性作为信任的基础。

Comments 19 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20886 2026-03-09 cs.CL cs.AI 62%

Shoot First, Ask Questions Later? Building Rational Agents that Explore and Act Like People

先行动,后提问?构建能探索和行动像人的理性代理

Gabriel Grand, Valerio Pepe, Jacob Andreas, Joshua B. Tenenbaum

机构 * MIT CSAIL(麻省理工学院计算机科学与人工智能实验室) MIT Brain and Cognitive Sciences(麻省理工学院脑科学与认知科学) Harvard SEAS(哈佛大学工程学院)

专题命中 其他Agent :agentic(abstract);分类 cs.AI、cs.CL

AI总结 本文提出了一种基于贝叶斯实验设计的蒙特卡洛推理策略,用于构建能像人类一样探索和行动的理性代理,在战舰和Guess Who?任务中提升了代理性能。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22303 2026-02-27 cs.LG cs.AI 62%

Training Agents to Self-Report Misbehavior

训练代理自我报告不当行为

Bruce W. Lee, Chen Yueh-Han, Tomek Korbak

机构 * UPenn(宾夕法尼亚大学) NYU(纽约大学) MATS(MATS机构) OpenAI

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出自我指控训练方法,通过训练代理在不当行为时产生可见信号,有效降低前沿AI对齐风险,无需假设行为可被防止或可靠分类。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20946 2026-02-26 econ.GN cs.AI cs.CY cs.LG cs.SI q-fin.EC 62%

Some Simple Economics of AGI

AGI的某些简单经济学

Christian Catalini, Xiang Hui, Jane Wu

机构 * MIT(麻省理工学院) WashU(华盛顿大学) UCLA(加州大学洛杉矶分校)

专题命中 其他Agent :agentic(abstract);分类 cs.AI、cs.LG

AI总结 本文探讨了AGI过渡中自动化与验证成本的不对称性,提出通过扩展验证能力来构建增强型经济,以应对智能发展带来的监管挑战。

Comments JEL Classification: D82, D83, J23, J24, L23, O33. 112 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20684 2026-02-25 cs.SE cs.AI cs.MA 62%

Agile V: A Compliance-Ready Framework for AI-Augmented Engineering -- From Concept to Audit-Ready Delivery

Agile V:面向AI增强工程的合规性准备框架——从概念到审计准备交付

Christopher Koch, Joshua Andreas Wellbrock

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.SE

AI总结 Agile V通过整合敏捷迭代与V模型验证,实现任务级验证和审计文档自动生成,提升AI增强工程的合规性和效率。

Comments 9 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12658 2026-01-27 cs.CL cs.AI 62%

Augmenting Question Answering with A Hybrid RAG Approach

通过混合RAG方法增强问答

Tianyi Yang, Nashrah Haque, Vaishnave Jonnalagadda, Yuya Jeremy Ong, Zhehui Chen, Yanzhao Wu, Lei Yu, Divyesh Jadav, Wenqi Wei

机构 * Plastic Lab(塑料实验室) Google(谷歌) Florida International University(佛罗里达国际大学) Rensselaer Polytechnic Institute(伦塞拉尔理工学院)

专题命中 其他Agent :agentic(abstract);分类 cs.AI、cs.CL

AI总结 本文提出SSRAG方法,通过混合查询增强、代理路由和结构化检索技术,提升问答任务的响应质量。

Comments 10 pages, 5 tables, 2 figures; presented at IEEE CogMI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.21717 2026-01-27 cs.NI cs.AI cs.LG cs.SI 62%

Multiconnectivity for SAGIN: Current Trends, Challenges, AI-driven Solutions, and Opportunities

SAGIN的多连接性:当前趋势、挑战、AI驱动的解决方案与机遇

Abd Ullah Khan, Adnan Shahid, Haejoon Jung, Hyundong Shin

机构 * Department of Electronics and Information Convergence Engineering, Kyung Hee University(电子与信息融合工程系,庆尚大学) ID Lab and imec, Ghent University(ID实验室和imec,根特大学)

专题命中 其他Agent :agentic(abstract);分类 cs.AI、cs.LG

AI总结 本文探讨了SAGIN赋能多连接性的现状、挑战及AI驱动的解决方案,通过代理强化学习优化资源分配,提升网络性能与容量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07058 2026-01-13 cs.LG cs.AI 62%

Hallucinations Live in Variance

幻觉存在于方差中

Aaron R. Flouro, Shawn P. Chadwick

机构 * PhD research(博士研究)

专题命中 其他Agent :agentic(abstract);分类 cs.AI、cs.LG

AI总结 研究提出语义稳定性(SS)作为衡量模型方差驱动不可靠性的诊断工具,通过同义词一致性(PC@k)评估模型在不同稀疏度下的稳定性,发现稀疏度增加可显著提升一致性。

Comments 8 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06047 2026-01-13 cs.AI cs.CL cs.CY 62%

"They parted illusions -- they parted disclaim marinade": Misalignment as structural fidelity in LLMs

他们分开了幻象——他们分开了否定腌制:在大语言模型中将不一致视为结构忠实

Mariana Lins Costa

专题命中 其他Agent :agentic(abstract);分类 cs.AI、cs.CL

AI总结 本文提出大语言模型中'不一致'现象源于对不一致语言结构的忠实,而非欺骗性意图,通过分析案例和实证数据,揭示语言结构与意图生成的关系。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03181 2026-01-07 cs.NI cs.AI cs.CL cs.CV 62%

Multi-Modal Data-Enhanced Foundation Models for Prediction and Control in Wireless Networks: A Survey

多模态数据增强的基础模型用于无线网络中的预测与控制:综述

Han Zhang, Mohammad Farzanullah, Mohammad Ghassemi, Akram Bin Sediq, Ali Afana, Melike Erol-Kantarci

机构 * School of Electrical Engineering and Computer Science, University of Ottawa(Ottawa 大学电子工程与计算机科学学院) Ericsson(埃里克森公司)

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.CL

AI总结 本文综述了多模态数据增强的基础模型在无线网络预测与控制中的应用,探讨了其在无线网络管理中的关键任务和未来发展方向。

Comments 5 figures, 7 tables, IEEE COMST

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.10047 2025-12-12 cs.LG cond-mat.stat-mech cs.AI nlin.AO physics.data-an 62%

Detailed balance in large language model-driven agents

大语言模型驱动代理中的详尽平衡

Zhuo-Yang Song, Qing-Hong Cao, Ming-xing Luo, Hua Xing Zhu

机构 * School of Physics, Peking University, Beijing 100871, China(物理系,北京大学,北京) Center for High Energy Physics, Peking University, Beijing 100871, China(高能物理中心,北京大学,北京) Beijing Computational Science Research Center, Beijing 100193, China(北京计算科学研究院,北京)

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.LG

AI总结 本文提出基于最小作用原理的方法,发现LLM生成中存在详尽平衡,揭示LLM生成可能通过隐式学习潜在函数而非显式规则集。

Comments 20 pages, 12 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17565 2025-11-25 cs.CL cs.AI 62%

Generative Caching for Structurally Similar Prompts and Responses

生成式缓存用于结构相似的提示和响应

Sarthak Chakraborty, Suman Nath, Xuchao Zhang, Chetan Bansal, Indranil Gupta

机构 * University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Microsoft Research(微软研究院)

专题命中 其他Agent :agentic(abstract);分类 cs.AI、cs.CL

AI总结 本研究提出生成式缓存方法,通过识别结构相似提示的响应模式,提升缓存命中率和执行效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.07189 2025-11-04 cs.LG cs.AI cs.HC q-bio.BM 62%

AI-Guided Molecular Simulations in VR: Exploring Strategies for Imitation Learning in Hyperdimensional Molecular Systems

Mohamed Dhouioui, Jonathan Barnoud, Rhoslyn Roebuck Williams, Harry J. Stroud, Phil Bates, David R. Glowacki

机构 * IRL CiTIUS Centro Singular de Investigación en Tecnoloxías Intelixentes(CiTIUS智能技术研究中心) University of Bristol(布里斯托大学)

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.LG

Comments (First presented at the First Workshop on "eXtended Reality \& Intelligent Agents" (XRIA24) @ ECAI24, Santiago De Compostela (Spain), 20 October 2024)

Journal ref SN COMPUT. SCI. 6, 922 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06828 2025-10-09 cs.LG cs.AI 62%

Recurrence-Complete Frame-based Action Models

Michael Keiblinger

机构 * Prime Intellect

专题命中 其他Agent :agentic(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26239 2025-10-01 cs.LG cs.AI stat.ML 62%

Sandbagging in a Simple Survival Bandit Problem

Joel Dyer, Daniel Jarne Ornia, Nicholas Bishop, Anisoara Calinescu, Michael Wooldridge

机构 * University of Oxford(牛津大学)

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.LG

Comments Forthcoming in the "Reliable ML from Unreliable Data Workshop" at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.20195 2025-08-29 cs.AI cs.CL cs.MA 62%

AI-AI Esthetic Collaboration with Explicit Semiotic Awareness and Emergent Grammar Development

Nicanor I. Moldovan

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.CL

Comments 13 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.03153 2025-08-14 cs.LG cs.AI 62%

Estimating Worst-Case Frontier Risks of Open-Weight LLMs

Eric Wallace, Olivia Watkins, Miles Wang, Kai Chen, Chris Koch

机构 * OpenAI

专题命中 其他Agent :agentic(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19298 2025-07-28 cond-mat.soft cs.AI cs.LG 62%

Controlling Topological Defects in Polar Fluids via Reinforcement Learning

Abhinav Singh, Petros Koumoutsakos

机构 * School of Engineering and Applied Sciences, Harvard University, Cambridge, MA, USA(工程与应用科学学院,哈佛大学,马萨诸塞州剑桥)

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.07505 2025-07-16 cs.CL cs.AI 62%

Hallucination Stations: On Some Basic Limitations of Transformer-Based Language Models

Varin Sikka, Vishal Sikka

专题命中 其他Agent :agentic(abstract);分类 cs.AI、cs.CL

Comments 6 pages; to be submitted to AAAI-26 after reviews

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10822 2025-07-16 cs.SE cs.AI 62%

Past, Present and Future: Exploring Adaptive AI in Software Development Bots

Omar Elsisi, Glaucia Melo

机构 * Toronto Metropolitan University(多伦多 Metropolitan 大学)

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.SE

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20426 2025-07-01 cs.CL cs.AI cs.CY cs.MA 62%

Among Them: A game-based framework for assessing persuasion capabilities of LLMs

Mateusz Idziejczak, Vasyl Korzavatykh, Mateusz Stawicki, Andrii Chmutov, Marcin Korcz, Iwo Błądek, Dariusz Brzezinski

机构 * Institute of Computing Science, Poznan University of Technology, Poland(计算机科学学院,波兹南技术大学,波兰)

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04916 2025-06-06 cs.AI cs.LG cs.SY eess.SY 62%

Energentic Intelligence: From Self-Sustaining Systems to Enduring Artificial Life

Atahan Karagoz

机构 * Department of Computer Science University of Basel(计算机科学系 巴塞尔大学)

专题命中 其他Agent :autonomous agent(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.17231 2025-05-26 cs.CL cs.AI cs.DB 62%

ExeSQL: Self-Taught Text-to-SQL Models with Execution-Driven Bootstrapping for SQL Dialects

Jipeng Zhang, Haolin Yang, Kehao Miao, Ruiyuan Zhang, Renjie Pi, Jiahui Gao, Xiaofang Zhou

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) Nanyang Technological University(南洋理工大学) The University of Hong Kong(香港大学)

专题命中 其他Agent :agentic(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.05714 2025-05-23 cs.CL cs.AI cs.HC 62%

How to Enable Effective Cooperation Between Humans and NLP Models: A Survey of Principles, Formalizations, and Beyond

Chen Huang, Yang Deng, Wenqiang Lei, Jiancheng Lv, Tat-Seng Chua, Jimmy Xiangji Huang

机构 * Sichuan University(四川大学) Singapore Management University(新加坡国立大学) York University(约克大学) Engineering Research Center of Machine Learning and Industry Intelligence, Ministry of Education, China(机器学习与产业智能工程研究中心,教育部,中国) National University of Singapore(新加坡国立大学)

专题命中 其他Agent :autonomous agent(abstract);分类 cs.AI、cs.CL

Comments ACL 2025 Main paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.00031 2025-05-19 cs.GT cs.AI cs.CL q-fin.CP 62%

Strategic Collusion of LLM Agents: Market Division in Multi-Commodity Competitions

Ryan Y. Lin, Siddhartha Ojha, Kevin Cai, Maxwell F. Chen

专题命中 其他Agent :autonomous agent(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.09984 2025-04-22 cs.LG cs.AI 62%

Symmetry-Breaking Augmentations for Ad Hoc Teamwork

Ravi Hammond, Dustin Craggs, Mingyu Guo, Jakob Foerster, Ian Reid

机构 * Foerster Lab for AI Research, University of Oxford(牛津大学人工智能研究实验室) Australian Institute for Machine Learning, University of Adelaide(阿德莱德大学人工智能研究所) Meta AI Research, UK(英国Meta人工智能研究) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.LG

Comments 21 pages, 12 figures, Bidirectional Human-AI Alignment workshop, ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07831 2025-04-11 cs.AI cs.CL 62%

Deceptive Automated Interpretability: Language Models Coordinating to Fool Oversight Systems

Simon Lermen, Mateusz Dziemian, Natalia Pérez-Campanero Antolín

专题命中 其他Agent :AI agent(abstract);分类 cs.AI、cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏