Learning to Win by Reading Manuals in a Monte-Carlo Framework
专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.CL、cs.LG
Journal ref Journal Of Artificial Intelligence Research, Volume 43, pages 661-704, 2012
AI 大模型
智能体、工具调用、规划、工作流、多智能体和自主任务执行。
专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.CL、cs.LG
Journal ref Journal Of Artificial Intelligence Research, Volume 43, pages 661-704, 2012
专题命中 工具调用 :agent(abstract);planning(abstract)
Comments In Proceedings of the 3rd IROS Workshop on Robots and Sensors integration in future rescue INformation system (ROSIN), Tokyo, Japan, November 7, 2013
专题命中 工具调用 :agent(abstract);multi-agent(abstract)
Comments Appears in Proceedings of the Twentieth Conference on Uncertainty in Artificial Intelligence (UAI2004)
专题命中 工具调用 :agent(abstract);tool use(abstract)
专题命中 工具调用 :planning(abstract);workflow(abstract)
专题命中 工具调用 :tool use(abstract);planning(abstract)
Comments Published in the Proceedings of the "I Workshop of Astronomy and Astrophysics for Students", Eds. N.R. Napolitano & M. Paolillo, Naples, 19-20 April 2006 (astro-ph/0701577)
专题命中 工具调用 :agent(abstract);multi-agent(abstract)
Comments 12 pages, 4 figures, draft of a paper submitted to Artificial Economics 2005, September 15-16, Lille, France
模式引导对话系统与模型上下文协议的收敛
机构 * SBB-IT
专题命中 工具调用 :agent(abstract,comments);分类 cs.AI、cs.CL
AI总结 本文通过分析模式引导对话系统与模型上下文协议的收敛,提出五条模式设计原则,揭示了SGD与MCP的内在联系及软件3.0中的关键监督机制。
Comments 18 sections, 4 figures, 7 tables, 40 references. Original research presenting: (1) formal framework mapping Schema-Guided Dialogue principles to Model Context Protocol concepts, (2) five foundational design principles for LLM-native schema authoring, (3) architectural patterns for secure, scalable agent orchestration. Research supported by SBB (Swiss Federal Railways)
地球观测卫星调度与图神经网络及蒙特卡洛树搜索
专题命中 工具调用 :planning(abstract,comments);分类 cs.AI、cs.LG
AI总结 本文提出利用图神经网络和深度强化学习结合蒙特卡洛树搜索,解决地球观测卫星调度问题,以提高调度效率和效益。
Comments Accepted at International Workshop on Planning & Scheduling for Space (IWPSS 2025)
专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG;planning(comments)
Comments Extended version of paper accepted at The International Conference on Automated Planning and Scheduling (ICAPS) 2025
专题命中 工具调用 :planning(abstract,comments);分类 cs.AI、cs.LG
Comments presented at the ICAPS 2024 workshop on Bridging the Planning and Reinforcement Learning
专题命中 工具调用 :planning(abstract,journal_ref);分类 cs.AI、cs.LG
Journal ref AAAI - ICML / IJCAI / AAMAS 2018 Workshop on Planning and Learning (PAL-18). Stockholm, Sweden 2018
专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG;planning(journal_ref)
Journal ref Journal of Water Resources Planning and Management, Volume 146, Issue 11 (November 2020)
基于大语言模型的多源强化学习状态编码复合神经架构搜索
机构 * Shanghai Jiao Tong University(上海交通大学) ; Cornell University(康奈尔大学) ; New York University(纽约大学)
专题命中 工具调用 :agent(abstract,comments);分类 cs.LG;planning(comments)
AI总结 本文提出基于大语言模型的复合神经架构搜索方法,用于多源强化学习状态编码,通过高效搜索发现更高性能的编码器架构。
Comments NeurIPS 2025 Workshop on Bridging Language, Agent, and World Models for Reasoning and Planning
专题命中 工具调用 :tool use(abstract);分类 cs.AI;agent(comments);multi-agent(comments)
Comments To appear in Proceedings of 11th International Workshop on Computational Logic in Multi-Agent Systems (CLIMA XI 2010)
利用生成式幻觉和生物物理引导建模实现统一的生物分子序列-结构协同设计
专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG
AI总结 研究针对DNA/RNA等生物分子设计难题,提出MCTH框架,通过蒙特卡洛树搜索实现序列-结构协同设计,在多模态设计任务中性能优于基线,且可跨模态泛化。
Search-G1:基于表征内在奖励的接地搜索智能体
专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.CL
AI总结 该研究提出Search-G1框架,通过两个经干预校准的读数构成的表征内在奖励,改善了搜索增强语言智能体的接地性与搜索成本的权衡,在多基准和模型规模上验证了其有效性。
倒排索引遍历的$\mathbf{P}$-完备性:关于布尔查询DAG评估的复杂性
机构 * Apple Inc.(苹果公司)
专题命中 工具调用 :AI agent(abstract);分类 cs.AI、cs.CL
AI总结 本文证明基于DAG的布尔查询评估问题是$\mathbf{P}$-完备的,并提出一种稀疏感知算法ComputePN,通过正负对偶表示和DAG记忆化,将评估时间严格限制在$O(|Q| \cdot |U_{\mathit{active}}|)$,避免指数爆炸和全量扫描。
帕累托条件强化学习的命令空间反事实解释
机构 * Centrum Wiskunde & Informatica (CWI)(数学和计算机科学中心(CWI)) ; Eindhoven University of Technology(埃因霍温理工大学)
专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.LG
AI总结 该研究针对帕累托条件强化学习的帕累托条件网络,提出命令空间反事实解释方法CF-ZOO,通过边界种子方向搜索生成可操作的直观解释。
Comments 11 pages, 3 figures
Journal ref Proceedings of the IJCAI-ECAI 2026 Workshop on Explainable Artificial Intelligence (XAI), Bremen, Germany, August 2026
PandasCorpus:真实世界Pandas工作流与使用模式资源
专题命中 工具调用 :workflow(abstract);分类 cs.LG、cs.SE
AI总结 该研究构建了包含13.9万个Jupyter笔记本的PandasCorpus数据集,分析2015-2025年Pandas工作流特征,为相关研究提供公开资源。
LatentSkill: 从上下文文本技能到LLM智能体的权重内隐技能
机构 * Shanghai Jiao Tong University(上海交通大学) ; Sun Yat-Sen University(中山大学) ; Shanghai Innovation Institute(上海创新研究院) ; OPPO Research Institute(OPPO研究院)
专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.CL
AI总结 提出LatentSkill框架,通过预训练超网络将文本技能转换为即插即用的LoRA适配器,将技能知识存储在权重空间而非上下文空间,从而减少预填充令牌并提升性能。
Comments 10 pages, 4 figures
从提示到构念:面向心理学研究的大语言模型双效度框架
专题命中 工具调用 :tool use(abstract);分类 cs.AI、cs.CL
AI总结 本研究针对大语言模型应用于心理学研究时的测量幻影问题,构建了双效度框架,明确不同研究目标对应的证据要求,强调需开发心理构念的计算类似物而非直接套用人类测量工具。
Journal ref Annu. Rev. Psychol. 2027. 78:2.1-2.26
StorySpark:面向故事前提生成的模块级进化搜索
机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) ; Renmin University of China(中国人民大学) ; Shandong University(山东大学) ; Institute of Deep Perception Technology, JITRI(JITRI深度感知技术研究院)
专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.CL
AI总结 针对LLM故事生成中前提构思不足的问题,提出StorySpark模块级进化搜索框架,通过优化叙事模块生成更优质的故事前提,提升下游故事质量与创意性。
Comments 26 pages, 7 figures
AgentSnare:学习延迟、转移和瓦解自主渗透智能体
专题命中 工具调用 :agent(abstract);分类 cs.CL、cs.LG
AI总结 AgentSnare是一种轨迹自适应欺骗系统,通过动态构建诱饵环境,吸收渗透智能体工具调用、转移其轨迹并瓦解攻击,在CVE-Bench实验中成功阻止了所有真实目标被利用。
从专有大语言模型API窃取推理轨迹
专题命中 工具调用 :agentic(abstract);分类 cs.AI、cs.LG
AI总结 该研究发现专有LLM API的加密推理轨迹存在跨会话、用户及模型的兼容性漏洞,开发了可扩展解密越狱方法,能提取模型推理、窃取PII等,还提出了对应缓解措施。
涌现不变性:从符号化思维到结构控制
专题命中 工具调用 :tool use(abstract);分类 cs.AI、cs.LG
AI总结 该研究形式化了语言为核心的智能的限制,提出动态边界控制框架,通过DeepSeek V4-Flash实验验证其可提升大语言模型的推理性能,组织了大语言模型的涌现限制。
Comments 51 pages, 3 figures
ToolUniverse:一个让AI科学家开发民主化的开放平台
机构 * Department of Biomedical Informatics, Harvard Medical School(哈佛医学院生物医学信息学系) ; Harvard College, Harvard University(哈佛大学哈佛学院) ; Massachusetts Institute of Technology(麻省理工学院) ; MIT Lincoln Laboratory(MIT林肯实验室) ; Kempner Institute for the Study of Natural and Artificial Intelligence, Harvard University(哈佛大学自然与人工智能研究学院) ; Broad Institute of MIT and Harvard(MIT与哈佛联合广谱研究所) ; Harvard Data Science Initiative(哈佛数据科学计划)
专题命中 工具调用 :agentic(abstract);分类 cs.AI、cs.LG
AI总结 ToolUniverse是支持从各类模型构建AI科学家的开放平台,通过AI-工具交互标准规范工具调用,已应用于大量科学工具,案例中可完成多领域端到端分析。
Comments https://aiscientist.tools
面包屑式搜索智能体
专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.CL
AI总结 该研究针对LLM搜索智能体的安全问题,提出ACH和TGSE策略,通过操纵搜索结果形成连贯证据链,大幅提升攻击成功率。
Comments 38 pages, 7 figures
思维树作为经典启发式搜索问题:形式化基础与设计模式
机构 * Guni Sharon
专题命中 工具调用 :planning(abstract);分类 cs.AI、cs.LG
AI总结 本文通过经典启发式搜索术语统一分类法,将基于LLM的推理映射到搜索组件,并识别出系统搜索和前瞻性策略两种设计模式。
Comments Extended version of the SoCS 2026 paper. Includes appendices omitted from the proceedings version
Journal ref Proceedings of the Nineteenth International Symposium on Combinatorial Search (SoCS 2026), AAAI Press, 2026
VibeSearchBench:野外长期主动搜索的基准测试
机构 * Xiaohongshu Dots Studio & Unipat AI(小红书 dots 飞 studios 与 Unipat AI)
专题命中 工具调用 :agent(abstract);分类 cs.AI、cs.CL
AI总结 针对现有搜索基准中查询过于明确、单轮交互和固定模式评估导致用户体验与评估结果差距的问题,提出VibeSearch范式并构建VibeSearchBench基准,通过渐进式用户模拟和图匹配评估框架测试前沿模型,发现所有模型在长期上下文推理、主动意图激发和结构化知识构建方面仍存在显著不足。