DualAD: Dual-Layer Planning for Reasoning in Autonomous Driving
专题命中 规划推理 :reasoning(title,abstract);planning(title);分类 cs.AI
Comments Autonomous Driving, Large Language Models (LLMs), Human Reasoning, Critical Scenario
AI 大模型
大模型数学、逻辑、规划、多步推理和测试时计算能力。
专题命中 规划推理 :reasoning(title,abstract);planning(title);分类 cs.AI
Comments Autonomous Driving, Large Language Models (LLMs), Human Reasoning, Critical Scenario
ForgeryVCR: 通过高效的取证工具在MLLMs中实现视觉中心推理用于图像伪造检测与定位
机构 * Shenzhen University(深圳大学) ; Tencent Youtu Lab(腾讯优图实验室)
专题命中 规划推理 :reasoning(title,abstract);CoT(abstract,abstract_cn);chain-of-thought(abstract)
AI总结 ForgeryVCR通过高效的取证工具实现视觉中心推理,提升图像伪造检测与定位的性能。
用于有效规划和工具使用的流内智能体系统优化
机构 * Stanford University(斯坦福大学) ; Texas A&M University(德克萨斯农工大学) ; UC San Diego(圣地亚哥大学) ; Lambda Website(Lambda网站)
专题命中 规划推理 :planning(title,abstract);reasoning(abstract);verifier(abstract);分类 cs.CL、cs.AI、cs.LG
AI总结 研究针对大语言模型中工具增强方法的不足,提出可训练的流内智能体框架AgentFlow及Flow-GRPO训练方法,通过协调模块优化规划器,在多基准测试中表现优异,展现流内优化优势。
Comments 47 pages, 12 figures. ICLR 2026 Oral. Project website: https://agentflow.stanford.edu/
DrugAgent:用于药物-靶点相互作用评估的冲突生物医学证据的可靠多智能体整合
机构 * Department of Computer Science and Engineering, University of Minnesota(计算机科学与工程系,明尼苏达大学) ; Computational Biology Branch, National Library of Medicine(国家医学图书馆计算生物学分支) ; Khoury College of Computer Sciences, Northeastern University(东北大学计算机科学学院) ; State Key Laboratory for Novel Software Technology at Nanjing University, School of Computer Science, Nanjing University(南京大学新型软件技术国家重点实验室,南京大学计算机科学学院) ; Developmental Therapeutics Branch, National Cancer Institute(国家癌症研究所发育治疗分支)
专题命中 规划推理 :reasoning(title,abstract);chain-of-thought(abstract);CoT(abstract);分类 cs.CL、cs.AI、cs.LG
AI总结 研究药物-靶点相互作用评估中整合异构数据问题,核心方法是基于大语言模型的多智能体系统DrugAgent,贡献是能进行异构证据支持的DTI评估,补充独立DTI预测,还提供相关策略。
推理模型拒绝发生在何处?
机构 * The Alan Turing Institute(艾伦·图灵研究所) ; University of Oxford(牛津大学) ; Northeastern University(东北大学)
专题命中 规划推理 :reasoning(title,abstract);chain-of-thought(abstract);CoT(abstract);分类 cs.CL、cs.AI、cs.LG
AI总结 研究探讨推理模型在生成响应前拒绝决策的位置,发现推理链中的初始句子对拒绝决定有显著影响,并通过激活方向分析揭示了拒绝机制。
Comments v1 accepted to the ICML 2025 Workshop on Reliable and Responsible Foundation Models (R2FM). 20 pages, 12 figures
Act2See:面向视频推理的涌现式主动视觉感知
机构 * Carnegie Mellon University(卡内基梅隆大学) ; MIT(麻省理工学院)
专题命中 规划推理 :reasoning(title,abstract);CoT(abstract,abstract_cn);chain-of-thought(abstract)
AI总结 本文提出Act2See框架,通过使VLMs在文本推理中主动交错视频帧,实现视频推理中的主动视觉感知,提升了推理质量并优于现有方法。
Comments CVPR 2026
$AutoDrive\text{-}P^3$:通过强化微调实现感知-预测-规划统一链式推理
机构 * School of Electronic and Computer Engineering, Peking University(北京大学电子与计算机工程学院)
专题命中 规划推理 :planning(title,abstract);reasoning(abstract);chain-of-thought(abstract);CoT(abstract)
AI总结 本文提出$AutoDrive\text{-}P^3$框架,通过结构化推理整合感知、预测和规划,引入$P^3\text{-}CoT$数据集和$P^3\text{-}GRPO$算法,实现端到端自动驾驶的高效决策与安全规划。
Comments Accepted at ICLR 2026 (International Conference on Learning Representations)
TraceGuard: 用于对抗大语言模型推理后门的过程引导防火墙
专题命中 规划推理 :reasoning(title,abstract);chain-of-thought(abstract);CoT(abstract);verifier(abstract)
AI总结 TraceGuard通过过程引导框架提升大语言模型的推理安全性,有效对抗推理后门。
Comments 20 pages,10 figures,6 tables
MindDriver: 引入渐进多模态推理用于自动驾驶
机构 * Amap, Alibaba Group(阿里集团蚂巴公司) ; The Hong Kong University of Science and Technology(香港科学与技术大学) ; The Chinese University of Hong Kong(香港中文大学) ; Xi’an Jiaotong University(西安交通大学)
专题命中 规划推理 :reasoning(title,abstract);chain-of-thought(abstract);CoT(abstract);planning(abstract)
AI总结 MindDriver通过渐进多模态推理框架提升自动驾驶系统的推理能力,实现语义到物理空间的转化与轨迹规划,展现优异的性能表现。
Comments CVPR2026; Yujian Yuan and Lingjun Zhang contributed equally with random order
TIR-Flow:基于冻结视频语言模型的主动视频搜索与推理
机构 * School of Electronic and Computer Engineering, Peking University(电子与计算机工程学院,北京大学)
专题命中 规划推理 :reasoning(title,abstract);chain-of-thought(abstract);CoT(abstract);logical reasoning(abstract)
AI总结 TIR-Flow通过三个协同模块实现冻结视频语言模型的主动视频搜索与推理,显著提升性能并拓展长视界视频推理能力。
ViC-Bench: 用自由式中间状态表示法对多模态大语言模型的视觉交错链式思维能力进行基准测试
机构 * School of Computer Science and Technology, Xi’an Jiaotong University(西安交通大学计算机科学与技术学院) ; Meituan Inc.(美团公司) ; Institute of Advanced Technology, University of Science and Technology of China(中国科学技术大学先进技术研究院) ; School of Computer Science and Technology, East China Normal University(华东师范大学计算机科学与技术学院) ; Research Center for Space Computing System, Zhejiang Lab(浙江实验室空间计算系统研究中心) ; College of Electronic and Information Engineering, Tongji University(同济大学电子与信息工程学院) ; School of Robotics and Automation, Nanjing University(南京大学机器人与自动化学院) ; College of Intelligent Robotics and Advanced Manufacturing, Fudan University(复旦大学智能机器人与先进制造学院)
专题命中 规划推理 :chain-of-thought(title,abstract);reasoning(abstract);CoT(abstract);planning(abstract)
AI总结 ViC-Bench 通过自由式中间状态表示法,系统评估多模态大语言模型的视觉交错链式思维能力,涵盖四个任务并提出新评估指标。
理解链式思维在代码生成中的有效性:一项经验性与信息论分析
专题命中 规划推理 :chain-of-thought(title,abstract);reasoning(abstract);CoT(abstract);planning(abstract)
AI总结 本研究通过经验分析和信息论方法,揭示了链式思维在代码生成中有效性的影响因素,发现结构化CoT在提升生成质量方面表现更优,且其效果受语言系统和模型容量影响。
OpenREAD: 基于LLM作为批评者的开放性推理端到端自动驾驶框架
机构 * Nanyang Technological University, Singapore(南洋理工大学) ; Harvard University, USA(哈佛大学)
专题命中 规划推理 :reasoning(title,abstract);chain-of-thought(abstract);CoT(abstract);planning(abstract)
AI总结 OpenREAD通过端到端强化微调框架,结合LLM作为批评者,提升自动驾驶中的推理与规划能力,实现从高层推理到低层轨迹规划的全面优化。
机构 * Michigan State University(密歇根州立大学) ; Amazon(亚马逊) ; Pennsylvania State University(宾夕法尼亚州立大学)
专题命中 规划推理 :reasoning(title,abstract);math reasoning(abstract);planning(abstract);分类 cs.CL、cs.AI、cs.LG
专题命中 规划推理 :planning(title,abstract);reasoning(abstract);chain-of-thought(abstract);分类 cs.CL、cs.AI、cs.LG
Comments Accepted at NAACL 2025 Findings
专题命中 规划推理 :planning(title,abstract);reasoning(abstract);chain-of-thought(abstract);分类 cs.CL、cs.AI、cs.LG
专题命中 规划推理 :reasoning(title,abstract);chain-of-thought(abstract);CoT(abstract);planning(abstract)
专题命中 规划推理 :reasoning(title,abstract);chain-of-thought(abstract);planning(abstract);分类 cs.CL、cs.AI、cs.LG
动态变化规范的推理与规划
机构 * University of Iowa(爱荷华大学) ; Northwestern University(西北大学)
专题命中 规划推理 :planning(title,abstract);reasoning(title);分类 cs.AI
AI总结 本文提出一种在人类-AI环境中使用动态变化规范引导规划的方法,通过可废止演算解决规范冲突并将规范作为规划护栏,理论证明与对话任务实验验证了有效性。
Comments 8 pages, 1 figure, dataset included in anc
REFLEX:基于大语言模型的反思零样本机器人规划的元认知推理
机构 * School of Applied and Creative Computing, Purdue University(应用与创意计算学院,普渡大学) ; Bowen School of Construction, Purdue University(建设学院,普渡大学) ; School of Engineering Technology, Purdue University(工程技术学院,普渡大学) ; Department of Technology Leadership and Innovation, Purdue University(技术领导与创新部门,普渡大学)
专题命中 规划推理 :planning(title,abstract);reasoning(title);分类 cs.CL
AI总结 REFLEX通过引入元认知学习,提升大语言模型在机器人任务中的推理、反思与创造力,实现零样本下的高效规划。
机构 * Department of Computer Science(计算机科学系) ; San Francisco State University(旧金山州立大学) ; Department of Mechanical & Industrial Engineering(机械与工业工程系) ; Louisiana State University(路易斯安那州立大学)
专题命中 规划推理 :planning(title,abstract);reasoning(title);分类 cs.AI
Comments Manuscript previously submitted to the NeurIPS 2025 Workshop on Bridging Language, Agent, and World Models (LAW 2025)
专题命中 规划推理 :reasoning(title);chain-of-thought(title);planning(abstract);分类 cs.AI
Comments We have decided to withdraw this paper as the work is still undergoing further refinement. To ensure the clarity of the results, we prefer to make additional improvements before resubmission. We appreciate the readers' understanding
机构 * Meta FAIR
专题命中 规划推理 :planning(title,abstract);reasoning(title);分类 cs.AI
机构 * Center for Vision, Automation & Control(视觉、自动化与控制中心) ; AIT Austrian Institute of Technology GmbH(奥地利技术研究院)
专题命中 规划推理 :planning(title,abstract);reasoning(title);分类 cs.AI
机构 * Intelligent Space Robotics Laboratory, Center for Digital Engineering, Skolkovo Institute of Science and Technology(斯克尔科沃科学与技术研究所)
专题命中 规划推理 :planning(title,abstract);reasoning(title);分类 cs.AI
Comments Submitted
专题命中 规划推理 :planning(title,abstract);reasoning(title);分类 cs.AI
Comments 10 pages, 2 figures, 1 table, AI-HRI AAAI 2022 Fall Symposium Series
面向长时域船舶轨迹与目的地预测的推理型大语言模型
机构 * Institute of High Performance Computing (IHPC), A*STAR, Singapore(新加坡科技研究局高性能计算研究所) ; The Key Laboratory of Road and Traffic Engineering, Ministry of Education, Tongji University(同济大学道路与交通工程教育部重点实验室) ; Meituan Inc., Shenzhen, China(美团(深圳)) ; Centre for Frontier AI Research (CFAR), A*STAR, Singapore(新加坡科技研究局前沿人工智能研究中心) ; Nankai University(南开大学) ; School of Artificial Intelligence, Jilin University(吉林大学人工智能学院)
专题命中 规划推理 :reasoning(title,abstract);planning(abstract);verifier(abstract);分类 cs.AI、cs.LG
AI总结 提出基于可验证奖励强化学习(RLVR)的Maritime LLM后训练框架,将轨迹转化为语义文本,通过物理有效性约束和层次匹配提升长时域(30天)预测精度,4B模型表现最优。
Comments The IEEE International Conference on Intelligent Transportation Systems (ITSC) 2026, Naples, Italy
PaT:在试后进行规划以实现高效的测试时间代码生成
机构 * Department of Computer Science and Engineering, POSTECH, South Korea(韩国POSTECH计算机科学与工程系) ; Graduate School of Artificial Intelligence, POSTECH, South Korea(韩国POSTECH人工智能研究生院) ; Microsoft Research Asia, Beijing, China(中国北京微软亚洲研究院)
专题命中 规划推理 :planning(title,abstract);reasoning(abstract);test-time compute(abstract);分类 cs.CL、cs.LG
AI总结 本文提出PaT方法,通过在试后规划来提高测试时间代码生成的效率,采用异构模型配置在多个基准和模型家族中显著提升了成本性能帕累托前沿。
Comments Accepted to ACL 2026 main conference
大型语言模型在规划问题中的最优性分析
专题命中 规划推理 :planning(title,abstract);reasoning(abstract);logical reasoning(abstract);分类 cs.CL、cs.AI
AI总结 本文研究了大型语言模型在规划问题中的最优性,通过Blocksworld领域和Path-Star图任务,发现增强推理的LLM在复杂多目标配置中显著优于传统规划器,且能精确跟踪理论最优限。
CoME:赋能移动专家的 informative hybrid-capabilities 推理
机构 * ruc(中国人民大学人工智能学院) ; xiaomi(小米公司) ; ntu(南洋理工大学) ; whu(武汉大学)
专题命中 规划推理 :reasoning(title,abstract);CoT(abstract);planning(abstract);分类 cs.CL、cs.AI
AI总结 CoME 通过四种专家架构和渐进训练策略,实现移动代理混合能力推理的解耦增强与平衡优化。