arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

共收录 5792 信号源:cs.CL, cs.AI, cs.LG

1. 其他推理 5792 篇

2509.08826 2025-09-11 cs.CV 75%

RewardDance: Reward Scaling in Visual Generation

Jie Wu, Yu Gao, Zilyu Ye, Ming Li, Liang Li, Hanzhong Guo, Jie Liu, Zeyue Xue, Xiaoxia Hou, Wei Liu, Yan Zeng, Weilin Huang

机构 * ByteDance Seed(字节跳动种子)

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);CoT(abstract)

Comments Bytedance Seed Technical Report

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.12736 2025-09-05 cs.CL cs.AI cs.LG cs.SY eess.SY math.OC 75%

ACING: Actor-Critic for Instruction Learning in Black-Box LLMs

Salma Kharrat, Fares Fourati, Marco Canini

机构 * KAUST(卡斯土尼亚大学)

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Accepted at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06980 2025-07-10 cs.SE 75%

Are They All Good? Evaluating the Quality of CoTs in LLM-based Code Generation

Binquan Zhang, Li Zhang, Zhiwen Luo, Yuxin Du, Fang Liu, Song Wang, Lin Shi

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);CoT(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19184 2025-06-10 cs.CL cs.AI cs.LG 75%

When Two LLMs Debate, Both Think They'll Win

Pradyumna Shyama Prasad, Minh Nhat Nguyen

机构 * School of Computing National University of Singapore(计算机学院新加坡国立大学)

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03161 2025-05-08 cs.CR 75%

An LLM-based Self-Evolving Security Framework for 6G Space-Air-Ground Integrated Networks

Qi Qin, Xinye Cao, Guoshun Nan, Sihan Chen, Rushan Li, Li Su, Haitao Du, Qimei Cui, Pengxuan Mao, Xiaofeng Tao, Tony Q. S. Quek

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);CoT(abstract)

Comments Accepted by IEEE Communications Magazine

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.12279 2025-03-11 cs.CV 75%

HouseTune: Two-Stage Floorplan Generation with LLM Assistance

Ziyang Zong, Guanying Chen, Zhaohuan Zhan, Fengcheng Yu, Guang Tan

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);CoT(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12215 2025-03-04 cs.LG cs.AI cs.CL 75%

Revisiting the Test-Time Scaling of o1-like Models: Do they Truly Possess Test-Time Scaling Capabilities?

Zhiyuan Zeng, Qinyuan Cheng, Zhangyue Yin, Yunhua Zhou, Xipeng Qiu

专题命中 其他推理 :reasoning(abstract);CoT(abstract);分类 cs.CL、cs.AI、cs.LG

Comments Add the github link

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09058 2025-02-14 cs.IR 75%

Unleashing the Power of Large Language Model for Denoising Recommendation

Shuyao Wang, Zhi Zheng, Yongduo Sui, Hui Xiong

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);CoT(abstract)

Comments 12 pages, 5 figures, 4 tables. Accecpted by WWW 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16383 2025-02-04 cs.LG cs.AI cs.CL 75%

RotateKV: Accurate and Robust 2-Bit KV Cache Quantization for LLMs via Outlier-Aware Adaptive Rotations

Zunhai Su, Zhe Chen, Wang Shen, Hanyu Wei, Linge Li, Huangqi Yu, Kehong Yuan

专题命中 其他推理 :reasoning(abstract);CoT(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.03936 2024-08-08 cs.CL cs.AI cs.LG 75%

SLIM-RAFT: A Novel Fine-Tuning Approach to Improve Cross-Linguistic Performance for Mercosur Common Nomenclature

Vinícius Di Oliveira, Yuri Façanha Bezerra, Li Weigang, Pedro Carvalho Brom, Victor Rafael R. Celestino

专题命中 其他推理 :chain-of-thought(abstract);CoT(abstract);分类 cs.CL、cs.AI、cs.LG

Comments 13 pages, 1 figure, to be publish in International Conference on Web Information Systems and Technologies - WEBIST 2024 proceedings

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.01502 2024-05-03 cs.CL cs.AI cs.LG 75%

Analyzing the Role of Semantic Representations in the Era of Large Language Models

Zhijing Jin, Yuen Chen, Fernando Gonzalez, Jiarui Liu, Jiayi Zhang, Julian Michael, Bernhard Schölkopf, Mona Diab

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.CL、cs.AI、cs.LG

Comments NAACL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.11189 2024-01-30 cs.CL cs.AI cs.LG 75%

Diversity Measures: Domain-Independent Proxies for Failure in Language Model Queries

Noel Ngu, Nathaniel Lee, Paulo Shakarian

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.02045 2024-01-09 cs.CL cs.AI cs.LG 75%

Enhance Multi-domain Sentiment Analysis of Review Texts through Prompting Strategies

Yajing Wang, Zongwei Luo

专题命中 其他推理 :chain-of-thought(abstract);CoT(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.10472 2023-07-21 cs.CL cs.AI cs.CY cs.LG 75%

Can Instruction Fine-Tuned Language Models Identify Social Bias through Prompting?

Omkar Dige, Jacob-Junqi Tian, David Emerson, Faiza Khan Khattak

专题命中 其他推理 :chain-of-thought(abstract);CoT(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.17003 2023-03-31 cs.CL cs.AI cs.LG 75%

Evaluating GPT-3.5 and GPT-4 Models on Brazilian University Admission Exams

Desnes Nunes, Ricardo Primi, Ramon Pires, Roberto Lotufo, Rodrigo Nogueira

专题命中 其他推理 :chain-of-thought(abstract);CoT(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.11311 2022-12-23 cs.CL cs.AI cs.LG cs.SI 75%

What do LLMs Know about Financial Markets? A Case Study on Reddit Market Sentiment Analysis

Xiang Deng, Vasilisa Bashlovkina, Feng Han, Simon Baumgartner, Michael Bendersky

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.11399 2022-11-17 cs.CL cs.AI cs.LG 75%

Transcending Scaling Laws with 0.1% Extra Compute

Yi Tay, Jason Wei, Hyung Won Chung, Vinh Q. Tran, David R. So, Siamak Shakeri, Xavier Garcia, Huaixiu Steven Zheng, Jinfeng Rao, Aakanksha Chowdhery, Denny Zhou, Donald Metzler, Slav Petrov, Neil Houlsby, Quoc V. Le, Mostafa Dehghani

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);分类 cs.CL、cs.AI、cs.LG

Comments V2 has updated references/related work

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.21558 2026-07-24 cs.AI 新提交 74%

Beyond Sycophancy: Structured Resistance and Compliance in LLM Moral Reasoning

超越谄媚:大语言模型道德推理中的结构化抵抗与顺从

Baihui Wang, Bernard Koch

专题命中 其他推理 :reasoning(title);分类 cs.AI

AI总结 研究大语言模型道德推理中超越谄媚的结构化抵抗与顺从,通过三项研究揭示其判断修正沿与人类社会心理学现象平行的三个维度结构化,为区分建设性信念修正与谄媚顺从提供原则基础,助力道德交互更好对齐。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.18536 2026-07-22 cs.AI cs.RO 新提交 74%

MAGE: Human-Like Macro Placement via Agentic Multimodal Reasoning

MAGE:通过智能多模态推理实现类人宏布局

Andrew B. Kahng, Sayak Kundu, Bodhisatta Pramanik

专题命中 其他推理 :reasoning(title);分类 cs.AI

AI总结 研究针对工业物理设计流程中宏布局需大量人工优化的问题,提出MAGE框架。该框架通过多模态多智能体实现宏布局优化,结合多种规则与检查,引入量化类人性指标。实验表明其相比商业工具及基线有显著提升,且能转移到新布局设置。

Comments 16 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.00434 2026-07-02 cs.CV cs.LG 新提交 74%

Information-Regularized Attention for Visual-Centric Reasoning

信息正则化注意力用于以视觉为中心的推理

Guohao Sun, Xiaofang Wang, Yash Patel, Mengchen Liu, Zhiqiang Tao, Praveen Krishnan

机构 * FAIR at Meta(Meta 的 FAIR 部门) Rochester Institute of Technology(罗切斯特理工学院)

专题命中 其他推理 :reasoning(title);分类 cs.LG

AI总结 针对视觉语言模型中的对象幻觉、视觉基础弱和灾难性遗忘问题,提出信息正则化注意力机制,通过随机注意力显式调控视觉信息注入,改善表示学习稳定性。

Comments Accepted by ECCV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.17682 2026-06-17 cs.CL 新提交 74%

From Trainee to Trainer: LLM-Designed Training Environment for RL with Multi-Agent Reasoning

从受训者到训练者:用于多智能体推理的LLM设计的强化学习训练环境

Chao Chen, Chengzu Li, Zhiwei Li, Yinhong Liu, Zhijiang Guo

机构 * LARK, HKUST (GZ)(香港科技大学(广州)LARK实验室) University of Cambridge(剑桥大学) HKUST(香港科技大学)

专题命中 其他推理 :reasoning(title);分类 cs.CL

AI总结 提出LLM-as-Environment-Engineer框架,让策略模型自动分析失败轨迹并修改训练环境配置,在MAPF-FrozenLake测试平台上用Qwen3-4B实现最优性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.01394 2026-06-02 cs.CL 74%

UniD$^3$: A Knowledge Graph-Enhanced RAG Framework for Drug-Disease Discovery and Reasoning

UniD$^3$:一种用于药物-疾病发现与推理的知识图谱增强RAG框架

Qing Wang, Tianshi Liu, Minghao Zhou, Jialu Liang, Sen Guo, Guangyu Wang, Jing Su, Qianqian Song

机构 * Department of Health Outcomes and Biomedical Informatics, University of Florida(佛罗里达大学健康成果与生物医学信息学系) Department of Hematology, H. Lee Moffitt Cancer Center and Research Institute(血液科,H. Lee Moffitt癌症中心与研究院) Center for Bioinformatics and Computational Biology, Houston Methodist Research Institute(生物信息学与计算生物学中心,休斯顿方法主义研究学院) Department of Cardiothoracic Surgery, Weill Cornell Medicine, Cornell University(心胸外科,Weill Cornell医学,康奈尔大学) Department of Biostatistics and Health Data Science, Indiana University School of Medicine(生物统计学与健康数据科学系,印第安纳大学医学院)

专题命中 其他推理 :reasoning(title);分类 cs.CL

AI总结 提出UniD$^3$框架,结合大语言模型与知识图谱增强检索生成(KG-RAG),从生物医学文献中提取、组织和验证药物-疾病知识,生成结构化数据集并提升推理可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.09769 2026-05-13 cs.AI 74%

UTS at PsyDefDetect: Multi-Agent Councils and Absence-Based Reasoning for Defense Mechanism Classification

在PsyDefDetect上使用UTS:多智能体委员会与基于缺席的推理用于防御机制分类

Dima Galat, Marian-Andrei Rizoiu

机构 * University of Technology Sydney(技术大学悉尼)

专题命中 其他推理 :reasoning(title);分类 cs.AI

AI总结 本文提出利用DMRS对情感支持对话中的防御机制进行分类,通过多阶段 deliberative 委员会架构,采用特定类别的倡导者评估证据强度,实现F1 0.382的高精度,同时通过针对性的重写集提升F1至0.410。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.10828 2026-05-12 cs.AI 74%

The First Drop of Ink: Nonlinear Impact of Misleading Information in Long-Context Reasoning

第一滴墨水:误导信息在长上下文推理中的非线性影响

Muhan Gao, Zih-Ching Chen, Kuan-Hao Huang

机构 * Department of Computer Science Engineering, Texas A\&M University, College Station, TX, USA NVIDIA AI Technology Center, NVIDIA Corporation, Santa Clara, CA, USA

专题命中 其他推理 :reasoning(title);分类 cs.AI

AI总结 研究探讨了误导信息在长上下文推理中的非线性影响,发现误导信息比例增加时,性能在初期急剧下降,后续变化较小。通过理论和实证分析,揭示了误导信息对注意力的 disproportionate 影响,并指出过滤收益主要来自减少上下文长度而非去除误导信息。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.22855 2026-03-25 cs.AR cs.LG 74%

TorR: Towards Brain-Inspired Task-Oriented Reasoning via Cache-Oriented Algorithm-Architecture Co-design

TorR: 向基于大脑的面向任务推理迈进:通过面向缓存的算法-架构联合设计

Hyunwoo Oh, SungHeon Jeong, Suyeon Jang, Hanning Chen, Sanggeon Yun, Tamoghno Das, Mohsen Imani

机构 * Department of Computer Science, University of California, Irvine(加州大学尔湾分校计算机科学系)

专题命中 其他推理 :reasoning(title);分类 cs.LG

AI总结 TorR通过将CLIP式的密集对齐替换为超维(HDC)关联推理,实现了高效的实时推理。在算法层面引入部分相似性重用,在架构层面设计了可扩展的位切内存和轻量控制器,以满足实时性需求,同时在能耗和精度上优于现有基线。

Comments Accepted to DAC 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.01212 2026-03-03 cs.CL 74%

XAI-enhanced Comparative Opinion Mining via Aspect-based Scoring and Semantic Reasoning

基于XAI的比较意见挖掘:通过基于方面的评分和语义推理

Ngoc-Quang Le, T. Thanh-Lam Nguyen, Quoc-Trung Phu, Thi-Phuong Le, Duy-Cat Can, Hoang-Quynh Le

机构 * VNU University of Engineering(越南工程大学) Singapore Management University(新加坡管理大学) Centre Hospitalier Universitaire Vaudois(瓦乌多大学医院) University of Lausanne(洛桑大学)

专题命中 其他推理 :reasoning(title);分类 cs.CL

AI总结 本文提出XCom模型,通过基于方面的评分和语义推理提升比较性意见挖掘的可解释性与性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.07611 2026-02-19 cs.AI 74%

DIAGPaper: Diagnosing Valid and Specific Weaknesses in Scientific Papers via Multi-Agent Reasoning

DIAGPaper: 通过多智能体推理诊断科学论文中的有效且具体弱点

Zhuoyang Zou, Abolfazl Ansari, Delvin Ce Zhang, Dongwon Lee, Wenpeng Yin

机构 * Penn State University(宾夕法尼亚州立大学) University of Sheffield(谢菲尔德大学)

专题命中 其他推理 :reasoning(title);分类 cs.AI

AI总结 DIAGPaper通过多智能体推理框架有效识别科学论文中的弱点,结合定制、反驳和优先模块提升弱点识别的准确性和优先级排序。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13666 2026-02-11 cs.RO cs.AI 74%

DREAM: Domain-aware Reasoning for Efficient Autonomous Underwater Monitoring

DREAM:面向高效自主水下监测的领域感知方法

Zhenqi Wu, Abhinav Modi, Angelos Mavrogiannis, Kaustubh Joshi, Nikhil Chopra, Yiannis Aloimonos, Nare Karapetyan, Ioannis Rekleitis, Xiaomin Lin

机构 * Electrical Engineering, University of South Florida(佛罗里达州立大学电气工程系) Maryland Robotics Center (MRC), University of Maryland(马里兰大学机器人中心) Woods Hole Oceanographic Institution (WHOI)(伍兹霍尔海洋研究所) Mechanical Engineering, University of Delaware(德雷克塞尔大学机械工程系)

专题命中 其他推理 :reasoning(title);分类 cs.AI

AI总结 DREAM通过视觉语言模型引导的自主框架,实现了高效、低耗的水下长期监测,显著提升了目标物体探测效率与覆盖范围。

Comments In Proceeding of ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.17043 2026-02-09 cs.AI 74%

UniRel: Relation-Centric Knowledge Graph Question Answering with RL-Tuned LLM Reasoning

UniRel: 基于强化学习调优的LLM推理关系中心知识图谱问答

Yinxu Tang, Chengsong Huang, Jiaxin Huang, William Yeoh

机构 * Washington University in St. Louis(圣路易斯华盛顿大学)

专题命中 其他推理 :reasoning(title);分类 cs.AI

AI总结 UniRel通过强化学习调优的LLM推理框架,实现关系中心的知识图谱问答,有效识别信息丰富的子图,提升问答性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.18582 2026-01-27 cs.CL 74%

From Classification to Ranking: Enhancing LLM Reasoning Capabilities for MBTI Personality Detection

从分类到排序:提升LLM推理能力以进行MBTI性格检测

Yuan Cao, Feixiang Liu, Xinyue Wang, Yihan Zhu, Hui Xu, Zheng Wang, Qiang Qiu

专题命中 其他推理 :reasoning(title);分类 cs.CL

AI总结 本文提出将性格检测视为排序任务,通过改进的强化学习方法提升LLM在MBTI性格检测中的推理能力。

Comments 9 pages, 4 figures, AAAI 2026 Bridge

详情

展开后加载摘要…

URL PDF HTML 收藏