arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

共收录 5813 信号源:cs.CL, cs.AI, cs.LG

1. 其他推理 5813 篇

2509.23412 2025-09-30 cs.CL cs.LG 62%

Comparison of Scoring Rationales Between Large Language Models and Human Raters

Haowei Hua, Hong Jiao, Dan Song

机构 * Princeton University(普林斯顿大学) University of Maryland(马里兰大学) University of Iowa(爱荷华大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.LG

Comments 23 Pages, 4 Tables, 13 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18521 2025-09-30 cs.LG cs.AI 62%

APRIL: Active Partial Rollouts in Reinforcement Learning to Tame Long-tail Generation

Yuzhen Zhou, Jiajun Li, Yusheng Su, Gowtham Ramesh, Zilin Zhu, Xiang Long, Chenyang Zhao, Jin Pan, Xiaodong Yu, Ze Wang, Kangrui Du, Jialian Wu, Ximeng Sun, Jiang Liu, Qiaolin Yu, Hao Chen, Zicheng Liu, Emad Barsoum

机构 * Advanced Micro Devices, Inc. (AMD)(先进微器件公司) Carnegie Mellon University (CMU)(卡内基梅隆大学) LMSYS Org(LMSYS组织) University of California, Los Angeles (UCLA)(加州大学洛杉矶分校)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14783 2025-09-30 cs.LG cs.AI 62%

Omni-Thinker: Scaling Multi-Task RL in LLMs with Hybrid Reward and Task Scheduling

Derek Li, Jiaming Zhou, Leo Maxime Brunswic, Abbas Ghaddar, Qianyi Sun, Liheng Ma, Yu Luo, Dong Li, Mark Coates, Jianye Hao, Yingxue Zhang

机构 * Huawei Noah’s Ark Lab, Montréal, Canada(华为诺亚实验室,加拿大蒙特利尔) McGill University and Mila - Québec AI Institute(麦吉尔大学和魁北克人工智能研究所) Huawei Noah’s Ark Lab, Beijing, China(华为诺亚实验室,中国北京)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24357 2025-09-30 cs.LG cs.AI 62%

ReCalKV: Low-Rank KV Cache Compression via Head Reordering and Offline Calibration

Xianglong Yan, Zhiteng Li, Tianao Zhang, Haotong Qin, Linghe Kong, Yulun Zhang, Xiaokang Yang

机构 * Shanghai Jiao Tong University(上海交通大学) ETH Zürich(苏黎世联邦理工学院)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22251 2025-09-29 cs.CL cs.AI 62%

Beyond Textual Context: Structural Graph Encoding with Adaptive Space Alignment to alleviate the hallucination of LLMs

Yifang Zhang, Pengfei Duan, Yiwen Yang, Shengwu Xiong

机构 * Wuhan University of Technology(武汉理工大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

Comments 11 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21625 2025-09-29 cs.SD cs.AI cs.LG eess.AS 62%

Guiding Audio Editing with Audio Language Model

Zitong Lan, Yiduo Hao, Mingmin Zhao

机构 * University of Pennsylvania(宾夕法尼亚大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15253 2025-09-29 cs.CL cs.AI 62%

Conflict-Aware Soft Prompting for Retrieval-Augmented Generation

Eunseong Choi, June Park, Hyeri Lee, Jongwuk Lee

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

Comments Accepted to EMNLP 2025; 15 pages; 5 figures, 11 tables; Code available at https://github.com/eunseongc/CARE

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01902 2025-09-29 cs.CL cs.AI 62%

How LLMs Fail to Support Fact-Checking

Adiba Mahbub Proma, Neeley Pate, James Druckman, Gourab Ghoshal, Hangfeng He, Ehsan Hoque

机构 * University of Rochester(罗切斯特大学)

专题命中 其他推理 :chain-of-thought(abstract);分类 cs.CL、cs.AI

Comments Adiba and Neeley contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21241 2025-09-26 cs.LG cs.AI 62%

Explaining Fine Tuned LLMs via Counterfactuals A Knowledge Graph Driven Framework

Yucheng Wang, Ziyang Chen, Md Faisal Kabir

机构 * Penn State Harrisburg(宾夕法尼亚州立大学哈里斯堡分校)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

Comments 16 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18158 2025-09-26 cs.CL cs.LG 62%

ZERA: Zero-init Instruction Evolving Refinement Agent -- From Zero Instructions to Structured Prompts via Principle-based Optimization

Seungyoun Yi, Minsoo Khang, Sungrae Park

机构 * Upstage AI Research(Upstage人工智能研究院)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.LG

Comments 9 pages, 4 figures. To appear in EMNLP 2025 Main Conference (Oral Presentation)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.19215 2025-09-26 cs.LG cs.AI 62%

Strassen Attention, Split VC Dimension and Compositionality in Transformers

Alexander Kozachinskiy, Felipe Urrutia, Hector Jimenez, Tomasz Steifer, Germán Pizarro, Matías Fuentes, Francisco Meza, Cristian B. Calderon, Cristóbal Rojas

机构 * University of Chile(智利大学) CENIA IPPT PAN(波兰国家物理与天文学研究所) IMC UC(UC大学医学中心)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

Comments Accepted at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.15141 2025-09-26 cs.AI cs.LG physics.chem-ph 62%

Text-Augmented Multimodal LLMs for Chemical Reaction Condition Recommendation

Yu Zhang, Ruijie Yu, Kaipeng Zeng, Ding Li, Feng Zhu, Xiaokang Yang, Yaohui Jin, Yanyan Xu

机构 * Institute for Clarity in Documentation(清晰文档研究所) Inria Paris-Rocquencourt(巴黎-罗克琴克研究所) Rajiv Gandhi University(拉吉夫·甘地大学) Tsinghua University(清华大学) Palmer Research Laboratories(帕勒尔研究实验室)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20051 2025-09-25 cs.LG cs.AI 62%

One Filters All: A Generalist Filter for State Estimation

Shiqi Liu, Wenhan Cao, Chang Liu, Zeyu He, Tianyi Zhang, Shengbo Eben Li

机构 * School of Vehicle and Mobility, Tsinghua University(清华大学车辆与移动系统学院) College of Engineering, Peking University(北京大学工程学院) College of AI, Tsinghua University(清华大学人工智能学院)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.19593 2025-09-25 cs.CL cs.AI 62%

GuessingGame: Measuring the Informativeness of Open-Ended Questions in Large Language Models

Dylan Hutson, Daniel Vennemeyer, Aneesh Deshmukh, Justin Zhan, Tianyu Jiang

机构 * University of Cincinnati(辛辛那提大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

Comments EMNLP 2025, 17 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18123 2025-09-24 cs.AI cs.LG 62%

SPADE: A Large Language Model Framework for Soil Moisture Pattern Recognition and Anomaly Detection in Precision Agriculture

Yeonju Lee, Rui Qi Chen, Joseph Oboamah, Po Nien Su, Wei-zhen Liang, Yeyin Shi, Lu Gan, Yongsheng Chen, Xin Qiao, Jing Li

机构 * organization= H. Milton Stewart School of Industrial Systems Engineering, Georgia Institute of Technology , city= Atlanta , state= GA , country= USA organization= Panhandle Research Extension Center, University of Nebraska-Lincoln , city= Scottsbluff , state= NE , country= USA organization= Department of Computer Science Engineering, University of Nebraska-Lincoln , city= Lincoln , state= NE , country= USA organization= Department of Biological Systems Engineering, University of Nebraska-Lincoln , city= Lincoln , state= NE , country= USA organization= Institute for Robotics Intelligent Machines, Georgia Institute of Technology , city= Atlanta , state= GA , country= USA organization= School of Civil \& Environmental Engineering, Georgia Institute of Technology, Georgia Institute of Technology , city= Atlanta , state= GA , country= USA

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18113 2025-09-24 cs.CL cs.LG 62%

Dynamic Prompt Fusion for Multi-Task and Cross-Domain Adaptation in LLMs

Xin Hu, Yue Kang, Guanzi Yao, Tianze Kang, Mengjie Wang, Heyao Liu

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.18156 2025-09-24 cs.CL cs.AI 62%

Can LLMs Explain Themselves Counterfactually?

Zahra Dehghanighobadi, Asja Fischer, Muhammad Bilal Zafar

机构 * Ruhr University Bochum(博尔塔伦大学博赫姆分校) UAR Research Center for Trustworthy Data Science and Security(UAR可信数据科学与安全研究中心)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.17765 2025-09-23 cs.CL cs.AI cs.CV eess.AS 62%

Qwen3-Omni Technical Report

Jin Xu, Zhifang Guo, Hangrui Hu, Yunfei Chu, Xiong Wang, Jinzheng He, Yuxuan Wang, Xian Shi, Ting He, Xinfa Zhu, Yuanjun Lv, Yongqi Wang, Dake Guo, He Wang, Linhan Ma, Pei Zhang, Xinyu Zhang, Hongkun Hao, Zishan Guo, Baosong Yang, Bin Zhang, Ziyang Ma, Xipin Wei, Shuai Bai, Keqin Chen, Xuejing Liu, Peng Wang, Mingkun Yang, Dayiheng Liu, Xingzhang Ren, Bo Zheng, Rui Men, Fan Zhou, Bowen Yu, Jianxin Yang, Le Yu, Jingren Zhou, Junyang Lin

机构 * Qwen Team(通义实验室)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

Comments https://github.com/QwenLM/Qwen3-Omni

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.17730 2025-09-23 cs.LG cs.CL 62%

ConfClip: Confidence-Weighted and Clipped Reward for Reinforcement Learning in LLMs

Bonan Zhang, Zhongqi Chen, Bowen Song, Qinya Li, Fan Wu, Guihai Chen

机构 * Shanghai Jiao Tong University Ant Group(上海交通大学-蚂蚁集团)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16633 2025-09-23 cs.CV cs.AI cs.CL 62%

When Big Models Train Small Ones: Label-Free Model Parity Alignment for Efficient Visual Question Answering using Small VLMs

Abhirama Subramanyam Penamakuri, Navlika Singh, Piyush Arora, Anand Mishra

机构 * Indian Institute of Technology Jodhpur(印度理工学院朱诺尔)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

Comments Accepted to EMNLP (Main) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16297 2025-09-23 cs.CY cs.AI cs.CL 62%

How Large Language Models are Designed to Hallucinate

Richard Ackermann, Simeon Emanuilov

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

Comments 23 pages, 2 tables, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15557 2025-09-22 cs.LG cs.AI 62%

Reward Hacking Mitigation using Verifiable Composite Rewards

Mirza Farhan Bin Tarek, Rahmatollah Beheshti

机构 * University of Delaware(特拉华大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

Comments Accepted at the 16th ACM Conference on Bioinformatics, Computational Biology, and Health Informatics (ACM-BCB 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.20422 2025-09-22 cs.CL cs.AI 62%

SEMMA: A Semantic Aware Knowledge Graph Foundation Model

Arvindh Arun, Sumit Kumar, Mojtaba Nayyeri, Bo Xiong, Ponnurangam Kumaraguru, Antonio Vergari, Steffen Staab

机构 * Institute for AI, University of Stuttgart(斯图加特大学人工智能研究所) IIIT Hyderabad(海得拉巴国家理工学院) Stanford University(斯坦福大学) University of Edinburgh(爱丁堡大学) University of Southampton(南安普顿大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

Comments EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18475 2025-09-22 cs.LG cs.AI 62%

A Survey of Large Language Models for Data Challenges in Graphs

Mengran Li, Pengyu Zhang, Wenbin Xing, Yijia Zheng, Klim Zaporojets, Junzhou Chen, Ronghui Zhang, Yong Zhang, Siyuan Gong, Jia Hu, Xiaolei Ma, Zhiyuan Liu, Paul Groth, Marcel Worring

机构 * Guangdong Key Laboratory of Intelligent Transportation System, School of Intelligent Systems Engineering, Shenzhen Campus of Sun Yat-sen University(广东智能交通系统重点实验室,智能系统工程学院,中山大学深圳校区) University of Amsterdam(阿姆斯特丹大学) Aarhus University(阿arhus大学) Beijing Institute of Artificial Intelligence, Beijing University of Technology(北京人工智能研究院,北京工业大学) School of Information and Engineering, Chang’an University(信息工程学院,长安大学) Key Laboratory of Road and Traffic Engineering of the Ministry of Education, Tongji University(交通工程教育部长实验室,同济大学) Key Laboratory of Intelligent Transportation Technology and System, School of Transportation Science and Engineering, Beihang University(智能交通技术与系统重点实验室,交通运输科学与工程学院,北航) Jiangsu Key Laboratory of Urban ITS, Jiangsu Province Collaborative Innovation Center of Modern Urban Traffic Technologies, School of Transportation, Southeast University(江苏城市ITS重点实验室,江苏省现代城市交通技术协同创新中心,交通学院,东南大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

Comments Accepted by Expert Systems with Applications

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19546 2025-09-18 cs.CL cs.AI 62%

Language Models Identify Ambiguities and Exploit Loopholes

Jio Choi, Mohit Bansal, Elias Stengel-Eskin

机构 * UNC Chapel Hill(北卡罗来纳大学教堂山分校) The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

Comments EMNLP 2025 camera-ready; Code: https://github.com/esteng/ambiguous-loophole-exploitation

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12754 2025-09-17 cs.RO cs.AI cs.HC cs.LG 62%

Toward Ownership Understanding of Objects: Active Question Generation with Large Language Model and Probabilistic Generative Model

Saki Hashimoto, Shoichi Hasegawa, Tomochika Ishikawa, Akira Taniguchi, Yoshinobu Hagiwara, Lotfi El Hafi, Tadahiro Taniguchi

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

Comments Submitted to AROB-ISBC 2026 (Journal Track option)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10753 2025-09-16 cs.LG cs.AI 62%

HalluField: Detecting LLM Hallucinations via Field-Theoretic Modeling

Minh Vu, Brian K. Tran, Syed A. Shah, Geigh Zollicoffer, Nhat Hoang-Xuan, Manish Bhattarai

机构 * Center for Nonlinear Studies(非线性研究中心) Los Alamos National Laboratory(洛斯阿拉莫斯国家实验室) T-1 Los Alamos National Laboratory(洛斯阿拉莫斯国家实验室T-1部门) Applied Mathematics University of Colorado Boulder(科罗拉多大学博尔德分校应用数学系) T-4 Los Alamos National Laboratory(洛斯阿拉莫斯国家实验室T-4部门)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09931 2025-09-16 cs.LG cs.AI 62%

Mechanistic Interpretability of LoRA-Adapted Language Models for Nuclear Reactor Safety Applications

Yoon Pyo Lee

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

Comments Accepted for publication in Nuclear Technology. 24 pages, 2 tables, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.13632 2025-09-15 cs.LG cs.AI cs.CV cs.DC cs.SE 62%

TraceFL: Interpretability-Driven Debugging in Federated Learning via Neuron Provenance

Waris Gill, Ali Anwar, Muhammad Ali Gulzar

机构 * Computer Science Department Virginia Tech Blacksburg, USA(弗吉尼亚理工大学计算机科学系) Engineering Department University of Minnesota Minneapolis, USA(明尼苏达大学工程系)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI、cs.LG

Comments Accepted at 2025 IEEE/ACM 47th International Conference on Software Engineering (ICSE)

Journal ref 2025 IEEE/ACM 47th International Conference on Software Engineering (ICSE), pp. 2264--2276

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.09174 2025-09-12 cs.CL cs.AI cs.SD 62%

EchoX: Towards Mitigating Acoustic-Semantic Gap via Echo Training for Speech-to-Speech LLMs

Yuhao Zhang, Yuhao Du, Zhanchen Dai, Xiangnan Ma, Kaiqi Kou, Benyou Wang, Haizhou Li

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏