arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

共收录 45051 信号源:cs.CL, cs.AI, cs.LG

1. 代码与定理证明 1129 篇

2412.16725 2025-11-14 cs.AI 57%

Enhancing Conflict Resolution in Language Models via Abstract Argumentation

Zhaoqun Li, Xiaotong Fang, Chen Chen, Mengze Li, Beishui Liao

机构 * School of Philosophy, Zhejiang University(浙江大学哲学系) College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) The State Key Lab of Brain-Machine Intelligence(脑机智能国家重点实验室)

专题命中 代码与定理证明 :chain-of-thought(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.09092 2025-11-13 cs.AI math.OC 57%

OR-R1: Automating Modeling and Solving of Operations Research Optimization Problem via Test-Time Reinforcement Learning

Zezhen Ding, Zhen Tan, Jiheng Zhang, Tianlong Chen

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

Comments 9 pages, 5 figures, AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07690 2025-11-12 cs.AI 57%

Towards AI-Assisted Generation of Military Training Scenarios

Soham Hans, Volkan Ustun, Benjamin Nye, James Sterrett, Matthew Green

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07083 2025-11-11 cs.AI 57%

Increasing AI Explainability by LLM Driven Standard Processes

Marc Jansen, Marcel Pehlke

机构 * Computer Science Institute University of Applied Sciences Ruhr West Bottrop(应用科学鲁尔西贝特大学计算机科学研究所)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04706 2025-11-10 cs.CY cs.AI cs.HC 57%

Prioritize Economy or Climate Action? Investigating ChatGPT Response Differences Based on Inferred Political Orientation

Pelin Karadal, Dilara Kekulluoglu

机构 * Sabanci University(萨班奇大学)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03108 2025-11-06 cs.AI 57%

miniF2F-Lean Revisited: Reviewing Limitations and Charting a Path Forward

Azim Ospanov, Farzan Farnia, Roozbeh Yousefzadeh

机构 * Huawei Hong Kong Research Center(华为香港研究中心) Department of Computer Science & Engineering, The Chinese University of Hong Kong(香港中文大学计算机科学与工程系)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23081 2025-11-05 cs.CL 57%

A Survey on LLM Mid-Training

Chengying Tu, Xuemiao Zhang, Rongxiang Weng, Rumei Li, Chen Zhang, Yang Bai, Hongfei Yan, Jingang Wang, Xunliang Cai

机构 * Peking University(北京大学) Meituan(美团)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00125 2025-11-04 cs.SE cs.AI cs.LO cs.PL 57%

Inferring multiple helper Dafny assertions with LLMs

Álvaro Silva, Alexandra Mendes, Ruben Martins

机构 * INESC TEC, Faculty of Engineering, University of Porto(葡萄牙波尔图大学工程学院INESC TEC) Computer Science Department of Carnegie Mellon University(卡内基梅隆大学计算机科学系)

专题命中 代码与定理证明 :verifier(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.20621 2025-11-03 cs.AI 57%

Towards the Formalization of a Trustworthy AI for Mining Interpretable Models explOiting Sophisticated Algorithms

Riccardo Guidotti, Martina Cinquini, Marta Marchiori Manerba, Mattia Setzu, Francesco Spinnato

机构 * University of Pisa(比萨大学)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21717 2025-10-28 cs.HC cs.AI cs.SE 57%

AI-Enhanced Operator Assistance for UNICOS Applications

Bernard Tam, Jean-Charles Tournier, Fernando Varela Rodriguez

机构 * The University of Sydney(悉尼大学)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

Comments Prepared as part of the CERN openlab programme 2025. Also available on Zenodo, a repository operated by CERN and co-funded by the European Union

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06239 2025-10-22 cs.AI 57%

Proof2Silicon: Prompt Repair for Verified Code and Hardware Generation via Reinforcement Learning

Manvi Jha, Jiaxin Wan, Deming Chen

机构 * Electrical and Computer Engineering(电气与计算机工程系) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 代码与定理证明 :verifier(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17919 2025-10-22 cs.CR cs.AI 57%

ParaVul: A Parallel Large Language Model and Retrieval-Augmented Framework for Smart Contract Vulnerability Detection

Tenghui Huang, Jinbo Wen, Jiawen Kang, Siyong Chen, Zhengtao Li, Tao Zhang, Dongning Liu, Jiacheng Wang, Chengjun Cai, Yinqiu Liu, Dusit Niyato

机构 * School of Automation, Guangdong University of Technology(广东科技大學自動化學院) College of Computer Science and Technology, Nanjing University of Aeronautics and Astronautics(南京航空航天大學計算機科學與技術學院) School of Cyberspace Science and Technology, Beijing Jiaotong University(北京交通大學網絡空間科學與技術學院) School of Computer Science and Technology, Guangdong University of Technology(廣東科技大學計算機科學與技術學院) College of Computing and Data Science, Nanyang Technological University(南洋理工大學計算與數據科學學院)

专题命中 代码与定理证明 :chain-of-thought(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14252 2025-10-17 cs.CL 57%

MoM: Mixtures of Scenario-Aware Document Memories for Retrieval-Augmented Generation Systems

Jihao Zhao, Zhiyuan Ji, Simin Niu, Hanyu Wang, Feiyu Xiong, Zhiyu Li

机构 * School of Information, Renmin University of China, Beijing, China(中国人民大学信息学院) MemTensor (Shanghai) Technology Co., Ltd.(MemTensor(上海)技术有限公司) Institute for Advanced Algorithms Research, Shanghai(上海先进算法研究所)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03178 2025-10-06 cs.SE cs.CL 57%

When Names Disappear: Revealing What LLMs Actually Understand About Code

Cuong Chi Le, Minh V. T. Pham, Cuong Duc Van, Hoang N. Phan, Huy N. Phan, Tien N. Nguyen

机构 * FPT Software AI Center(FPT软件AI中心) Nanyang Technological University(南洋理工大学) University of Texas at Dallas(德克萨斯大学达拉斯分校)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.00732 2025-10-02 cs.AI 57%

EvolProver: Advancing Automated Theorem Proving by Evolving Formalized Problems via Symmetry and Difficulty

Yuchen Tian, Ruiyuan Huang, Xuanwu Wang, Jing Ma, Zengfeng Huang, Ziyang Luo, Hongzhan Lin, Da Zheng, Lun Du

机构 * Hong Kong Baptist University(香港 Baptist 大学) Ant Group(蚂蚁集团) School of Data Science, Fudan University(复旦大学数据科学学院)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24127 2025-09-30 cs.AI cs.DB 57%

Transparent, Evaluable, and Accessible Data Agents: A Proof-of-Concept Framework

Nooshin Bahador

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

Comments 20 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22834 2025-09-30 cs.NI cs.AI 57%

Bridging Language Models and Formal Methods for Intent-Driven Optical Network Design

Anis Bekri, Amar Abane, Abdella Battou, Saddek Bensalem

机构 * National Institute of Standards and Technology(美国国家标准技术研究院)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

Comments Accepted at AICCSA 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20182 2025-09-25 cs.AR cs.AI 57%

Automated Multi-Agent Workflows for RTL Design

Amulya Bhattaram, Janani Ramamoorthy, Ranit Gupta, Diana Marculescu, Dimitrios Stamoulis

机构 * Chandra Family Department of Electrical and Computer Engineering(查克拉家族电子与计算机工程系)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

Comments Accepted: ML for Systems Workshop NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01830 2025-09-23 cs.CL 57%

From Language to Cognition: How LLMs Outgrow the Human Language Network

Badr AlKhamissi, Greta Tuckute, Yingtian Tang, Taha Binhuraib, Antoine Bosselut, Martin Schrimpf

机构 * EPFL(苏黎世联邦理工学院) MIT(麻省理工学院) Georgia Institute of Technology(佐治亚理工学院)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.CL

Comments EMNLP 2025. Project Page at https://language-to-cognition.epfl.ch

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16533 2025-09-23 cs.CL 57%

Challenging the Evaluator: LLM Sycophancy Under User Rebuttal

Sungwon Kim, Daniel Khashabi

机构 * Johns Hopkins University(约翰霍普金斯大学)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.CL

Comments Accepted to EMNLP 2025 Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.15195 2025-09-19 cs.SE cs.AI cs.CR 57%

Orion: Fuzzing Workflow Automation

Max Bazalii, Marius Fleischer

机构 * NVIDIA

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

Comments 11 pages, 3 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.13597 2025-09-18 cs.CR cs.AI 57%

Agentic JWT: A Secure Delegation Protocol for Autonomous AI Agents

Abhishek Goswami

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

Comments 17 pages, 6 figures, 2 Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.11131 2025-09-16 cs.AI cs.MA q-bio.OT 57%

Neural cellular automata: applications to biology and beyond classical AI

Benedikt Hartl, Michael Levin, Léo Pio-Lopez

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07026 2025-09-10 cs.LO cs.AI 57%

Contradictions

Yang Xu, Shuwei Chen, Xiaomei Zhong, Jun Liu, Xingxing He

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

Comments 37 Pages,9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00110 2025-09-03 cs.CY cs.AI 57%

The Application of Virtual Environments and Artificial Intelligence in Higher Education: Experimental Findings in Philosophy Teaching

Adel Vehrer, Zsolt Palfalusi

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11302 2025-08-29 cs.CL 57%

Are formal and functional linguistic mechanisms dissociated in language models?

Michael Hanna, Yonatan Belinkov, Sandro Pezzelle

机构 * Institute for Logic, Language and Computation University of Amsterdam(逻辑、语言与计算研究所 阿姆斯特丹大学) Technion – Israel Institute of Technology(技术ion-以色列理工学院)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.CL

Comments To appear in Computational Linguistics. Pre-MIT Press publication version. 40 pages, 14 figures, 3 tables. Code available at https://github.com/hannamw/formal-functional-dissociation

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.16245 2025-08-25 cs.GT cs.LG cs.MA econ.TH 57%

Limit-Computable Grains of Truth for Arbitrary Computable Extensive-Form (Un)Known Games

Cole Wyeth, Marcus Hutter, Jan Leike, Jessica Taylor

机构 * David R. Cheriton School of Computer Science, University of Waterloo(多伦多大学大卫·R·切里顿计算机科学学院) Google DeepMind and Australian National University(谷歌DeepMind和澳大利亚国立大学) Anthropic(Anthropic公司) Median Group(Median集团)

专题命中 代码与定理证明 :planning(abstract);分类 cs.LG

Comments 42 pages; 2 figures; 7 algorithms

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14927 2025-08-22 cs.GT cs.AI 57%

AI Testing Should Account for Sophisticated Strategic Behaviour

Vojtech Kovarik, Eric Olav Chen, Sami Petersen, Alexis Ghersengorin, Vincent Conitzer

机构 * Department of Computer Science(计算机科学系) Czech Technical University Prague(捷克技术大学布拉格) Global Priorities Institute(全球优先研究所) University of Oxford(牛津大学) Foundations of Cooperative AI Lab(合作人工智能基础实验室) Carnegie Mellon University(卡内基梅隆大学)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14644 2025-08-21 cs.AI 57%

LeanGeo: Formalizing Competitional Geometry problems in Lean

Chendong Song, Zihan Wang, Frederick Pu, Haiming Wang, Xiaohan Lin, Junqi Liu, Jia Li, Zhengying Liu

机构 * Moonshot AI Numina Peking University(北京大学) University of Toronto(多伦多大学)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

Comments 28 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12416 2025-08-19 cs.HC cs.AI 57%

fCrit: A Visual Explanation System for Furniture Design Creative Support

Vuong Nguyen, Gabriel Vigliensoni

机构 * Concordia University(康科德大学)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

Comments In Proceedings of Explainable AI for the Arts Workshop 2025 (XAIxArts 2025) arXiv:2406.14485

详情

展开后加载摘要…

URL PDF HTML 收藏