arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

2026-08-19 至 2026-08-19 共收录 3 信号源:cs.CL, cs.AI, cs.LG

1. 代码与定理证明 3 篇

2608.04457 2026-08-19 cs.DB cs.AI cs.LO 版本更新 74%

Eigenius: A Typed Knowledge-Graph DBMS with Epistemic Stratification and Institution-Mediated Reasoning

Eigenius:具备认知分层与机构介导推理的类型化知识图数据库管理系统

Hans-Martin Will, Allen L. Brown Jr., Matthew Fuchs

专题命中 代码与定理证明 :reasoning(title);分类 cs.AI

AI总结 Eigenius是一款开源类型化知识图DBMS,通过耦合类型系统等实现数据溯源不变,统一科学认识论,在Nature研究复现中验证了52个结论并发现4处差异。

Comments Minor corrections to the previous version of the manuscript

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.17127 2026-08-19 cs.CL 版本更新 57%

The Emergence of Lab-Driven Alignment Signatures: A Psychometric Framework for Auditing Latent Bias and Compounding Risk in Generative AI

实验室驱动的对齐签名的出现:一种心理测量框架,用于审计生成AI中的潜在偏见和叠加风险

Dusan Bosnjakovic

机构 * AI Researcher(人工智能研究员)

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.CL

AI总结 本文提出了一种心理测量框架,用于审计生成AI中的潜在偏见和叠加风险,通过分析九个领先模型的实验室信号,揭示了持续行为聚类的成因。

Comments v2: expanded from 9 to 18 behavioral dimensions and from 4 to 6 developer organizations; revised statistical methodology (rank-based inference with effect-size criterion, replacing variance-decomposition approach); model-level results now reported; references corrected throughout

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10818 2026-08-19 cs.AI cs.HC 版本更新 57%

LLM Enhancement with Domain Expert Mental Model to Reduce LLM Hallucination with Causal Prompt Engineering

结合领域专家心智模型的大语言模型增强:通过因果提示工程减少大语言模型幻觉

Boris Kovalerchuk, Brent D. Fegley

专题命中 代码与定理证明 :reasoning(abstract);分类 cs.AI

AI总结 本文提出结合专家心智模型(EMM)的因果提示工程框架,通过形式化相关前置流程构建EMM,减少LLM幻觉,在三类任务中验证了方法有效性。

Comments 42 pages,4 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏