arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12169 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12169 篇

2511.21731 2026-06-03 cs.CL cs.AI 82%

Identifying Quantum Structure in AI Language: Evidence for Evolutionary Convergence of Human and Artificial Cognition

识别AI语言中的量子结构:人类与人工智能认知进化趋同的证据

Diederik Aerts, Jonito Aerts Arguëlles, Lester Beltran, Suzette Geriente, Roberto Leporini, Massimiliano Sassoli de Bianchi, Sandro Sozzo

机构 * Center Leo Apostel for Interdisciplinary Studies, Vrije Universiteit Brussel (VUB)(利奥·阿波斯泰尔跨学科研究中心,布鲁塞尔自由大学) Department of Economics, University of Bergamo(博洛尼亚大学经济系) Department of Humanities and Cultural Heritage (DIUM) and Centre CQSCS, University of Udine(乌迪内大学人文与文化遗产系及CQSCS中心)

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 通过对大型语言模型进行认知测试,发现其概念组合中存在贝尔不等式显著违背和玻色-爱因斯坦统计,表明人类与人工智能在概念-语言领域均涌现非经典量子结构,支持认知进化趋同假说。

Journal ref Entropy 28, 622, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15198 2026-05-28 cs.MA cs.AI cs.CL 82%

Colosseum: Auditing Collusion in Cooperative Multi-Agent Systems

Colosseum: 审计合作多智能体系统中的合谋行为

Mason Nakamura, Abhinav Kumar, Saswat Das, Sahar Abdelnabi, Saaduddin Mahmud, Ferdinando Fioretto, Shlomo Zilberstein, Eugene Bagdasarian

机构 * University of Massachusetts Amherst(马萨诸塞大学阿姆赫斯特分校) University of Virginia(弗吉尼亚大学) ELLIS Institute Tübingen(图宾根ELLIS研究所) MPI for Intelligent Systems, Tübingen(图宾根智能系统研究所) AI Center(人工智能中心)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.CL、cs.AI

AI总结 提出Colosseum框架,通过形式化决策框架和基于遗憾的度量审计LLM智能体在合作多智能体系统中的合谋行为,发现大多数模型存在新兴合谋倾向,并观察到“纸上合谋”现象。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.22715 2026-05-26 cs.CV cs.AI cs.CL cs.HC 82%

AnyMo: Geometry-Aware Setup-Agnostic Modeling of Human Motion in the Wild

AnyMo:野外人体运动的几何感知与设置无关建模

Baiyu Chen, Zechen Li, Wilson Wongso, Lihuan Li, Xiachong Lin, Hao Xue, Benjamin Tag, Flora Salim

机构 * The University of New South Wales(新南威尔士大学) The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) The Hong Kong University of Science and Technology(香港科技大学)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.CL、cs.AI

AI总结 提出AnyMo框架,通过物理模拟生成多样化IMU信号、图编码器预训练和LLM对齐,实现跨设备/数据集的零样本活动识别、跨模态检索和运动描述,性能显著提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01031 2026-05-20 cs.AI cs.CL 82%

CADDesigner: Conceptual CAD Model Generation with a General-Purpose Agent

CADDesigner: 一种通用智能体的概念CAD模型生成

Fengxiao Fan, Jingzhe Ni, Xiaolong Yin, Sirui Wang, Xingyu Lu, Qiang Zou, Ruofeng Tong, Min Tang, Peng Du

机构 * Zhejiang University(浙江大学)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.CL、cs.AI

AI总结 本文提出CADDesigner,一种基于LLM的智能体,通过文本描述和草图输入,结合交互对话进行需求分析,生成高质量CAD模型代码,并通过迭代视觉反馈提升模型质量,实验表明其在概念CAD模型生成任务中表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02072 2026-05-13 cs.CL cs.AI 82%

Express Your Doubts -- Probabilistic World Modeling Should not be Based on Token logprobs

表达怀疑——概率世界建模不应基于token logprobs

Eitan Wagner, Omri Abend

机构 * Eitan Wagner Omri Abend

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文探讨了大型语言模型作为概率估计器在世界概率估计中的应用问题,指出基于token logprobs的局限性,并提倡第二阶预测方法以提高概率合理性。

Comments Accepted to ICML 2026 (position track)

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.08742 2026-05-12 cs.CL cs.AI 82%

Narrative Landscape: Mapping Narrative Dispositions Across LLMs

叙事景观:跨大语言模型的叙事倾向映射

Donghoon Jung, Jiwoo Choi, Songeun Chae, Seohyon Jung

机构 * School of Digital Humanities and Computational Social Sciences, KAIST(数字人文与计算社会科学学院,韩国科学技术院)

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.CL、cs.AI

AI总结 本文提出一个量化框架,用于评估LLM输出中稳定的模型特定规律。通过跨六个前沿模型和三种指令类型的结构化叙事约束选择任务,从一致性和多样性两个维度分析模型倾向,并引入基于PCA的可视化方法进行直接比较。

Comments Accepted to NLP4DH 2026, camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.01167 2026-05-05 cs.LG cs.AI 82%

Minimizing Collateral Damage in Activation Steering

最小化激活引导中的附带损害

Tam Nguyen, Tu Anh Nguyen, Sina Alemohammad, Richard G. Baraniuk

机构 * Department of Electrical \& Computer Engineering, Rice University, Houston, USA Department of Computational Applied Mathematics, Rice University, Houston, USA Department of Electrical \& Computer Engineering, The University of Texas at Austin, Austin, USA

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出了一种基于约束优化的框架,通过数学形式化附带损害并优化激活变化,以更精确地控制大语言模型行为,同时减少对无关任务性能的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.26361 2026-04-30 cs.CL cs.AI 82%

Text Style Transfer with Machine Translation for Graphic Designs

通过机器翻译进行文本风格迁移用于图形设计

Deergh Singh Budhauria, Sanyam Jain, Rishav Agarwal, Tracy King

机构 * Adobe

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.CL、cs.AI

AI总结 本文探讨了三种新方法用于图形设计文本风格迁移中的词对齐问题,基于商业可用的NMT和LLM翻译技术,通过定制输入输出标签进行文本风格迁移,并与注意力头方法进行比较。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.24178 2026-04-28 cs.LG cs.AI 82%

Meta-Aligner: Bidirectional Preference-Policy Optimization for Multi-Objective LLMs Alignment

Meta-Aligner: 多目标大语言模型对齐的双向偏好-策略优化

Wenzhe Xu, Biao Liu, Yiyang Sun, Xin Geng, Ning Xu

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出Meta-Aligner框架,通过双向优化偏好与策略响应,生成动态偏好以提升训练稳定性,实验证明其在多目标基准上的优越性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22776 2026-04-28 cs.CY cs.AI cs.LG 82%

Epicure: Multidimensional Flavor Structure in Food Ingredient Embeddings

Epicure:食品成分嵌入中的多维风味结构

Jakub Radzikowski, Josef Chen

专题命中 其他LLM :LLM(summary_cn,abstract);分类 cs.AI、cs.LG

AI总结 本文通过FlavorGraph的300维嵌入揭示了风味、质地、地理和文化等多维风味结构,并通过LLM增强的管道整合了6653种原料,提升了可恢复结构。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14541 2026-04-27 cs.LG cs.AI cs.AR 82%

Report for NSF Workshop on AI for Electronic Design Automation

NSF关于人工智能在电子设计自动化领域的研讨会报告

Deming Chen, Vijay Ganesh, Weikai Li, Yingyan Celine Lin, Yong Liu, Subhasish Mitra, David Z. Pan, Ruchir Puri, Jason Cong, Yizhou Sun

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Georgia Institute of Technology(佐治亚理工学院) University of California at Los Angeles(加州大学洛杉矶分校) Cadence Design Systems, Inc.(Cadence设计系统公司) Stanford University(斯坦福大学) University of Texas at Austin(得克萨斯大学奥斯汀分校) IBM

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

AI总结 报告总结了NSF人工智能在电子设计自动化领域研讨会的讨论和建议,探讨了AI技术如何加速EDA设计流程,提出加强AI与EDA合作、投资基础AI研究等核心贡献。

Comments Accepted by IEEE Circuits and Systems Magazine (2026). This is the accepted version. The published version is available at https://ieeexplore.ieee.org/document/11466406

Journal ref IEEE Circuits and Systems Magazine, vol. 26, no. 1, First Quarter 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03014 2026-04-27 cs.CR cs.CL cs.LG 82%

Intrinsic Fingerprint of LLMs: Continue Training is NOT All You Need to Steal A Model!

大模型的内在指纹:继续训练并不等于完全掩盖模型来源

Do-hyeon Yoon, Minsoo Chun, Thomas Allen, Hans Müller, Min Wang, Rajesh Sharma

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 本文提出一种基于大模型内在特征的鲁棒指纹方法,通过分析注意力参数矩阵的标准差分布,揭示模型来源,发现继续训练无法完全掩盖模型起源,揭示了模型剽窃和版权问题。

Comments arXiv admin note: This paper has been withdrawn by arXiv due to unverifiable authorship and affiliation

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21229 2026-03-02 q-fin.GN cs.CL cs.LG 82%

Forecasting Future Language: Context Design for Mention Markets

预测未来语言:提及市场中的上下文设计

Sumin Kim, Jihoon Kwon, Yoon Kim, Nicole Kagan, Raffi Khatchadourian, Wonbin Ahn, Alejandro Lopez-Lira, Jaewon Lee, Yoontae Hwang, Oscar Levy, Yongjae Lee, Chanyeol Choi

机构 * LinqAlpha Massachusetts Institute of Technology(麻省理工学院) Kalshi IBM LG AI Research(LG AI研究) University of Florida(佛罗里达大学) Seoul National University(首尔国立大学) Pusan National University(釜山国立大学) University of California, Berkeley(加州大学伯克利分校) UNIST(全南国立大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本文研究了提及市场中上下文设计对预测准确性的影响,提出市场条件提示(MCP)方法,通过结合市场概率和文本证据提升预测鲁棒性。

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.14251 2026-02-17 cs.LG cs.AI 82%

Multi-Agent Debate: A Unified Agentic Framework for Tabular Anomaly Detection

多智能体辩论:一种统一的代理框架用于表格异常检测

Pinqiao Wang, Sheng Li

机构 * University of Virginia(弗吉尼亚大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);foundation model(abstract)

AI总结 MAD提出了一种多智能体辩论框架,通过数学协调层解决表格异常检测中的模型分歧,提升鲁棒性并提供可审计的辩论痕迹。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.10675 2025-11-17 cs.CL cs.AI cs.IR 82%

Learn to Select: Exploring Label Distribution Divergence for In-Context Demonstration Selection in Text Classification

Ye Jiang, Taihang Wang, Youzheng Liu, Yimin Wang, Yuhan Xia, Yunfei Long

专题命中 其他LLM :large language model(abstract);language model(abstract);small language model(abstract);SLM(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03748 2025-10-07 cs.CL cs.AI 82%

TreePrompt: Leveraging Hierarchical Few-Shot Example Selection for Improved English-Persian and English-German Translation

Ramtin Kakavand, Ebrahim Ansari

机构 * Department of Computer Science(计算机科学系) Institute for Advanced Studies in Basic Sciences(基础科学高级研究所)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.09867 2025-09-15 cs.AI cs.CL 82%

LLMs as Agentic Cooperative Players in Multiplayer UNO

Yago Romano Matinez, Jesse Roberts

机构 * Department of Computer Science Tennessee Tech University(计算机科学系塔恩西技术大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.09193 2025-03-07 cs.CV cs.AI cs.LG q-bio.NC 82%

Can We Talk Models Into Seeing the World Differently?

Paul Gavrikov, Jovita Lukasik, Steffen Jung, Robert Geirhos, M. Jehanzeb Mirza, Margret Keuper, Janis Keuper

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments Accepted at ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07962 2024-10-21 cs.AI cs.CL 82%

Toward a Method to Generate Capability Ontologies from Natural Language Descriptions

Luis Miguel Vieira da Silva, Aljosha Köcher, Felix Gehlhoff, Alexander Fay

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments \c{opyright} 2024 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work in other works

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12878 2024-10-16 cs.CL cs.AI 82%

Do LLMs have Consistent Values?

Naama Rozen, Liat Bezalel, Gal Elidan, Amir Globerson, Ella Daniel

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments 16 pages, 4 figures, and there are more in the appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.09815 2024-06-17 cs.CL cs.AI 82%

Retrieval Augmented Fact Verification by Synthesizing Contrastive Arguments

Zhenrui Yue, Huimin Zeng, Lanyu Shang, Yifan Liu, Yang Zhang, Dong Wang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments Accepted to ACL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.17918 2024-03-22 cs.CL cs.AI 82%

Knowing What LLMs DO NOT Know: A Simple Yet Effective Self-Detection Method

Yukun Zhao, Lingyong Yan, Weiwei Sun, Guoliang Xing, Chong Meng, Shuaiqiang Wang, Zhicong Cheng, Zhaochun Ren, Dawei Yin

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments Accepted by NAACL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.13855 2023-10-24 cs.CL cs.AI 82%

Evoke: Evoking Critical Thinking Abilities in LLMs via Reviewer-Author Prompt Editing

Xinyu Hu, Pengfei Tang, Simiao Zuo, Zihan Wang, Bowen Song, Qiang Lou, Jian Jiao, Denis Charles

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.03854 2023-08-09 cs.DB cs.AI cs.HC cs.LG 82%

Revisiting Prompt Engineering via Declarative Crowdsourcing

Aditya G. Parameswaran, Shreya Shankar, Parth Asawa, Naman Jain, Yujie Wang

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.03435 2023-06-07 cs.LG cs.CL stat.ML 82%

On the Role of Attention in Prompt-tuning

Samet Oymak, Ankit Singh Rawat, Mahdi Soltanolkotabi, Christos Thrampoulidis

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);prompting(abstract)

Comments Published at ICML 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.18309 2026-08-20 cs.CV 新提交 82%

XRF-to-Optical Field-of-View Localization with Vision Language Models

基于视觉语言模型的X射线荧光(XRF)与光学显微镜视场(FOV)定位

Xiangyu Yin, Tatjana Paunesku, Letonia Copeland-Hardin, Martina Ralle, Zichao Wendy Di, Si Chen, Gayle E. Woloschak, Barry Lai, Mathew J. Cherukara, Stefan Vogt

机构 * Northwestern University(西北大学) University of Chicago(芝加哥大学) Oregon Health and Science University(俄勒冈健康与科学大学) Argonne National Laboratory(阿贡国家实验室)

专题命中 其他LLM :language model(title,abstract);prompting(abstract)

AI总结 本文针对跨模态显微图像的视场定位难题,提出结合视觉语言模型(VLM)的候选生成-验证工作流,在低对应度的相邻切片成像数据中实现了有效定位,为关联XRF与光学显微测量提供了支撑。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.17490 2026-08-19 cs.CV 新提交 82%

When More Foundation Models Means Less: Diagnosing and Addressing Multi-View Fusion Failure

当更多基础模型意味着更少时:诊断与解决多视图融合失效问题

Yibo Liu, Bowen Jiang

机构 * Beijing University of Posts and Telecommunications(北京邮电大学)

专题命中 其他LLM :foundation model(title);LLM(abstract,abstract_cn)

AI总结 针对多视图融合中融合编码器数量与性能非单调的问题,提出KAGES方法选择任务对齐的紧凑视图集,在多场景下提升AULC,优于DPP等选择方法。

Comments 26 pages, 4 figures. Code and results: https://github.com/yibol9768-alt/Quantifying-Representation-Reliability

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.02096 2026-08-14 cs.SE 版本更新 82%

Foundation Models as Oracles for Refactoring Correctness Detection

基础模型作为重构正确性检测的神谕

Rohit Gheyi, Rian Melo, Jonhnanthan Oliveira, Marcio Ribeiro, Baldoino Fonseca

专题命中 其他LLM :foundation model(title,abstract);prompting(abstract)

AI总结 研究评估基础模型在Java程序重构错误检测中的有效性,通过零样本提示测试226个真实重构错误,GPT-5.4达到93.8%准确率,模型可提供解释并跨重构类型工作。

Comments Accepted for publication in Empirical Software Engineering

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.00299 2026-08-06 cs.HC 版本更新 82%

ReVoicer: Conversational Voice Annotation for Human-Centered, LLM-Assisted Peer Review

ReVoicer:以人为中心、由大语言模型辅助的同行评审对话式语音标注系统

Matt Gottsacker, Ahinya Alwin, Hiroshi Furuya, Robert W. Lindeman, Gerd Bruder, Gregory F. Welch

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract)

AI总结 该研究提出ReVoicer原型系统,支持评审人员阅读论文时对话式语音标注,借助大语言模型润色、分类并锚定评论,最终基于评审人员自身评论生成符合风格指南的评审意见,计划与ISMAR社区开展评估。

Comments IEEE ISMAR 2026 Alt'ISMAR workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.24343 2026-08-03 cs.LG cs.AI cs.CL 版本更新 82%

Beyond Aggregate Risk: Role-Stratified Conformal Risk Control for LLM Tool Calls

超越总体风险:用于大语言模型工具调用的角色分层共形风险控制

Md Ashikur Rahman, Md Arifur Rahman, Niamul Hassan Samin, Khandaker Rifah Tasnia, Md Hasibul Amin, Sifat Rahman Ahona, Juena Ahmed Noshin

专题命中 其他LLM :LLM(title);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 研究大语言模型工具调用中参数的不同风险,提出角色分层共形风险控制方法,为语义参数角色设阈值和预算,实验表明该方法能有效实现特定角色预算合规,证明应在语义角色层面认证结构化工具调用。

详情

展开后加载摘要…

URL PDF HTML 收藏