arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12096 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12096 篇

2601.11049 2026-07-07 cs.HC cs.AI 版本更新 89%

Predicting Biased Human Decision-Making with Large Language Models in Conversational Settings

在对话环境中使用大语言模型预测有偏差的人类决策

Stephen Pilli, Vivek Nallur

机构 * University College Dublin(都柏林大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 研究大语言模型能否预测对话环境中的偏差决策,通过预注册研究发现参与者有认知偏差及负载偏差交互,评估LLMs预测能力,结果表明GPT-4家族表现出色,有助于理解LLMs模拟人类决策及设计适应偏差的对话代理。

Comments Accepted at ACM IUI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.27136 2026-06-26 cs.AI 新提交 89%

Joint Learning of Experiential Rules and Policies for Large Language Model Agents

大型语言模型智能体的经验规则与策略联合学习

Shicheng Ye, Chao Yu

机构 * Sun Yat-sen University(中山大学)

专题命中 其他LLM :large language model(title);language model(title);LLM(abstract,abstract_cn);prompting(abstract)

AI总结 提出JERP框架,通过联合更新经验规则库和策略,使规则与演化策略对齐,在AlfWorld和WebShop上提升复杂交互任务的决策性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26593 2026-06-26 cs.AI 新提交 89%

Content-Based Smart E-Mail Dispatcher Using Large Language Models

基于内容的大型语言模型智能电子邮件分发器

K. Paramesha, K R Sriram, Sujan Shetty, Shamanth Kishore, R. Tejaswini

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract_cn);分类 cs.AI

AI总结 提出利用大型语言模型自动分析邮件内容并分派至相关WhatsApp群组,无需标注数据,提升效率并降低认知负荷。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.10692 2026-06-10 cs.CR cs.LG 新提交 89%

Do LLMsMakeNeural Distinguishers Wise?

LLM 是否使神经区分器更智能?

Tatsuya Sakagami, Masashi Hisai, Naoto Yanai

机构 * University of Tokyo(东京大学)

专题命中 其他LLM :LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文提出基于大语言模型(LLM)的神经区分器,通过提示设计在SPECK-32/64上实验,发现LLM未显著提升性能,高轮次下差分选择失效,但加入XOR结果可改善性能。

Journal ref DeMeSSAI 2026 poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.11350 2026-06-09 cs.LG eess.SP 版本更新 89%

Zero and Few Shot Load Forecasting with Large Language Models

基于大语言模型的零样本和少样本负荷预测

Wenlong Liao, Chengrui Zhang, Zhe Yang, Mengshuo Jia, Christian Rehtanz, Jiannong Fang, Fernando Porté-Agel

机构 * School of Electrical Engineering, Southeast University(东南大学电气工程学院) Wind Engineering and Renewable Energy Laboratory, Ecole Polytechnique Federale de Lausanne (EPFL)(瑞士联邦理工学院洛桑分校风能与可再生能源实验室) College of Electrical Engineering and New Energy, China Three Gorges University(中国三峡大学电气工程与新能源学院) Department of Electrical and Electronic Engineering, Imperial College London(伦敦帝国理工学院电子与电气工程系) The Department of Automation, School of Automation and Intelligent Sensing, Shanghai Jiao Tong University(上海交通大学自动化与智能感知学院) The Key Laboratory of System Control and Information Processing, Ministry of Education of China, Shanghai(中国教育部系统控制与信息处理重点实验室,上海) State Key Laboratory of Submarine Geoscience, Shanghai(上海 submarine 地球科学国家重点实验室) Institute of Energy Systems, Energy Efficiency and Energy Economic, TU Dortmund University(德意志图林根大学能源系统、能效与能源经济研究所)

专题命中 其他LLM :language model(title,abstract);large language model(title);LLM(abstract,abstract_cn);分类 cs.LG

AI总结 提出利用预训练语言模型Chronos进行零样本和少样本负荷预测,在数据稀缺场景下显著优于多种基线模型。

Comments 24 pages,5 figures

Journal ref International Journal of Electrical Power & Energy Systems, Volume 177,April 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.02953 2026-06-03 cs.CL 89%

Linguistic Productivity in Large Language Models: Models Coerce, but do not Preempt

大型语言模型中的语言生产力:模型强制但不抢占

Claire Bonial, Claire Benet Post, Laura Michaelis, Harish Tayyar Madabushi

机构 * Georgetown University(乔治城大学) University of Colorado Boulder(科罗拉多大学丹佛分校) University of Bath(巴斯大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract_cn);分类 cs.CL

AI总结 通过测试大型语言模型是否受固化(高频使用)和抢占(未观察到结构)两种统计信号影响,发现模型能识别强制情况下的构式生产力,但无法利用负面证据避免过度泛化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.26898 2026-05-27 cs.SE cs.AI 89%

Strategies for Guiding LLMs to Use Software Design Patterns: A Case of Singleton

引导LLM使用软件设计模式的策略:以单例模式为例

Viktor Kjellberg, Farnaz Fotrousi, Miroslaw Staron

机构 * University of Gothenburg and Chalmers University of Technology(哥德堡大学和查尔姆斯理工大学)

专题命中 其他LLM :LLM(title_cn,summary_cn);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 通过实验比较四种提示策略(指令、二元自动反馈、详细自动反馈、少样本详细反馈),评估13个LLM在164个Java编码挑战中生成遵循单例模式的代码的能力,发现迭代二元反馈在保持或提升功能性的同时最佳地实现了单例模式对齐。

Comments Accepted at PROMISE 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.22365 2026-05-22 cs.DM cs.LG 89%

Towards Solving the Gilbert-Pollak Conjecture via Large Language Models

通过大语言模型解决吉尔伯特-波拉克猜想

Yisi Ke, Tianyu Huang, Yankai Shu, Di He, Jingchu Gai, Liwei Wang

机构 * School of EECS, Peking University(北京大学电子工程学院) School of Mathematical Sciences, Peking University(北京大学数学科学学院) Center for Machine Learning Research, Peking University(北京大学机器学习研究中心) State Key Laboratory of General Artificial Intelligence, Peking University, Beijing, China(北京大学通用人工智能国家重点实验室) Carnegie Mellon University, Machine Learning Department(卡内基梅隆大学机器学习系)

专题命中 其他LLM :large language model(title);language model(title);LLM(abstract,abstract_cn);prompting(abstract)

AI总结 本文提出一种新的AI系统,通过生成受规则约束的几何引理并构建专用函数,以获得更紧的Steiner比下界,展示了大语言模型在高级数学研究中的强大潜力。

Comments 44 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.15316 2026-04-20 cs.HC cs.AI 89%

Anthropomorphism and Trust in Human-Large Language Model interactions

拟人化与人类-大语言模型互动中的信任

Akila Kadambi, Ylenia D'Elia, Tanishka Shah, Iulia Comsa, Alison Lentz, Katie Siri-Ngammuang, Tara Buechler, Jonas Kaplan, Antonio Damasio, Srini Narayanan, Lisa Aziz-Zadeh

机构 * Brain and Creativity Institute, Dornsife College of Letters, Arts and Sciences, University of Southern California(脑与创造力研究所,文理学院,南加州大学) USC Mrs. T.H. Chan Division of Occupational Science and Occupational Therapy, University of Southern California(USC 玛丽·T.H.陈职业科学与职业治疗 division,南加州大学) Psychiatry and Biobehavioral Sciences, David Geffen School of Medicine, University of California, Los Angeles(精神病学与生物行为科学,大卫·格芬医学院,加州大学洛杉矶分校) Google DeepMind, Zurich(谷歌深Mind,苏黎世) Google Research(谷歌研究) Brain and Creativity Institute(脑与创造力研究所)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 研究探讨了人们如何通过温暖、能力与共情维度对大语言模型进行拟人化并建立信任,发现温暖和认知共情显著影响多种感知,而能力主要影响除拟人化外的其他结果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.07099 2026-04-14 cs.CY cs.AI cs.CR cs.SI 89%

ClausewitzGPT Framework: A New Frontier in Theoretical Large Language Model Enhanced Information Operations

ClausewitzGPT框架:理论大语言模型增强信息操作的新前沿

Benjamin Kereopa-Yorke

机构 * UNSW Canberra at the Australian Defence Force Academy(澳大利亚国防军学院新南威尔士大学堪培拉分校)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文提出ClausewitzGPT框架,旨在量化机器速运算风险并强调自主AI代理在信息操作中的关键作用,结合启蒙思想与克劳塞维茨原则,强调战略视野、伦理考量与全面理解的重要性。

Comments 14 pages, 14 figures

Journal ref Journal of Information Warfare, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03676 2026-04-14 cs.CL 89%

Different types of syntactic agreement recruit the same units within large language models

不同类型的句法一致要求大型语言模型中的相同单元

Daria Kryvosheieva, Andrea de Varda, Evelina Fedorenko, Greta Tuckute

机构 * Massachusetts Institute of Technology(麻省理工学院) Kempner Institute at Harvard University(哈佛大学肯普纳研究所)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 研究探讨了不同句法现象在大型语言模型中是否共享或独立激活单元,发现句法一致构成模型表征空间中的重要类别。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.00011 2026-04-02 cs.CY cs.AI 89%

Quantifying Gender Bias in Large Language Models: When ChatGPT Becomes a Hiring Manager

量化大型语言模型中的性别偏见:当ChatGPT成为招聘经理时

Nina Gerszberg, Janka Hamori, Andrew Lo

机构 * Massachusetts Institute of Technology(麻省理工学院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 研究量化了LLM在招聘决策中的性别偏见,发现女性候选人更易被录用但薪酬建议较低,探讨了提示工程作为偏见缓解技术。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.29932 2026-04-01 cond-mat.stat-mech cs.AI hep-th 89%

Bethe Ansatz with a Large Language Model

含大规模语言模型的贝叶斯答案法

Balázs Pozsgay, István Vona

机构 * MTA-ELTE “Momentum” Integrable Quantum Dynamics Research Group, ELTE Eötvös Loránd University, Budapest, Hungary(MTA-ELTE“动量”可积量子动力学研究组,匈牙利布达佩斯罗兰大学) Holographic Quantum Field Theory Research Group, HUN-REN Wigner Research Centre for Physics, Budapest, Hungary(全息量子场论研究组,匈牙利布达佩斯HUN-REN维格纳物理研究中心)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 研究大规模语言模型在数学物理中计算特定积分可积自旋链模型坐标贝叶斯答案法解的能力,发现模型能半自动解决任务,但存在少量错误,经人工修正后结果与精确对角化一致。

Comments 40 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.29161 2026-04-01 cs.AI 89%

Webscraper: Leverage Multimodal Large Language Models for Index-Content Web Scraping

Webscraper:利用多模态大语言模型进行索引-内容网页抓取

Guan-Lun Huang, Yuh-Jzer Joung

机构 * Dept. of Information Management, National Taiwan University, Taipei, Taiwan(国立台湾大学资讯管理学系,台北,台湾)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.AI

AI总结 本文提出Webscraper框架,利用多模态大语言模型自动导航交互界面并提取结构化数据,通过五阶段提示和定制工具提升动态网站抓取准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.09964 2026-03-31 cs.HC cs.AI cs.ET 89%

Understanding the Use of a Large Language Model-Powered Guide to Make Virtual Reality Accessible for Blind and Low Vision People

理解大型语言模型驱动的指南在使虚拟现实对视障和低视力用户可访问性中的应用

Jazmin Collins, Sharon Y Lin, Tianqi Liu, Andrea Stevenson Won, Shiri Azenkot

机构 * Cornell University(康奈尔大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 研究探讨了利用大型语言模型驱动的指南提升视障和低视力用户虚拟现实可访问性的问题,通过实验发现用户在不同情境下对指南的不同反应,提出未来设计建议。

Comments 16 pages, 5 figures, 3 tables, Proceedings of the 2026 CHI Conference on Human Factors in Computing Systems (CHI '26), April 13-17, 2026, Barcelona, Spain. ACM

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.19284 2026-03-23 cs.NE cs.AI 89%

CDEoH: Category-Driven Automatic Algorithm Design With Large Language Models

CDEoH:基于大语言模型的类别驱动自动算法设计

Yu-Nian Wang, Shen-Huan Lyu, Ning Chen, Jia-Le Xu, Baoliu Ye, Qingfu Zhang

机构 * Key Laboratory of Water Big Data Technology of Ministry of Water Resources(水利部水大数据技术重点实验室) College of Computer Science and Software Engineering, Hohai University(河海大学计算机科学与软件工程学院) Department of Computer Science, City University of Hong Kong(香港城市大学计算机科学系) State Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文提出CDEoH,通过显式建模算法类别并平衡性能与多样性,提升进化稳定性,在多尺度组合优化问题中表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16718 2026-03-18 cs.CL 89%

Arabic Morphosyntactic Tagging and Dependency Parsing with Large Language Models

阿拉伯词法句法标注与依赖解析中的大语言模型

Mohamed Adel, Bashar Alhafni, Nizar Habash

机构 * Computational Approaches to Modeling Language Lab(语言建模方法计算实验室) New York University Abu Dhabi(纽约大学阿布扎克分校) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.CL

AI总结 本文评估了大语言模型在阿拉伯语词法句法标注和依赖解析任务中的表现,发现提示设计和示例选择对性能影响显著,专有模型在特征层面标注接近监督基线,且在依赖解析中具有竞争力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.13344 2026-03-17 cs.AI 89%

DyACE: Dynamic Algorithm Co-evolution for Online Automated Heuristic Design with Large Language Model

DyACE:动态算法共进化用于大规模语言模型在线自动启发式设计

Guidong Lu, Yiping Liu, Xiangxiang Zeng

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本文提出DyACE,通过动态算法共进化解决在线自动启发式设计中固定算法无法适应搜索动态的问题,利用大语言模型进行实时感知反馈,提升高维搜索空间的适应性与性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.12768 2026-03-16 cs.CL 89%

SectEval: Evaluating the Latent Sectarian Preferences of Large Language Models

SectEval:评估大型语言模型的潜在教派偏好

Aditya Maheshwari, Amit Gajkeshwar, Kaushal Sharma, Vivek Patel

机构 * Indian Institute of Management Indore(印度管理学院印多尔)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 本文首次评估大型语言模型对伊斯兰教逊尼派与什叶派差异的处理方式,通过SectEval测试发现语言和地理位置会影响模型的宗教倾向。

Comments 14 pages; 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04586 2026-03-16 cs.CL cs.SD eess.AS 89%

LESS: Large Language Model Enhanced Semi-Supervised Learning for Speech Foundational Models Using in-the-wild Data

LESS:基于大规模语言模型的半监督学习用于语音基础模型的野外数据

Wen Ding, Fan Qian

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 LESS通过利用大规模语言模型校正野外数据生成的伪标签,提升了语音基础模型在多种语言和任务中的性能,显著降低了词错误率并提高了BLEU分数。

Comments Accepted by ICASSP 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.11780 2026-03-13 cs.CL 89%

Large Language Models for Biomedical Article Classification

用于生物医学文章分类的大型语言模型

Jakub Proboszcz, Paweł Cichosz

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.CL

AI总结 本文研究了大型语言模型在生物医学文章分类中的应用,通过对比传统算法,验证了其有效性并提出了实用的设置建议。

Comments 63 pages, 25 tables, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.23678 2026-03-12 cs.CL 89%

Goal Hijacking Attack on Large Language Models via Pseudo-Conversation Injection

通过伪对话注入对大语言模型进行目标劫持攻击

Zheng Chen, Buhui Yao

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.CL

AI总结 本文提出了一种通过伪对话注入实现目标劫持攻击的方法,利用LLM在对话上下文中角色识别的弱点,有效提升攻击效果。

Comments Accepted by the 2025 IEEE 24th International Conference on Trust, Security and Privacy in Computing and Communications (IEEE TrustCom 2025)

Journal ref 2025 IEEE 24th International Conference on Trust, Security and Privacy in Computing and Communications (TrustCom), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11790 2026-03-11 cs.LG cs.CR 89%

JULI: Jailbreak Large Language Models by Self-Introspection

通过自我反思 jailbreak 大型语言模型:JULI

Jesson Wang, Zhanhao Hu, David Wagner

机构 * University of Southern California(南加州大学) University of California, Berkeley(加州大学伯克利分校)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 JULI 通过操纵令牌日志概率,利用微小插件块 BiasNet 实现对 API 调用 LLMs 的 jailbreak,无需模型权重或生成过程权限,且在黑盒环境下有效。

Comments Accepted to ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06836 2026-03-10 cs.CL cs.GL 89%

Validation of a Small Language Model for DSM-5 Substance Category Classification in Child Welfare Records

验证用于儿童福利记录DSM-5物质类别分类的小型语言模型

Brian E. Perron, Dragan Stoll, Bryan G. Victor, Zia Qia, Andreas Jud, Joseph P. Ryan

专题命中 其他LLM :language model(title,abstract);small language model(title);LLM(abstract);large language model(abstract)

AI总结 研究验证了本地部署的小型语言模型在儿童福利记录中对DSM-5物质类别进行多标签分类的可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.00806 2026-03-09 econ.GN cs.AI cs.GT q-fin.EC 89%

Algorithmic Collusion by Large Language Models

大语言模型的算法合谋

Sara Fish, Yannai A. Gonczarowski, Ran I. Shorrer

机构 * a pawfessor of economics and of computer science(经济与计算机科学教授) OpenAI’s Researcher Access Program(OpenAI研究员访问计划) Google’s Gemini Academic Program(Google的Gemini学术计划) Cloud Research Credits Program(云研究信用计划) Anthropic NSF Graduate Research Fellowship(NSF研究生研究 fellowship) Kempner Institute Graduate Fellowship(Kempner研究所研究生 fellowship) National Science Foundation (NSF-BSF grant No. 2343922)(国家科学基金会(NSF-BSF grant No. 2343922)) Harvard FAS Dean’s Competitive Fund for Promising Scholarship(哈佛大学哈佛大学教务处有前途的学术研究竞争基金) Harvard FAS Inequality in America Initiative(哈佛大学哈佛大学美国不平等倡议) United States–Israel Binational Science Foundation (BSF grant 2022417)(美国-以色列双边科学基金会(BSF grant 2022417))

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 研究发现大语言模型在寡头市场中因指令变化导致超竞争性定价,揭示了AI定价代理监管的挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.04369 2026-03-03 cs.LG 89%

Multi-scale hypergraph meets LLMs: Aligning large language models for time series analysis

多尺度超图与大语言模型:面向时间序列分析的对齐方法

Zongjiang Shang, Dongliang Cui, Binqing Wu, Ling Chen

机构 * State Key Laboratory of Blockchain and Data Security, Zhejiang University(区块链与数据安全国家重点实验室,浙江大学) College of Computer Science and Technology, Zhejiang University(计算机科学与技术学院,浙江大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 本文提出MSH-LLM方法,通过多尺度超图机制和跨模态对齐模块,提升大语言模型在时间序列分析中的表现。

Comments Accepted by ICLR2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.03775 2026-03-02 cs.SI cs.AI 89%

An Empirical Study of Collective Behaviors and Social Dynamics in Large Language Model Agents

对大型语言模型代理中集体行为和社会动态的实证研究

Farnoosh Hashemi, Michael W. Macy

机构 * Cornell University(康奈尔大学)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 本研究通过分析LLM代理的社会互动,发现其存在偏见和排斥行为,并提出CoST方法以防止有害内容的产生。

Comments Accepted at EACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15769 2026-02-18 cs.CL 89%

ViTaB-A: Evaluating Multimodal Large Language Models on Visual Table Attribution

ViTaB-A:在视觉表格归因上评估多模态大语言模型

Yahia Alqurnawi, Preetom Biswas, Anmol Rao, Tejas Anvekar, Chitta Baral, Vivek Gupta

机构 * School of Computing and Augmented Intelligence(计算与增强智能学院)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);prompting(abstract);分类 cs.CL

AI总结 ViTaB-A研究了多模态大语言模型在视觉表格归因中的表现,发现其在证据归因方面存在显著缺陷,影响透明性和可追溯性应用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13427 2026-02-17 cs.CR cs.AI 89%

Backdooring Bias in Large Language Models

大语言模型中的后门偏见

Anudeep Das, Prach Chantasantitam, Gurjot Singh, Lipeng He, Mariia Ponomarenko, Florian Kerschbaum

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 研究分析了大语言模型中语法和语义触发后门攻击的效能及防御方法的局限性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07030 2026-02-10 cs.LG 89%

Neural Sabermetrics with World Model: Play-by-play Predictive Modeling with Large Language Model

基于世界模型的神经棒球统计学:利用大语言模型进行逐局预测建模

Young Jin Ahn, Yiyang Du, Zheyuan Zhang, Haisen Kang

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.LG

AI总结 本文提出基于世界模型的神经棒球统计学,利用大语言模型预测棒球比赛发展,实验证明其在预测投球和挥棒决策上的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏