arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-03-09 至 2026-03-09 共收录 11 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 11 篇

2404.00806 2026-03-09 econ.GN cs.AI cs.GT q-fin.EC 89%

Algorithmic Collusion by Large Language Models

大语言模型的算法合谋

Sara Fish, Yannai A. Gonczarowski, Ran I. Shorrer

机构 * a pawfessor of economics and of computer science(经济与计算机科学教授) OpenAI’s Researcher Access Program(OpenAI研究员访问计划) Google’s Gemini Academic Program(Google的Gemini学术计划) Cloud Research Credits Program(云研究信用计划) Anthropic NSF Graduate Research Fellowship(NSF研究生研究 fellowship) Kempner Institute Graduate Fellowship(Kempner研究所研究生 fellowship) National Science Foundation (NSF-BSF grant No. 2343922)(国家科学基金会(NSF-BSF grant No. 2343922)) Harvard FAS Dean’s Competitive Fund for Promising Scholarship(哈佛大学哈佛大学教务处有前途的学术研究竞争基金) Harvard FAS Inequality in America Initiative(哈佛大学哈佛大学美国不平等倡议) United States–Israel Binational Science Foundation (BSF grant 2022417)(美国-以色列双边科学基金会(BSF grant 2022417))

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 研究发现大语言模型在寡头市场中因指令变化导致超竞争性定价,揭示了AI定价代理监管的挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06330 2026-03-09 cs.HC cs.AI 85%

Structured Exploration vs. Generative Flexibility: A Field Study Comparing Bandit and LLM Architectures for Personalised Health Behaviour Interventions

结构探索与生成灵活性:一项比较带宽和LLM架构在个性化健康行为干预中的田野研究

Dominik P. Hofer, Haochen Song, Rania Islambouli, Laura Hawkins, Ananya Bhattacharjee, Meredith Franklin, Joseph Jay Williams, Jan D. Smeddinck

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文通过田野研究比较带宽和LLM架构在个性化健康行为干预中的效果,发现LLM方法更有效,但带宽优化未带来额外帮助性,揭示了结构探索与生成灵活性的权衡。

Comments Currently under review at a conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05996 2026-03-09 cs.CL 79%

Track-SQL: Enhancing Generative Language Models with Dual-Extractive Modules for Schema and Context Tracking in Multi-turn Text-to-SQL

Track-SQL: 通过双提取模块增强生成语言模型以在多轮文本到SQL中进行模式和上下文跟踪

Bingfeng Chen, Shaobin Shi, Yongqi Luo, Boyan Xu, Ruichu Cai, Zhifeng Hao

机构 * School of Computer Science, Guangdong University of Technology(广东技术大学计算机科学学院) Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东人工智能与数字经济实验室) Peng Cheng Laboratory(鹏城实验室) College of Science, Shantou University(汕头大学理学院)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

AI总结 Track-SQL通过双提取模块提升生成语言模型在多轮文本到SQL任务中的模式和上下文跟踪能力,实现性能显著提升。

Comments Accepted at the 2025 Conference of the Nations of the Americas Chapter of the Association for Computational Linguistics (NAACL 2025), Long Paper, 19 pages

Journal ref Proceedings of the 2025 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long Papers), pp. 10690-10708. Association for Computational Linguistics, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01653 2026-03-09 cs.CV cs.AI 79%

MAP: Mitigating Hallucinations in Large Vision-Language Models with Map-Level Attention Processing

MAP: 通过地图级注意力处理缓解大视觉-语言模型中的幻觉

Chenxi Li, Yichen Guo, Benfang Qian, Jinhao You, Kai Tang, Yaosong Du, Zonghao Zhang, Xiande Huang

机构 * DAIL Tech(DAIL科技)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

AI总结 MAP通过地图级注意力处理提升大视觉-语言模型的事实一致性与性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05883 2026-03-09 cs.CL 77%

VerChol -- Grammar-First Tokenization for Agglutinative Languages

VerChol -- 以语法优先的词法化方法用于黏着语言

Prabhu Raja

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 VerChol 提出了一种以语法优先的词法化方法,用于处理黏着语言中复杂的形态结构,以提高词法化效率和准确性。

Comments 13 pages. A Morphological Alternative to Statistical Subword Tokenization

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03738 2026-03-09 cs.CL 77%

Activation-Space Personality Steering: Hybrid Layer Selection for Stable Trait Control in LLMs

激活空间人格引导:用于LLMs中稳定特质控制的混合层选择

Pranav Bhandari, Nicolas Fay, Sanjeevan Selvaganapathy, Amitava Datta, Usman Naseem, Mehwish Nasim

机构 * Network Analysis and Social Influence Modelling (NASIM) Lab(网络分析与社会影响建模实验室) School of Physics Maths and Computing, The University of Western Australia(西澳大学物理数学与计算学院) School of Psychological Science, The University of Western Australia(西澳大学心理学科学学院) School of Computing, Macquarie University(麦考瑞大学计算机学院)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出通过混合层选择方法,在LLMs中实现稳定的人格控制,利用Big Five人格特质构建低秩子空间,以提升模型输出的可控性与实用性。

Comments Accepted to EACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.13127 2026-03-09 cs.CV cs.CR 71%

SPARK: Jailbreaking T2V Models by Synergistically Prompting Auditory and Recontextualized Knowledge

SPARK:通过协同提示音频与重新上下文化知识实现T2V模型的劫持

Zonghao Ying, Moyang Chen, Nizhang Li, Zhiqiang Wang, Wenxin Zhang, Quanchen Zou, Zonglei Jing, Aishan Liu, Xianglong Liu

机构 * State Key Laboratory of Complex \& Critical Software Environment, Beihang University College of Science, Mathematics Technology, Wenzhou-Kean University 360 AI Security Lab Faculty of Innovation Engineering, Macau University of Science Hong Kong University of Science University of Chinese Academy of Sciences

专题命中 其他LLM :prompting(title)

AI总结 SPARK通过模块化提示设计,利用音频与视觉关联模式,实现对T2V模型的高效劫持,提升攻击成功率23%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07967 2026-03-09 cs.HC 71%

Pre/Absence: Prompting Cultural Awareness and Understanding for Lost Architectural Heritage in Virtual Reality

预/缺席:在虚拟现实中提示文化意识与理解以应对失落的建筑遗产

Yaning Li, Ke Zhao, Shucheng Zheng, Xingyu Chen, Chenyi Chen, Wenxi Dai, Weile Jiang, Qi Dong, Yiqing Zhao, Meng Li, Lin-Ping Yuan

专题命中 其他LLM :prompting(title)

AI总结 本研究通过虚拟现实设计 Pre/Absence,探索如何在虚拟环境中提升用户对失落建筑遗产的文化意识与理解,通过对比实验发现 VR 能更有效地增强文化认知与情感参与。

Comments for further revisions

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.01285 2026-03-09 cs.CL cs.CY 70%

Do Prevalent Bias Metrics Capture Allocational Harms from LLMs?

流行的偏见度量标准是否能捕捉大语言模型造成的分配伤害?

Hannah Cyberey, Yangfeng Ji, David Evans

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文研究了现有偏见度量标准在评估大语言模型分配伤害时的可靠性,发现基于平均性能差距和分布距离的度量标准无法准确捕捉群体差异。

Comments Accepted to Workshop on Insights from Negative Results in NLP (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06516 2026-03-09 physics.ao-ph 50%

Evaluating the Predictability of Selected Weather Extremes with Aurora, an AI Weather Forecast Model

利用Aurora模型评估选定天气极端事件的可预测性

Qin Huang, Moyan Liu, Yeongbin Kwon, Upmanu Lall

专题命中 其他LLM :foundation model(abstract)

AI总结 Aurora模型在短期极端天气预报中表现优异,但在亚季节性预报中因动力结构丧失而受限。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17355 2026-03-09 cs.CV 50%

UAM: A Unified Attention-Mamba Backbone of Multimodal Framework for Tumor Cell Classification

UAM: 多模态框架中用于肿瘤细胞分类的统一注意力-马amba骨干

Taixi Chen, Jingyun Chen, Nancy Guo

机构 * State University of New York at Binghamton(纽约州立大学布林莫尔分校)

专题命中 其他LLM :foundation model(abstract)

AI总结 UAM提出了一种统一的注意力-马amba架构,用于多模态框架中的肿瘤细胞分类和图像分割,通过灵活结合两种模块提升性能。

详情

展开后加载摘要…

URL PDF HTML 收藏