arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-03-19 至 2026-03-19 共收录 255 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 19 篇

2512.02341 2026-03-19 cs.CV 78%

TALO: Pushing 3D Vision Foundation Models Towards Globally Consistent Online Reconstruction

TALO:推动3D视觉基础模型向全球一致的在线重建迈进

Fengyi Zhang, Tianjun Zhang, Kasra Khosoussi, Zheng Zhang, Zi Huang, Yadan Luo

机构 * UQMM Lab, The University of Queensland(昆士兰大学UQMM实验室) Shanghai Jiao Tong University(上海交通大学) Harbin Institute of Technology(哈尔滨工业大学)

专题命中 其他LLM :foundation model(title,abstract)

AI总结 本文提出基于薄板样条的高自由度长期对齐框架,通过全局传播控制点纠正空间变化不一致,并采用点无关子图注册设计提升鲁棒性,实验表明其在多数据集和相机配置下均能获得更一致的几何和更低的轨迹误差。

Comments CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17887 2026-03-19 cs.HC cs.AI 77%

AI-Assisted Goal Setting Improves Goal Progress Through Social Accountability

人工智能辅助的目标设定通过社会问责制提高目标进展

Michel Schimpf, Julian Voigt, Thomas Bohné

机构 * Technical University of Munich(慕尼黑技术大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究探讨了人工智能辅助目标设定对目标进展的影响,发现其通过增强社会问责感提升短期目标进展,但未显著优于结构化自我反思。

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09127 2026-03-19 cs.CL 77%

Multilingual LLMs Struggle to Link Orthography and Semantics in Bilingual Word Processing

多语言大语言模型在双语词处理中难以建立拼写与语义的联系

Eshaan Tanwar, Gayatri Oke, Tanmoy Chakraborty

机构 * Department of Electrical Engineering, Indian Institute of Technology Delhi(印度理工学院德里电气工程系) Yardi School of Artificial Intelligence, Indian Institute of Technology Delhi(印度理工学院德里人工智能学院)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究探讨多语言大语言模型在处理双语词时,如何通过拼写和语义特征进行区分,发现模型在处理双语同形词时存在显著困难,倾向于依赖拼写相似性而非语义理解。

Comments Code available at: https://github.com/EshaanT/Bilingual_processing_LLMs

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17174 2026-03-19 cs.CR cs.AI cs.SE 77%

Detecting Data Poisoning in Code Generation LLMs via Black-Box, Vulnerability-Oriented Scanning

通过黑盒、面向漏洞的扫描检测代码生成LLM中的数据中毒

Shenao Yan, Shimaa Ahmed, Shan Jin, Sunpreet S. Arora, Yiwei Cai, Yizhen Wang, Yuan Hong

机构 * University of Connecticut(康涅狄格大学) Visa Research(Visa研究)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出CodeScan框架,通过分析多生成代码的结构相似性,结合抽象语法树规范化,检测代码生成模型中的安全漏洞,实现高准确率的数据中毒检测。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10184 2026-03-19 cs.CL 77%

Incongruent Positivity: When Miscalibrated Positivity Undermines Online Supportive Conversations

不一致的积极性:当不准确的积极性削弱在线支持性对话

Leen Almajed, Abeer ALdayel

机构 * Computer Science Department, King Saud University, College of Computer and Information Sciences(计算机科学系,沙特王后大学,计算机与信息科学学院)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究探讨了在高压力情境中,LLM生成的不一致积极性如何导致消极回应,提出通过微调模型和开发多标签分类器来提升支持性对话的质量。

Comments To appear in ICWSM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17902 2026-03-19 cs.CR cs.AI 70%

Differential Privacy in Generative AI Agents: Analysis and Optimal Tradeoffs

生成AI代理中的差分隐私:分析与最优权衡

Ya-Ting Yang, Quanyan Zhu

机构 * Department of Electrical and Computer Engineering, New York University(纽约大学电气与计算机工程系)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文从企业数据角度分析生成AI代理的隐私泄露问题,提出基于差分隐私的概率框架,推导隐私界限并优化温度参数选择。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16900 2026-03-19 physics.soc-ph cs.AI cs.HC 70%

Social physics in the age of artificial intelligence

人工智能时代的社会物理学

The Anh Han, Joel Z. Leibo, Tom Lenaerts, Iyad Rahwan, Fernando Santos, Matjaž Perc, Valerio Capraro

机构 * School of Computing, Engineering and Digital Technologies, Teesside University(计算、工程与数字技术学院,泰赛大学) Google DeepMind(谷歌DeepMind) Machine Learning Group, Université Libre de Bruxelles(机器学习组,布鲁塞尔自由大学) AI Lab, Vrije Universiteit Brussel(人工智能实验室,布鲁塞尔自由大学) Center for Human-Compatible AI, UC Berkeley(人类兼容人工智能中心,伯克利大学) Max Planck Institute for Human Development, Center for Humans & Machines, Berlin(人类发展马克斯普朗克研究所,人类与机器中心,柏林) Informatics Institute, University of Amsterdam(信息学院,阿姆斯特丹大学) Faculty of Natural Sciences and Mathematics, University of Maribor(自然科学与数学学院,马里博大学) Community Healthcare Center Dr. Adolf Drolc Maribor(阿多夫·德罗尔博士社区医疗中心,马里博) Department of Physics, Kyung Hee University(物理系,庆熙大学) University College, Korea University(大学学院,韩国大学) University of Milan Biccoca(米兰Biccoca大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文探讨人工智能时代社会物理学的新研究方向,聚焦人类与机器的共演化,提出六个关键研究领域,包括社会行为的演化动力学、机器文化、语言与行为的共演化等。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17192 2026-03-19 cs.CY 67%

Narrative Frames: A New Approach to Analysing Metaphors in AI Ethics and Policy Discourse

叙事框架:一种分析人工智能伦理与政策 discourse 中隐喻的新方法

Daniel Stone

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出Narrative Frames框架,通过归纳编码和交叉参考,系统分析AI政策 discourse 中隐喻,解决现有方法定义不一致的问题,为研究者和政策制定者提供共同词汇。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17055 2026-03-19 cs.CV 67%

PaAgent: Portrait-Aware Image Restoration Agent via Subjective-Objective Reinforcement Learning

PaAgent:通过主观-客观强化学习实现的面向人物图像修复代理

Yijian Wang, Qingsen Yan, Jiantao Zhou, Duwei Dai, Wei Dong

机构 * School of Computer Science, Northwestern Polytechnical University(西北工业大学计算机学院) Shenzhen Research Institute of Northwestern Polytechnical University(西北工业大学深圳研究院) State Key Laboratory of Internet of Things for Smart City, University of Macau(澳门大学智慧城市物联网国家重点实验室) National-Local Joint Engineering Research Center of Biodiagnosis and Biotherapy, the Second Affiliated Hospital of Xi’an Jiaotong University(西安交通大学生物诊断与生物治疗国家地方联合工程研究中心) College of Information and Control Engineering, Xi’an University of Architecture and Technology(西安建筑科技大学信息与控制工程学院)

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 PaAgent通过结合自进化的人物银行和检索增强生成技术,提升图像修复任务中对复杂场景的感知能力,通过主观-客观强化学习策略优化修复工具选择,实验验证其在多种修复基准上的优越性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17970 2026-03-19 cs.LG cs.NA math.NA math.OC 57%

Beyond Muon: MUD (MomentUm Decorrelation) for Faster Transformer Training

超越缪子:MUD(动量去相关)用于更快的Transformer训练

Ben S. Southworth, Stephen Thomas

机构 * Theoretical Division, Los Alamos National Laboratory(洛斯阿拉莫斯国家实验室理论部) Lehigh University(莱斯大学)

专题命中 其他LLM :language model(abstract);分类 cs.LG

AI总结 MUD通过三角化白化替代缪子的极分解更新,提升Transformer训练效率,具有更低的优化器开销和更快的困惑度表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17912 2026-03-19 cs.CL stat.ML 57%

Pretrained Multilingual Transformers Reveal Quantitative Distance Between Human Languages

预训练多语言Transformer揭示人类语言之间的定量距离

Yue Zhao, Jiatao Gu, Paloma Jeretič, Weijie Su

专题命中 其他LLM :language model(abstract);分类 cs.CL

AI总结 本文提出利用预训练多语言模型中的注意力机制计算语言距离,通过Attention Transport Distance(ATD)方法揭示语言间的定量关系,并提升低资源机器翻译性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17901 2026-03-19 cs.SI cs.CY 50%

Grievance Politics vs. Policy Debates: A Cross-Platform Analysis of Conservative Discourse on Truth Social and Reddit

诉求政治与政策辩论:对保守派在Truth Social和Reddit上的跨平台分析

Yining Wang, Alhasan Abdellatif, Artemis Deligianni, Hannah Hok, Yusuf Mucahit Cetinkaya, Tugrulcan Elmas

专题命中 其他LLM :LLM(abstract)

AI总结 本文通过主题建模分析Truth Social与Reddit保守派社区,发现Truth Social以诉求和叙事内容为主,而Reddit更侧重政策辩论,且Reddit的毒性更高。

Comments Accepted at ICWSM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16736 2026-03-19 cs.CV 50%

World Reconstruction From Inconsistent Views

从不一致视角重建世界

Lukas Höllein, Matthias Nießner

机构 * Technical University of Munich, Germany(慕尼黑技术大学)

专题命中 其他LLM :foundation model(abstract)

AI总结 本文提出一种非刚性对齐方法,通过全局一致坐标框架生成清晰点云,提升视频扩散模型在3D世界重建中的性能。

Comments project website: https://lukashoel.github.io/video_to_world video: https://www.youtube.com/watch?v=qXnUwhVmBzA code: https://github.com/lukasHoel/video_to_world

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.17520 2026-03-19 cs.CV 50%

PCA-Seg: Revisiting Cost Aggregation for Open-Vocabulary Semantic and Part Segmentation

PCA-Seg:重新审视开放词汇语义和部分分割中的成本聚合

Jianjian Yin, Tao Chen, Yi Chen, Gensheng Pei, Xiangbo Shu, Yazhou Yao, Fumin Shen

机构 * Nanjing University of Science and Technology(南京理工大学) Nanjing Normal University(南京师范大学) Department of Electrical and Computer Engineering, Sungkyunkwan University(成均馆大学电子与计算机工程系) University of Electronic Science and Technology of China(电子科技大学) State Key Laboratory of Intelligent Manufacturing of Advanced Construction Machinery(先进施工机械智能制造国家重点实验室)

专题命中 其他LLM :language model(abstract)

AI总结 本文提出PCA-Seg方法,通过并行成本聚合缓解类级语义与空间上下文之间的知识干扰,提升开放词汇语义和部分分割性能。

Comments Accepted by CVPR2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21304 2026-03-19 cs.DB cs.MA 50%

ORCA: ORchestrating Causal Agent

ORCA:协调因果代理

Joanie Hayoun Chung, Sumin Lee, Sungbin Lim

专题命中 其他LLM :LLM(abstract)

AI总结 ORCA通过维护共享状态和引入人工检查点,实现关系数据库的协调因果分析,提升用户交互效率和因果结论可靠性。

Comments 35 pages, CHI EA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏