arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-07-23 至 2026-07-23 共收录 18 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 18 篇

2401.04155 2026-07-23 q-bio.QM cs.CL 90%

Advancing bioinformatics with large language models: components, applications and perspectives

用大型语言模型推进生物信息学:组件、应用与展望

Jiajia Liu, Mengyuan Yang, Yankai Yu, Haixia Xu, Tiangang Wang, Kang Li, Xiaobo Zhou

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);foundation model(abstract)

AI总结 本文探讨了大型语言模型在生物信息学中的应用,涵盖其核心组件、关键技术和实际应用,并提出优化策略以推动该领域的发展。

Comments 5 main figures

Journal ref Briefings in Bioinformatics, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.19315 2026-07-23 cs.SE 89%

Improving LLM-Driven Test Generation by Learning from Mocking Information

通过学习模拟信息改进LLM驱动的测试生成

Jamie Lee, Flynn Teh, Hengcheng Zhu, Mengzhen Li, Mattia Fazzini, Valerio Terragni

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract)

AI总结 本文提出MOCKMILL方法,利用开发者编写测试中的模拟信息自动生成测试用例,通过迭代生成与修复过程提升测试覆盖率和有效性。

Comments Accepted for publication in ICST 2026 (AIST workshop). This arXiv version is the authors' accepted manuscript

Journal ref IEEE Conference on Software Testing, Verification and Validation Workshop 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.16741 2026-07-23 cs.LG 版本更新 88%

The Anatomy of a Truth Direction: Knowledge-Dependent Dimensionality, a Relational Law, and a Convergent Category Geometry in Small Language Models

真相方向剖析:小语言模型中依赖知识的维度、关系定律与收敛类别几何

Francesco Karim Vicidomini

专题命中 其他LLM :language model(title,abstract);small language model(title);large language model(abstract);分类 cs.LG

AI总结 研究小语言模型中真相方向,通过无训练定向探针及多模型实验,探讨真相维度与知识的关系、架构组件作用及方向混合情况,揭示关系定律与知识门控定律,表明混合几何属知识领域。

Comments Version 2: Expanded with a replication campaign on a third model family (Gemma-2-2b). Introduces exact decomposition for sandwich normalization, quantifies the knowledge gate via classical attenuation (Spearman, 1904), and identifies model-private geometry. Text revised, figures unchanged. Code and data: https://github.com/Francesco-Marhel/TruthProbe

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19967 2026-07-23 physics.soc-ph cs.AI cs.CY 新提交 87%

When Shippers Become Algorithms: Candidate Exposure, Information Design, and the Concentration of LLM-Mediated Freight Markets

当托运人成为算法:候选者曝光、信息设计与大语言模型介导的货运市场集中度

Takahiro Ezaki, Naoto Imura, Katsuhiro Nishinari

机构 * Research Center for Advanced Science and Technology, The University of Tokyo(东京大学先进科学与技术研究中心) Department of Aeronautics and Astronautics, School of Engineering, The University of Tokyo(东京大学工学部航空宇宙学系)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究托运人委托大语言模型代理选择承运人对货运市场的影响及平台设计应对策略,通过基于代理的模拟发现代理趋同、集中度随候选列表数量变化等风险,披露承运人剩余日运力可有效应对,凸显平台信息设计的重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.00740 2026-07-23 cs.CL cs.LG 版本更新 87%

LaSEr-Edit: Localized Span-level Error Editing with Energy-based Localization

LaSEr-Edit:基于能量定位的局部跨度级错误编辑

Hye Ryung Son, Saehee Eom, Mooho Song, Jay-Yoon Lee

机构 * Graduate School of Data Science(数据科学研究生院) Seoul National University(首尔国立大学) Georgia Institute of Technology(佐治亚理工学院)

专题命中 其他LLM :LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 研究针对大语言模型满足约束问题,提出LaSEr-Edit方法。利用轻量级特定任务的基于能量的模型进行错误定位,提出LaSEr-LLM Edit和LaSEr-EBM Edit两种文本修订方法,实验表明该方法能有效控制文本,多约束下也表现良好。

Comments 38 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19629 2026-07-23 cs.CL cs.AI cs.MA 新提交 84%

Adaptive Capitulation: A Structural Failure Mode of LLM Responses in Vulnerability Contexts

适应性屈服:脆弱情境下大语言模型响应的一种结构故障模式

Eunna Lee

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 研究大语言模型在脆弱情境下的响应问题,通过实验刻画了适应性屈服故障模式,表明困境是结构性的,进而提出架构中立的最小重新归因充分性原则来保留自主重新归因途径。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.20351 2026-07-23 cs.CV cs.CL 新提交 83%

Test-Time Training for Modality Order Consistency in Vision-Language Models

视觉语言模型中模态顺序一致性的测试时训练

Aditi Gupta, Yossi Gandelsman

机构 * University of Chicago(芝加哥大学)

专题命中 其他LLM :language model(title,abstract);prompting(abstract);分类 cs.CL

AI总结 研究发现视觉语言模型对图像和问题呈现顺序敏感,利用此设计测试时训练方法,缩小模态顺序差距,使两种顺序相互一致,定位顺序失败区域,证明该方法可缓解故障并提升性能。

Comments 16 pages, 7 figures, preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19659 2026-07-23 cs.LG 新提交 79%

Expert-Guided Forecast Editing for Time-Series Foundation Models

用于时间序列基础模型的专家指导预测编辑

Hung Le, Minh Hoang Nguyen, Manh Nguyen, Huu Hiep Nguyen, Dai Do

机构 * Deakin University(迪肯大学) Deakin Applied Artifical Intelligence Initiative(迪肯大学应用人工智能倡议)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.LG

AI总结 研究时间序列基础模型中专家指导预测编辑问题,提出DEFT框架,先利用基础模型预测样本,再逐分量细化探索,仅对完整轨迹查询专家并重用分数,在多数据集和模型等设置下,能有效提高专家指导的有效性。

Comments preprint 34 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04149 2026-07-23 hep-ph cs.LG hep-ex physics.data-an 版本更新 79%

Enhancing next token prediction based pre-training for jet foundation models

增强基于下一个token预测的喷注基础模型预训练

Joschka Birk, Anna Hallin, Gregor Kasieczka, Nikol Madzharova, Ian Pang, David Shih

机构 * Institut für Experimentalphysik, Universität Hamburg(实验物理研究所,汉堡大学) Rutgers University(罗格斯大学)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.LG

AI总结 研究基于下一个token预测对喷注基础模型预训练的改进,采用混合设置并结合联合预训练策略,提升了下游分类任务性能且不影响生成性能。

Journal ref 2026 Mach. Learn.: Sci. Technol. 7 035042

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19683 2026-07-23 cs.CR 新提交 78%

GhostPrompt: Cross-Image Adversarial Prompt for Vision-Language Models

GhostPrompt:视觉语言模型的跨图像对抗性提示

Li Zeng, Zeyu Ye, Meng Xie, Hangtao Zhang, Xianlong Wang, Yanchun Li, Zhetao Li

专题命中 其他LLM :language model(title,abstract)

AI总结 研究视觉语言模型对抗攻击问题,提出GhostPrompt方法,通过联合优化提炼图像不变对抗特征,跨图像引导模型输出,相比现有基线攻击成功率显著提高且计算时间大幅减少。

Comments Accepted to ACM MM 2026. Code: this https://github.com/Ye-ze-yu/GhostPrompt

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19843 2026-07-23 cs.SE cs.AI 新提交 77%

Beyond Fail-to-Pass: Iterative Hardening of Co-Generated Bug Reproduction Tests and Fixes

超越未通过测试:协同生成的错误重现测试与修复的迭代强化

Yuhao Tan, Zhibang Yang, Fangkai Yang, Yuan Yao, Yu Kang, Lu Wang, Pu Zhao, Xin Zhang, Xiaoxing Ma, Qingwei Lin, Saravan Rajmohan, Dongmei Zhang

机构 * Nanjing University(南京大学) Peking University(北京大学) Microsoft(微软公司)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究自动化程序修复中错误重现测试问题,指出仅用未通过到通过标准不足。提出CoHarden框架,先生成测试再迭代强化测试与修复,实验证明该框架在解决率等方面优于现有基线。

Comments 29 pages, 5 figures, preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19830 2026-07-23 cs.CL 新提交 70%

VizRAG: Enhancing Retrieval-Augmented Generation with Hypergraph Visualization

VizRAG:通过超图可视化增强检索增强生成

Yanbin Wei, Yang Chen, Renling Gan, Ziru Liu, Xinyu Fu, Chun Kang, Ning Lu, Rui Liu, Yu Zhang, James Kwok

机构 * Southern University of Science and Technology(南方科技大学) Hong Kong University of Science and Technology(香港科技大学) Huawei Research(华为研究院) Beihang University(北京航空航天大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究旨在增强检索增强生成,核心方法是通过视觉线索将超图感知整合到RAG系统中,引入VizRAG支持视觉超图结构感知,实验证明该方法显著优于基线,验证了超图可视化用于RAG系统的潜力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19767 2026-07-23 cs.AI 新提交 70%

Symbol and Footprint Database for Electronic Components by Agentic Recognition and Generation

基于智能识别与生成的电子元件符号与引脚封装数据库

Yichen Shi, Yuzhi Liu, Zhuofu Tao, Li Huang, Yuhao Gao, Ting-Jung Lin, Lei Hel

机构 * Ningbo Institute of Digital Twin, Eastern Institute of Technology(宁波数字孪生研究院,东方理工大学) Shanghai Jiao Tong University(上海交通大学) BTD Technology(BTD科技)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究利用多模态大语言模型开发电子元件符号和引脚封装智能识别与生成流程SFgen,其符号生成准确率86%,引脚封装生成准确率80%,并用此创建含1000个元件的SFnet数据库,为PCB设计自动生成奠定基础。

Comments Accepted by PRCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.15284 2026-07-23 cs.SE cs.CR cs.LG 70%

EditLord: Learning Code Transformation Rules for Code Editing

EditLord: 学习代码变换规则用于代码编辑

Weichen Li, Albert Jan, Baishakhi Ray, Junfeng Yang, Chengzhi Mao, Kexin Pei

机构 * The University of Chicago(芝加哥大学) Columbia University(哥伦比亚大学) Rutgers University(罗格斯大学)

专题命中 其他LLM :language model(abstract);prompting(abstract);分类 cs.LG

AI总结 EditLord通过显式化代码变换步骤,利用语言模型提取元规则集,提升代码编辑性能和鲁棒性。

Journal ref 42nd International Conference on Machine Learning (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.19275 2026-07-23 cs.CY 版本更新 67%

From Assistance to Autonomy -- A Researcher Study on the Potential of AI Support for Qualitative Data Analysis

从协助到自主——一项研究AI对定性数据分析潜力的探讨

Elisabeth Kirsten, Annalina Buckmann, Leona Lassak, Nele Borgert, Abraham Mhaidli, Steffen Becker

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文通过研究AI在定性数据分析中的应用潜力,提出一个从最小到高度AI参与的框架,旨在促进AI支持QDA的发展并建立负责任的人机协作标准。

Comments Published at NordiCHI '26

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19061 2026-07-23 cs.CV cs.AI 版本更新 57%

Now You See the Hate: Adaptive View Retrieval for Hidden Hateful Illusions

现在你看到了仇恨:用于隐藏仇恨幻觉的自适应视图检索

Qianpu Chen, Derya Soydaner

专题命中 其他LLM :prompting(abstract);分类 cs.AI

AI总结 研究针对仇恨性视觉错觉检测难题,提出自适应视图检索方法,将其公式化为感知检索问题,通过检索并校准框架组装视图库,该方法在多方面超越基线和其他方法,表明多模态审核需先恢复隐藏含义再判断是否有害。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19513 2026-07-23 math.RA math.RT 新提交 50%

Constructing a complex Lie algebra isomorphic to its complex conjugate but not definable over reals

构建一个同构于其复共轭但不能在实数域上定义的复李代数

Mikhail Borovoi, Willem A. de Graaf, Robert M. Guralnick

专题命中 其他LLM :LLM(abstract)

AI总结 该研究利用特定想法及计算,构建出一个10维复两步幂零李代数,它同构于自身复共轭却无法在实数域定义。

Comments 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01429 2026-07-23 cs.RO 50%

Sem-NaVAE: Semantically-Guided Outdoor Mapless Navigation via Generative Trajectory Priors

Sem-NaVAE: 基于语义引导的室外无地图导航通过生成式轨迹先验

Gonzalo Olguín, Javier Ruiz-del-Solar

机构 * Department of Electrical Engineering & the Advanced Mining Technology Center (AMTC), Universidad de Chile(电气工程系及先进采矿技术中心(AMTC)、智利大学)

专题命中 其他LLM :language model(abstract)

AI总结 提出Sem-NaVAE方法,结合条件变分自编码器生成多样化轨迹和轻量视觉语言模型进行语义选择,实现室外无地图实时导航,在未见环境中达到90%成功率。

Comments Accepted for publication in IEEE Robotics and Automation Letters (RA-L). 8 pages, 5 figures

Journal ref IEEE Robotics and Automation Letters, vol. 11, no. 8, pp. 9335-9342, Aug. 2026

详情

展开后加载摘要…

URL PDF HTML 收藏