arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-07-30 至 2026-07-30 共收录 21 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 21 篇

2607.26642 2026-07-30 cs.AI 新提交 92%

AlphaSchema: Exploring the Space of Trading Semantics for LLM-Based Alpha Mining

AlphaSchema:探索基于大语言模型(LLM)的阿尔法挖掘的交易语义空间

Jingyang Yi, Jian Yang, Yifei Jin, Yuqi Li, Jian Li

机构 * X-Tech Monash University(莫纳什大学) PandaAI(熊猫AI) IIIS, Tsinghua University(清华大学智能产业研究院(IIIS))

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 AlphaSchema构建并探索结构化交易语义空间,通过解耦探索与实施、结合代理模型平衡探索与利用,在中国股市挖掘出强预测与投资组合表现的因子池,且对LLM选择具鲁棒性。

Comments Initial Version

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.27134 2026-07-30 cs.AI cs.CL cs.GT 新提交 91%

Linguistic Monoculture in LLM-Assisted Language Use

LLM辅助语言使用中的语言单一文化

Suhas Thejaswi, Juhi Kulshreshta, Lutz Oettershagen

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 该研究针对LLM辅助语言使用中的语言单一文化问题,构建数学框架分析作者与LLM的共同演化机制,发现个性化可保留语言多样性,且个体理性作者的过度一致性会产生负外部性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.22657 2026-07-30 cs.NE 版本更新 90%

Automatic programming via large language models with population self-evolution for dynamic fuzzy job shop scheduling problem

结合种群自进化大语言模型的动态模糊作业车间调度问题自动编程

Jin Huang, Qihao Liu, Xinyu Li, Liang Gao, Yue Teng

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn)

AI总结 该研究针对动态模糊作业车间调度问题,提出种群自进化框架结合大语言模型自动设计启发式调度规则,性能优于多种现有方法。

Comments 13 pages, 10 figures. Accepted for publication in IEEE Transactions on Fuzzy Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.18496 2026-07-30 cs.CR cs.AI cs.HC 版本更新 90%

Towards an Automated Test of LLM Security Knowledge

迈向大语言模型安全知识的自动化测试

Shufan Chai, Liangliang Sun, Jessica Staddon

机构 * Northeastern University(东北大学)

专题命中 其他LLM :LLM(title,summary_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究旨在实现大语言模型安全知识自动化测试,引入利用消费者保护机构权威信息识别LLM响应不稳定性的部分自动化方法,通过对身份盗窃和冒名顶替诈骗主题及Gemini和GPT家族中5个LLMs进行测试,可区分模型安全知识充足与否。

Comments v3: fixed typos in abstract metadata; no changes to the paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.23971 2026-07-30 cs.HC cs.AI 版本更新 88%

Ask don't tell: Reducing sycophancy in large language models

请勿告知:减少大语言模型的趋炎附势

Magda Dubois, Cozmin Ududec, Christopher Summerfield, Lennart Luettgau

机构 * UK AI Security Institute(英国人工智能安全研究所)

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);分类 cs.AI

AI总结 本文通过实验研究发现,非问题比问题引发更高的趋炎附势倾向,且趋炎附势随用户传达的可信度增加而增强,采用将非问题转为问题可有效降低趋炎附势。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26899 2026-07-30 cs.HC cs.AI cs.CY 新提交 86%

Human diversity fuels collective creativity that large language models cannot simulate or sustain

人类多样性催生了大型语言模型无法模拟或维持的集体创造力

Mengchen Dong, Hiromu Yakura

专题命中 其他LLM :large language model(title);language model(title);prompting(abstract);分类 cs.AI

AI总结 该研究通过实验发现,L2写作者的集体多样性优于L1写作者,AI构思会压缩集体多样性而AI润色可保留,模拟显示AI无法达到人类群体的创意多样性,人机协作设计影响人类多样性的存续。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26591 2026-07-30 cs.SE 新提交 86%

MultiFixer: A Coordinator-Proposer Based Multi-Agent Framework For Fixing Multi-Hunk Bugs

MultiFixer:一种用于修复多块漏洞的基于协调者-提议者的多智能体框架

Haichuan Hu, Chunrong Fang, Ye Shang, Jiawei Liu, Weifeng Sun, Guoqing Xie, Chenxing Zhong, Quanjun Zhang

专题命中 其他LLM :LLM(summary_cn,abstract);large language model(abstract);language model(abstract)

AI总结 该研究针对现有LLM-based APR方法难以修复多块漏洞的问题,提出MultiFixer多智能体框架,经多基准测试,其在Defects4J等数据集上达到最优修复效果,能有效修复多块漏洞。

Comments Accepted to 41st IEEE/ACM International Conference on Automated Software Engineering (ASE 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06114 2026-07-30 cs.CL cs.AI 版本更新 82%

Making Implicit Premises Explicit in Logical Understanding of Enthymemes

在演绎推理中显式化隐含前提

Xuyao Feng, Anthony Hunter

专题命中 其他LLM :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 本文提出了一种整合大型语言模型和神经符号推理器的管道,用于将隐含前提显式化并解码逻辑蕴含。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.27154 2026-07-30 cs.CV cs.AI 新提交 79%

Anatomy Contextualized Adaption of CT Foundation Models

CT基础模型的解剖学上下文适配

Roshan Kenia, Stephanie L McNamara, William Lotter

机构 * Harvard Medical School(哈佛医学院) Dana-Farber Cancer Institute(达纳-法伯癌症研究所) Massachusetts General Hospital(麻省总医院) Brigham and Women’s Hospital(布里格姆妇女医院)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI

AI总结 本研究提出轻量框架ACA,适配冻结CT基础模型实现解剖学级视觉-语言对齐,在Merlin、CT-RATE数据集上零样本发现分类性能优于基线,训练耗时短且可保留增强全局解剖学上下文。

Comments Accepted to the European Conference on Computer Vision (ECCV) 2026 MedFM-Bench Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26762 2026-07-30 cs.CL 新提交 79%

Relation Geometry in Semantic Space of Language Models

语言模型语义空间中的关系几何

Zhihan Cao, Hiroaki Yamada, Simone Teufel, Tatsuya Hiraoka, Kentaro Inui, Hitomi Yanaka, Takenobu Tokunaga

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

AI总结 本文研究语言模型语义空间中关系几何的表征情况,通过三类语言模型在6种语义关系上的实验,发现非对称关系的表征更清晰,且不同模型依赖的信息来源存在差异。

Comments Manuscript under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26288 2026-07-30 econ.EM 新提交 78%

The Innate Economic Preferences of Language Models

语言模型的固有经济偏好

Joy Buchanan, Joshua Foster

专题命中 其他LLM :language model(title,abstract)

AI总结 该研究揭示语言模型生成规则与离散选择随机效用模型同构,通过12种模型的投资组合任务发现其普遍存在异质风险厌恶,且微调可实现目标风险态度的工程化构建。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.16755 2026-07-30 cs.SE 版本更新 78%

An Empirical Study of Foundation Models for Variability-Induced Compilation Errors in Configurable C Code

具有变异性意识的编译错误检测与修复:利用基础模型在可配置系统中

Rohit Gheyi, Lucas Albuquerque, Márcio Ribeiro, Eduardo Almeida, Danyllo Albuquerque, Mirko Perkusich

专题命中 其他LLM :foundation model(title,abstract)

AI总结 本文研究利用基础模型检测和修复可配置系统中的编译错误,展示其在精度和召回率上的优势。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.27167 2026-07-30 cs.SE cs.CL 新提交 70%

SpecFirst: Behavioral Specification Elicitation as a First-Class Step in Agent-Based Program Synthesis from Scratch

SpecFirst:将行为规范提取作为从零开始的基于智能体的程序合成中的一等步骤

Yihao Chen, Shi Chang, Feng Lin, Khaled Chawa, Boyuan Chen, Shaowei Wang, Ahmed E. Hassan

专题命中 其他LLM :LLM(abstract,abstract_cn);分类 cs.CL

AI总结 SpecFirst是将行为规范提取设为独立前置阶段的两阶段框架,在ProgramBench上显著提升了从零开始的程序合成的测试通过率与二进制探索覆盖率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26735 2026-07-30 cs.CV cs.AI cs.MM 新提交 70%

Dual Inversion for Text-to-Image Diffusion Models: From Both Prompt and Noise Perspectives

文本到图像扩散模型的双重逆推:从提示与噪声双视角

Xiaolong Liu, Junjian Li, Yuan Xiao, Jiaqi Deng, Dayong Ye, Tianqing Zhu, Huan Huo

机构 * University of Technology Sydney(悉尼科技大学) Geely(吉利) City University of Hong Kong(香港城市大学) City University of Macau(澳门城市大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.AI

AI总结 该研究针对现有文本到图像扩散模型提示逆推方法的局限,提出联合恢复语义提示与潜在噪声的Dualin方法,实现了高质量逆推提示与最优图像保真度,为可控图像编辑奠定基础。

Comments Accepted by ACM MM 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26726 2026-07-30 cs.CL 新提交 70%

AtmosERC: Modeling Dialogue-Level Affective Atmosphere for Emotion Recognition in Conversation

AtmosERC:面向对话情感识别的对话级情感氛围建模

Weijie Feng, Tongwei Zhang, Binbin Liu, Zhiyong Cheng

专题命中 其他LLM :LLM(abstract,abstract_cn);分类 cs.CL

AI总结 本文提出AtmosERC框架,通过建模对话级情感氛围提升对话情感识别性能,其可作为插件线索增强基于大语言模型的ERC,且在局部情感偏差下预测更稳定。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26340 2026-07-30 cs.NI cs.DC cs.LG 新提交 70%

Incast-Free MoE Rate-Based Scheduling

无内爆的MoE基于速率的调度

Evyatar Cohen, Jose Yallouz, Alexander Shpiner, Mark Silberstein, Sylvia Ratnasamy, Isaac Keslassy

专题命中 其他LLM :large language model(abstract);language model(abstract);分类 cs.LG

AI总结 针对MoE架构轮询调度引发的指数级内爆瓶颈,提出主动公平调度框架,可消除内爆、维持高链路利用率并降低集体完成时间。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.18699 2026-07-30 cs.DS cs.SC 版本更新 67%

A more accurate rational non-commutative algorithm for multiplying 4x4 matrices using 48 multiplications

一种更精确的有理非交换算法用于使用48次乘法的4x4矩阵乘法

Jean-Guillaume Dumas, Clément Pernet, Alexandre Sedoglavic

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文提出一种更精确的4x4矩阵乘法算法,使用48次非交换乘法,改进了误差界指数并提升了实际精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.08951 2026-07-30 cs.SE 版本更新 67%

GenAI Is No Silver Bullet for Qualitative Research in Software Engineering

生成式AI并非软件工程定性研究的万能钥匙

Neil A. Ernst, Christoph Treude

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本文探讨生成式AI在软件工程定性研究中的应用,分析其优缺点及对研究质量的影响,旨在为研究人员提供关于该技术的全面理解。

Comments accepted to ACM TOSEM. Replication: https://doi.org/10.5281/zenodo.21631436

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26391 2026-07-30 cs.LG q-bio.BM 新提交 57%

Q-Steer: Action-Value Guidance for Molecular Policy Optimization

Q-Steer:分子策略优化的动作值引导

Xinyu Wang, Jinbo Bi, Minghu Song

机构 * University of Connecticut(康涅狄格大学) Institute of Health and Medicine, Hefei Comprehensive National Science Center(合肥综合性国家科学中心健康与医学技术研究所)

专题命中 其他LLM :language model(abstract);分类 cs.LG

AI总结 Q-Steer 是一种用于分子语言模型的生成时动作值引导原语,通过 PAVS-Q 评分器在固定在线 oracle 预算下提升分子优化性能,在所有测试组合中均取得正增益。

Comments 20 pages, 2 figures, and 25 tables. Includes supplementary experimental results

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.26375 2026-07-30 cs.CL cs.HC 新提交 57%

(Im)Paired Programming: Coding Agents Improve Productivity but Harm Understanding

(非)配对编程:编码智能体提升生产力但损害理解能力

Nishant Balepur, Connor Baumler, Valerie Chen, Eunsol Choi, Rachel Rudinger, Jordan Lee Boyd-Graber

专题命中 其他LLM :prompting(abstract);分类 cs.CL

AI总结 该研究通过54名学生的对照实验,发现编码智能体虽提升生产力但损害用户代码理解,低投入交互加剧该问题,用户仍偏好智能体,同时为开发者提出了优化方向。

Comments In-progress Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.27139 2026-07-30 cs.CV 新提交 50%

SeasonStereo: Robust Dense Stereo Matching for Multi-Date Satellite Imagery via Generative AI

SeasonStereo:基于生成式AI的多日期卫星图像鲁棒密集立体匹配

Álvaro Díaz-Laureano, Roger Marí, Elías Masquil, Pablo Arias, Gabriele Facciolo

机构 * Eurecat(欧罗卡特技术中心) IIE, Facultad de Ingeniería, Universidad de la República(乌拉圭共和国大学工程学院IIE) Universitat Pompeu Fabra(庞培法布拉大学) Université Paris-Saclay(巴黎萨克雷大学) CNRS(法国国家科学研究中心) ENS Paris-Saclay(巴黎萨克雷高等师范学校) Institut Universitaire de France(法国大学研究院)

专题命中 其他LLM :foundation model(abstract)

AI总结 SeasonStereo框架利用带可控季节外观变化的合成图像对训练并结合基础模型零样本几何先验,实现低成本、高精度的多日期卫星图像密集立体匹配,支持大规模三维重建。

详情

展开后加载摘要…

URL PDF HTML 收藏