arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-07-23 至 2026-07-23 共收录 221 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 14 篇

2607.19739 2026-07-23 cs.IR cs.AI 新提交 90%

Personalized Recommendation Tool Learning via Autonomous Language Agents

通过自主语言智能体进行个性化推荐工具学习

Mingdai Yang, Zhiwei Liu, Weizhi Zhang, Yibo Wang, Hao Peng, Philip Yu

机构 * Microsoft(微软公司)

专题命中 领域大模型 :LLM(summary_cn,abstract);language agent(title);large language model(abstract);language model(abstract)

AI总结 研究针对基于大语言模型的智能体在推荐系统中的局限性,提出PRTA框架,让LLM作中央规划器与多推荐模型交互,负责高级推理和个性化工具选择,传统模型进行全排序评分,设计反射机制,实验表明该框架提升了全排序推荐性能。

Comments 6 pages. Accepted by RecSys'26

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.16022 2026-07-23 cs.CL 版本更新 88%

Enhancing LLMs for Identifying and Prioritizing Important Medical Jargons from Electronic Health Record Notes Utilizing Data Augmentation: A Comparative Study

通过数据增强提升LLM识别和优先级排序电子健康记录笔记中的重要医学术语

Won Seok Jang, Sharmin Sultana, Zonghai Yao, Hieu Tran, Zhichao Yang, Sunjae Kwon, Hong Yu

机构 * Miner School of Computer and Information Sciences, UMass Lowell, MA, USA(UMass Lowell 计算机与信息科学学院) Manning College of Information and Computer Sciences, UMass Amherst, MA, USA(UMass Amherst 信息与计算机科学学院) Center for Healthcare Organization and Implementation Research, VA Bedford Health Care, MA, USA(VA Bedford医疗中心医疗组织与实施研究中心)

专题命中 领域大模型 :LLM(title_cn,summary_cn);prompting(abstract);分类 cs.CL

AI总结 本文通过数据增强提升LLM在低资源环境下识别和优先级排序电子健康记录中的医学术语,评估了闭源和开源模型的性能,发现微调和数据增强能提升模型表现。

Comments 28pages, 5 figures, 6 tables

Journal ref JMIR AI. 17.Jul.2026 in Vol 5 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.20284 2026-07-23 cs.CV 新提交 88%

Multimodal Large Language Models for Remote Sensing Image Understanding: Domain-Specific or General-Purpose?

用于遥感图像理解的多模态大语言模型:领域特定还是通用?

Qiwei Ma, Chunping Qiu, Xinjun Cheng, Xiaoyu Zhang, Puhong Duan, Ke Yang, Xudong Kang, Shutao Li

机构 * School of Artificial Intelligence and Robotics, Hunan University(湖南大学人工智能与机器人学院) Intelligent Game and Decision Lab (IGDL)(智能游戏与决策实验室) Yuelushan Center for Industrial Innovation(岳麓山工业创新中心)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract)

AI总结 本文针对遥感图像理解的多模态大语言模型展开研究,通过系统调查和评估,比较其与通用模型在不同任务上的表现,发现当前模型存在局限,进而给出未来方向,为开发相关模型提供系统参考。

Comments 27 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.09988 2026-07-23 cs.IR cs.AI 版本更新 85%

An LLM-powered Agentic Recommendation System for Connected TV Content Discovery

一种用于智能电视内容发现的由大语言模型驱动的智能推荐系统

Lei Shi, Di Wang, Harry Tran, Helsing Xu, Yuchen Lu, Dhara Ghodasara, Wilson Chaney, Xueting Liao, Jerry Yu, Huayu Ding, Reza Mirghaderi, David Fan, Qi Guo, Chongguang He, Warren Wang, Warren Deng, Mingze Gao, Shike Mei, Shuo Tang, Zhe Zhang, Jianming He, Abhishek Kumar, Haotian Wu, Hamed Firooz, Li Li

机构 * Meta

专题命中 领域大模型 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 针对推荐系统整合上下文信号的挑战,提出用于智能电视内容发现的大语言模型驱动的智能推荐系统,利用其推理能力处理多样信号,采用智能架构协调组件,克服大语言模型用于推荐的实际限制。

Comments 13 pages, 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.20058 2026-07-23 cs.AI cond-mat.mes-hall cond-mat.mtrl-sci cs.CL 新提交 84%

Reading and Steering Representations of Materials-Science Mechanisms in an Open-Weight Language Model

在开放权重语言模型中读取和引导材料科学机制的表示

Markus J. Buehler

机构 * Massachusetts Institute of Technology(麻省理工学院)

专题命中 领域大模型 :language model(title,abstract);large language model(abstract);分类 cs.CL、cs.AI

AI总结 研究大语言模型中材料科学机制信息的表示形式,结合多种方法,包括匹配读数、状态几何等,通过实验验证其三种形式,如概念可读、取向由状态变换承载等,还通过比较提示等发现物理关系在受控状态变化中更易显现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.20274 2026-07-23 cs.CV cs.AI cs.CL cs.LG 新提交 82%

Self-supervision drives representational convergence in medical foundation models more than clinical supervision

自我监督比临床监督更能推动医学基础模型中的表征趋同

Soroosh Tayebi Arasteh, Sebastian Ziegelmayer, Mahshad Lotfinia, Lisa Adams, Sven Nebelung, Jakob Nikolas Kather, Daniel Truhn

机构 * RWTH Aachen University(亚琛工业大学) University Hospital RWTH Aachen(亚琛工业大学附属医院) Technical University of Munich(慕尼黑工业大学) Friedrich-Alexander-Universität Erlangen-Nürnberg(埃尔朗根-纽伦堡弗里德里希-亚历山大大学) Technical University Dresden(德累斯顿工业大学) University Hospital Dresden(德累斯顿大学附属医院) University Hospital Heidelberg(海德堡大学附属医院)

专题命中 领域大模型 :foundation model(title);pretraining(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 研究探讨医学基础模型中表征趋同问题,通过对多个编码器剖析发现自我监督比临床监督更能驱动收敛,虽收敛有限但线性分类器可跨编码器转移,表明医学成像收敛由预训练目标决定,为互操作性设计与验证提供依据。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.20027 2026-07-23 cs.LG 新提交 79%

Zero-Shot Heart Rate Variability Forecasting from Consumer Wearables Using Time Series Foundation Models

使用时间序列基础模型从消费级可穿戴设备进行零样本心率变异性预测

Luukas Peräkylä, Fahad Sohrab, Ville Hautamäki, Merja Heinäniemi, Sui Huang, Pekka Abrahamsson

机构 * Tampere University(坦佩雷大学) University of Eastern Finland(东芬兰大学)

专题命中 领域大模型 :foundation model(title,abstract);分类 cs.LG

AI总结 研究利用时间序列基础模型从消费级可穿戴设备预测心率变异性,针对数据碎片化问题引入变异性保留插补方法,结果显示TSFMs在无需微调时优于传统基线模型,为其在真实数据集上的性能建立基线,凸显特定领域微调对临床部署的潜力。

Comments Accepted to Computing in Cardiology (CinC) 2026. 4 pages, 2 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.07643 2026-07-23 cs.CV 版本更新 78%

Universality Reconsidered: Rethinking the Validation of Foundation Models for General-Purpose 3D Medical Segmentation

揭示通用3D医学分割中的模态差异与泛化幻觉

Yichi Zhang, Le Xue, Feiyang Xiao, Wenbo Zhang, Gang Feng, Chenguang Zheng, Yuan Qi, Yuan Cheng, Zixin Hu

机构 * Fudan University, Shanghai, China.(复旦大学) Shanghai Academy of Artificial Intelligence for Science, Shanghai, China.(上海人工智能科学研究院) Shanghai Universal Medical Imaging Diagnostic Center, Shanghai, China.(上海通用医学影像诊断中心)

专题命中 领域大模型 :foundation model(title,abstract)

AI总结 本文通过UMD数据集揭示3D医学分割中模态差异与泛化幻觉问题,指出现有基础模型在实际应用中存在显著性能差异,需转向多模态训练以提升通用性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19822 2026-07-23 eess.SP 新提交 75%

WARA: A Closed-Loop Multi-Agent Framework for Wireless Optimization Autoresearch

WARA:用于无线优化自动研究的闭环多智能体框架

Yuan Guo, Yilong Chen, Chao Hu, Xianghao Yu, Liang Hong, Jie Xu

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract)

AI总结 本文针对无线领域提出端到端自动研究框架WARA,通过闭环多智能体系统将研究工作流程分三个阶段,经工件介导过程和控制器管理验证工件,设计评分智能体评估有效性,其性能优于一次性大语言模型生成,接近同行评审论文质量。

Comments 2026 IEEE/CIC International Conference on Communications in China (ICCC)

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19406 2026-07-23 cs.LG 新提交 57%

NMR Elucidation as an Agentic Search Problem, Not a Modeling Problem

核磁共振解析作为一个智能体搜索问题,而非建模问题

Irina Espejo Morales, Damon Hinz, Marvin Alberts, Geraud Krawezik, Haewon Jeong, Shirley Ho

专题命中 领域大模型 :LLM(abstract);分类 cs.LG

AI总结 研究针对NMR数据结构解析这一化学等领域的瓶颈,构建由冻结语言模型支持的自主智能体,将其解析过程视为受限搜索而非建模任务,在多个数据集上取得良好结果,为自动化光谱分析的多步骤编排框架提供了方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.28390 2026-07-23 cs.AI 版本更新 57%

You Live More Than Once: Towards Hierarchical Skill Meta-Evolving

你活不止一次:迈向分层技能元进化

Xujun Li, Kehan Zheng, Mingyuan Zhao, Yize Geng, Jinfeng Zhou, Qi Zhu, Fei Mi, Lifeng Shang, Minlie Huang, Hongning Wang

机构 * Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系) Huawei Foundation Model Department(华为基础模型部门)

专题命中 领域大模型 :LLM(abstract_cn);分类 cs.AI

AI总结 本文提出HiSME,一种轻量级分层技能元进化方法,通过从智能体任务执行轨迹中学习元技能,联合优化技能和技能进化策略,以持续提升部署的智能体系统在不同下游场景中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19988 2026-07-23 cs.CY 新提交 50%

What Does the Credential Still Certify? Cognitive Stewardship for AI-Mediated Education

证书仍然证明什么?人工智能介导教育的认知管理

Kai Yao

专题命中 领域大模型 :LLM(abstract)

AI总结 研究探讨生成式AI对教育评估前提的改变,开发认知管理框架,审核30所大学的AI评估指南,发现公共政策分类AI使用较好但解释证书有效性不足,强调大学需制定使认证逻辑可见的政策。

Comments Accepted at the Ninth AAAI/ACM Conference on AI, Ethics, and Society (AIES 2026), Malmö, Sweden, October 12--14, 2026. 13 pages, 3 figures; supplementary material is available as an ancillary file

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22649 2026-07-23 cs.CV 版本更新 50%

Interactive Medical-SAM2 GUI: A Napari-based semi-automatic annotation tool for medical images

交互式医学-SAM2 GUI:基于Napari的半自动标注工具用于医学图像

Woojae Hong, Jong Ha Hwang, Jiyong Chung, Joongyeon Choi, Hyunggun Kim, Yong Hwy Kim

机构 * Department of Biomechatronic Engineering, Sungkyunkwan University(生物机电工程系,全州大学) Department of Neurosurgery, Seoul National University Hospital, Seoul National University College of Medicine(神经外科,首尔国立大学医院,首尔国立大学医学院)

专题命中 领域大模型 :prompting(abstract)

AI总结 交互式医学-SAM2 GUI 提供基于Napari的本地工作流,实现高效的3D医学图像半自动标注与导出功能。

Comments Planning to submit JOSS (Journal of Open Source Software)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 知识编辑与模型理解 10 篇

2607.17890 2026-07-23 cs.AI 版本更新 89%

Stress Testing Concept Erasure with Large Language Model Agents

用大语言模型智能体对概念擦除进行压力测试

Yuyang Xue, Feng Chen, Zhihua Liu, Edward Moroshko, Jingyu Sun, Steven McDonagh, Sotirios A. Tsaftaris

机构 * School of Engineering, University of Edinburgh(爱丁堡大学工程学院) The University of Manchester(曼彻斯特大学) The University of Melbourne(墨尔本大学)

专题命中 知识编辑与模型理解 :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI

AI总结 研究针对概念擦除评估面临的验证挑战,提出STACE框架,利用大语言模型智能体迭代生成并验证压力测试假设,通过一套指标评估性能效率,经实验证明该框架在多方面表现优异且可拓展到其他领域。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19371 2026-07-23 cs.AI cs.CL cs.LG 新提交 86%

Mitigating Scaffolding Collapse in Socratic Tutors via Representation Alignment

通过表示对齐减轻苏格拉底式导师中的支架坍塌

Jing Shao, Qifeng Wu, Hanyu Zhang, Sixia Sun, Jun Zhuang

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);preference optimization(abstract)

AI总结 研究基于大语言模型的苏格拉底式导师的支架坍塌问题,提出支架保留表示对齐方法,先监督微调预热,再结合轨迹加权优化与表示损失,经多学科和策略评估,该方法能提升长程苏格拉底辅导的鲁棒性。

Comments preprint, under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19618 2026-07-23 q-bio.GN cs.AI cs.LG 新提交 84%

Causal dictionary learning reveals and validates transcription-factor binding features in genomic language models

因果字典学习揭示并验证基因组语言模型中的转录因子结合特征

Sarwan Ali

机构 * Columbia University Irving Medical Center(哥伦比亚大学伊文思医疗中心)

专题命中 知识编辑与模型理解 :language model(title,abstract);foundation model(abstract);分类 cs.AI、cs.LG

AI总结 研究针对基因组语言模型内部表示不透明问题,引入结合稀疏字典学习与因果干预的框架,训练自动编码器提取转录因子结合特征,开发去混淆协议并因果验证,为基因组深度学习可解释性提供计算标准。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.22864 2026-07-23 cs.LG 版本更新 79%

Reading Calibrated Uncertainty from Language Model Trajectories

从语言模型轨迹中读取校准的不确定性

Aliai Eusebi, Alexander Herzog, Xiaoyu Liang, Marie Vasek, Enrico Mariconti, Lorenzo Cavallaro

机构 * University College London, London, United Kingdom(伦敦大学学院)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.LG

AI总结 本文通过提取层间MLP更新的尺度不变几何特征,使用稀疏线性探针从语言模型内部激活轨迹中校准不确定性,在选择性弃权任务中优于最大softmax概率,最高提升21 AURC点。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.20202 2026-07-23 eess.SP 新提交 78%

JEPA-CFM: A Joint Embedding Predictive Architecture-based Channel Foundation Model for Robust Fluid Antenna Systems

JEPA-CFM:一种基于联合嵌入预测架构的稳健流体天线系统信道基础模型

Yuan Gao, Yiming Liu, Jun Jiang, Jianbo Du, Shunqing Zhang, Xiaoli Chu, Kai-Kit Wong

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

AI总结 针对流体天线系统获取信道信息及定位的障碍,提出基于联合嵌入预测架构的信道基础模型,通过提取潜在嵌入学习通用表示,结合互补损失项训练,经仿真验证其在信道外推和无线定位上优于传统基线。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19977 2026-07-23 physics.chem-ph physics.atm-clus physics.comp-ph 新提交 78%

Rem3Di: Learning smooth, chiral 3D molecular descriptors from atomistic foundation models

Rem3Di:从原子基础模型学习平滑、手性3D分子描述符

Steffen Wedig, Felix Burton, Rokas Elijošius, Christoph Schran, Lars L. Schaaf

专题命中 知识编辑与模型理解 :foundation model(title,abstract)

AI总结 Rem3Di是一种表征学习框架,利用原子基础模型潜在特征生成可转移分子描述符,用于性质预测和虚拟筛选。它能捕捉分子手性,在药物性质基准测试中表现出色,还能区分过渡金属配合物,为化学机器学习提供新途径。

Comments An earlier version of this work appeared at the NeurIPS 2025 Workshop on Symmetry and Geometry in Neural Representations (NeurReps). Workshop version: https://openreview.net/forum?id=jOmZsvXoK5

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.20817 2026-07-23 cs.CL 版本更新 77%

EssayCBM: Rubric-Aligned Concept Bottleneck Models for Transparent Essay Grading

EssayCBM: 基于评分标准的概念瓶颈模型用于透明的作文评分

Kumar Satvik Chaudhary, Chengshuai Zhao, Fan Zhang, Garima Agrawal, Yuli Deng, Huan Liu

机构 * Arizona State University(亚利桑那州立大学)

专题命中 知识编辑与模型理解 :LLM(abstract,abstract_cn);language model(abstract);分类 cs.CL

AI总结 EssayCBM通过分解为八种可解释的写作概念,提供透明的作文评分方法,使教师能够检查和调整评分标准级别的预测。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.20146 2026-07-23 cs.CL 新提交 70%

Gotta Catch them all: the modes of Sycophancy

必须抓住它们:谄媚的模式

Shreyans Jain, Alexandra Yost, Amirali Abdullah

机构 * Thoughtworks(Thoughtworks公司) Southern Utah University (SUU)(南犹他大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究大语言模型谄媚行为,通过分析948种社会压力情境下的三种假设模式,发现其并非单一倾向,而是结构化家族,内部表示从第14层起线性可分,出现在不同处理阶段,依赖不同注意力电路,为更精确测量和干预提供依据。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.04081 2026-07-23 cs.CL 版本更新 70%

Abstraction Induces the Brain Alignment of Language and Speech Models

抽象诱导语言和语音模型的脑对齐

Emily Cheng, Aditya R. Vaidya, Richard Antonello

机构 * The University of Texas at Austin, USA(德克萨斯大学奥斯汀分校) Zuckerman Mind Brain Behavior Institute, Columbia University, USA(祖克曼心智-大脑-行为研究所,哥伦比亚大学)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究发现语言和语音模型通过共享的意义抽象与大脑对齐,中间层的高内在维度和语义内容增强了大脑预测性。

Comments ICML 2026 camera-ready version

Journal ref ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.18543 2026-07-23 cs.AI cs.CL cs.SE 版本更新 62%

CEO-Bench: Can Agents Play the Long Game?

CEO-Bench:智能体能否玩转长期博弈?

Haozhe Chen, Karthik Narasimhan, Zhuang Liu

机构 * Princeton University(普林斯顿大学)

专题命中 知识编辑与模型理解 :language model(abstract);分类 cs.CL、cs.AI

AI总结 提出CEO-Bench,通过模拟500天运营初创公司的任务,评估语言模型智能体在长期、不确定、动态环境下的综合决策能力。

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 其他LLM 18 篇

2401.04155 2026-07-23 q-bio.QM cs.CL 90%

Advancing bioinformatics with large language models: components, applications and perspectives

用大型语言模型推进生物信息学:组件、应用与展望

Jiajia Liu, Mengyuan Yang, Yankai Yu, Haixia Xu, Tiangang Wang, Kang Li, Xiaobo Zhou

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);foundation model(abstract)

AI总结 本文探讨了大型语言模型在生物信息学中的应用,涵盖其核心组件、关键技术和实际应用,并提出优化策略以推动该领域的发展。

Comments 5 main figures

Journal ref Briefings in Bioinformatics, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.19315 2026-07-23 cs.SE 89%

Improving LLM-Driven Test Generation by Learning from Mocking Information

通过学习模拟信息改进LLM驱动的测试生成

Jamie Lee, Flynn Teh, Hengcheng Zhu, Mengzhen Li, Mattia Fazzini, Valerio Terragni

专题命中 其他LLM :LLM(title,title_cn);large language model(abstract);language model(abstract)

AI总结 本文提出MOCKMILL方法,利用开发者编写测试中的模拟信息自动生成测试用例,通过迭代生成与修复过程提升测试覆盖率和有效性。

Comments Accepted for publication in ICST 2026 (AIST workshop). This arXiv version is the authors' accepted manuscript

Journal ref IEEE Conference on Software Testing, Verification and Validation Workshop 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.16741 2026-07-23 cs.LG 版本更新 88%

The Anatomy of a Truth Direction: Knowledge-Dependent Dimensionality, a Relational Law, and a Convergent Category Geometry in Small Language Models

真相方向剖析:小语言模型中依赖知识的维度、关系定律与收敛类别几何

Francesco Karim Vicidomini

专题命中 其他LLM :language model(title,abstract);small language model(title);large language model(abstract);分类 cs.LG

AI总结 研究小语言模型中真相方向,通过无训练定向探针及多模型实验,探讨真相维度与知识的关系、架构组件作用及方向混合情况,揭示关系定律与知识门控定律,表明混合几何属知识领域。

Comments Version 2: Expanded with a replication campaign on a third model family (Gemma-2-2b). Introduces exact decomposition for sandwich normalization, quantifies the knowledge gate via classical attenuation (Spearman, 1904), and identifies model-private geometry. Text revised, figures unchanged. Code and data: https://github.com/Francesco-Marhel/TruthProbe

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19967 2026-07-23 physics.soc-ph cs.AI cs.CY 新提交 87%

When Shippers Become Algorithms: Candidate Exposure, Information Design, and the Concentration of LLM-Mediated Freight Markets

当托运人成为算法:候选者曝光、信息设计与大语言模型介导的货运市场集中度

Takahiro Ezaki, Naoto Imura, Katsuhiro Nishinari

机构 * Research Center for Advanced Science and Technology, The University of Tokyo(东京大学先进科学与技术研究中心) Department of Aeronautics and Astronautics, School of Engineering, The University of Tokyo(东京大学工学部航空宇宙学系)

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究托运人委托大语言模型代理选择承运人对货运市场的影响及平台设计应对策略,通过基于代理的模拟发现代理趋同、集中度随候选列表数量变化等风险,披露承运人剩余日运力可有效应对,凸显平台信息设计的重要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.00740 2026-07-23 cs.CL cs.LG 版本更新 87%

LaSEr-Edit: Localized Span-level Error Editing with Energy-based Localization

LaSEr-Edit:基于能量定位的局部跨度级错误编辑

Hye Ryung Son, Saehee Eom, Mooho Song, Jay-Yoon Lee

机构 * Graduate School of Data Science(数据科学研究生院) Seoul National University(首尔国立大学) Georgia Institute of Technology(佐治亚理工学院)

专题命中 其他LLM :LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

AI总结 研究针对大语言模型满足约束问题,提出LaSEr-Edit方法。利用轻量级特定任务的基于能量的模型进行错误定位,提出LaSEr-LLM Edit和LaSEr-EBM Edit两种文本修订方法,实验表明该方法能有效控制文本,多约束下也表现良好。

Comments 38 pages, 7 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19629 2026-07-23 cs.CL cs.AI cs.MA 新提交 84%

Adaptive Capitulation: A Structural Failure Mode of LLM Responses in Vulnerability Contexts

适应性屈服:脆弱情境下大语言模型响应的一种结构故障模式

Eunna Lee

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 研究大语言模型在脆弱情境下的响应问题,通过实验刻画了适应性屈服故障模式,表明困境是结构性的,进而提出架构中立的最小重新归因充分性原则来保留自主重新归因途径。

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.20351 2026-07-23 cs.CV cs.CL 新提交 83%

Test-Time Training for Modality Order Consistency in Vision-Language Models

视觉语言模型中模态顺序一致性的测试时训练

Aditi Gupta, Yossi Gandelsman

机构 * University of Chicago(芝加哥大学)

专题命中 其他LLM :language model(title,abstract);prompting(abstract);分类 cs.CL

AI总结 研究发现视觉语言模型对图像和问题呈现顺序敏感,利用此设计测试时训练方法,缩小模态顺序差距,使两种顺序相互一致,定位顺序失败区域,证明该方法可缓解故障并提升性能。

Comments 16 pages, 7 figures, preprint

详情

展开后加载摘要…

URL PDF HTML 收藏