arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-06-09 至 2026-06-09 共收录 55 信号源:cs.CL, cs.AI, cs.LG

1. 领域大模型 55 篇

2510.12171 2026-06-09 cs.AI 版本更新 92%

MatSciBench: Benchmarking the Reasoning Ability of Large Language Models in Materials Science

MatSciBench: 基准测试大型语言模型在材料科学中的推理能力

Junkai Zhang, Jingru Gan, Xiaoxuan Wang, Zian Jia, Changquan Gu, Jianpeng Chen, Yanqiao Zhu, Mingyu Derek Ma, Dawei Zhou, Ling Li, Wei Wang

机构 * University of California, Los Angeles Computer Science Department(加州大学洛杉矶分校计算机科学系) University of Pennsylvania Department of Materials Science and Engineering(宾夕法尼亚大学材料科学与工程系) Virginia Tech Department of Computer Science(弗吉尼亚理工大学计算机科学系)

专题命中 领域大模型 :LLM(summary_cn,abstract);large language model(title,abstract);language model(title,abstract);prompting(abstract)

AI总结 提出MatSciBench基准,包含1340道大学级材料科学问题,覆盖6个主领域和31个子领域,评估LLM推理能力,发现当前模型在领域知识、计算和图表理解方面存在局限。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07526 2026-06-09 cs.CL cs.AI 新提交 92%

GraphLoRA: Structure-Aware Low-Rank Adaptation for Large Language Model Recommendation

GraphLoRA: 面向大语言模型推荐的结构感知低秩适配

Lin Mu, Guoji Wang, Li Ni, Lei Sang, Zhize Wu, Peiquan Jin, Yiwen Zhang

机构 * Anhui University(安徽大学) Hefei University(合肥大学) University of Science and Technology of China(中国科学技术大学)

专题命中 领域大模型 :LLM(summary_cn,abstract);large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI

AI总结 提出GraphLoRA框架,通过在低秩适配路径中嵌入可训练的图消息传递网络,实现结构信号传播,从而深度融合图结构与文本语义,提升LLM推荐性能。

Comments ACL 2026 findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07521 2026-06-09 cs.CL cs.AI 新提交 91%

Evaluating Hallucinations in Domain-Adapted Large Language Models

评估领域自适应大语言模型中的幻觉现象

Sanchita Porwal, Sai Prasath S, Xingjian Bi, Madelyn Scandlen

机构 * College of Computing, Georgia Institute of Technology(佐治亚理工学院计算学院)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);LLM(abstract,abstract_cn);分类 cs.CL、cs.AI

AI总结 本研究通过微调Llama-2模型,测试其记忆、回忆和推理能力,发现领域自适应大语言模型在生成新领域特定信息时存在幻觉问题,表明仅靠微调难以有效缓解幻觉。

Comments 13 pages, 2 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.09738 2026-06-09 cs.CV 新提交 90%

HDSL: A Hierarchical Domain-Specific Language for Structured 3D Indoor Scene Generation and Localized Editing with LLM Agents

HDSL:一种用于结构化3D室内场景生成和基于LLM智能体局部编辑的层次化领域特定语言

Letian Li, Chao Shen, Shuzhao Xie, Chenghao Gu, ZhengXiao He, Yu Meng, Xin Yang, Wenyuan Jiang, Zhi Wang

机构 * SIGS, Tsinghua University(清华大学深圳国际研究生院) Nankai University(南开大学) University of Arizona(亚利桑那大学) Zhejiang University(浙江大学) ETH Zurich(苏黎世联邦理工学院)

专题命中 领域大模型 :LLM(title,title_cn);language model(abstract)

AI总结 提出HDSL语言,以树结构表示室内场景,结合LLM智能体生成、多模态检索和力导向布局优化,实现结构化场景生成与局部编辑,显著提升对象覆盖率和编辑效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07681 2026-06-09 cs.SE cs.AI cs.CE cs.MA 新提交 90%

Systematic LLM Translation of Legacy Scientific Code to Differentiable Frameworks: Application to a Land Surface Model

将遗留科学代码系统性地LLM翻译为可微分框架:以陆面模型为例

Aya Lahlou, Linnia Hawkins, Pierre Gentine

机构 * University of California, Los Angeles(加州大学洛杉矶分校) NASA Goddard Space Flight Center(国家航空航天局戈达德空间飞行中心)

专题命中 领域大模型 :LLM(title,title_cn);分类 cs.AI

AI总结 提出基于LLM的五阶段流水线,将遗留Fortran代码自动翻译为JAX可微分框架,在CLM-ml-v2模型上实现完整雅可比矩阵计算和24倍加速。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21239 2026-06-09 cs.CL 版本更新 90%

A Unified LLM-Adaptable Framework for Cold-Start Cognitive Diagnosis

面向冷启动认知诊断的统一LLM可适配框架

Zihan Yao, Chentao Song, Yu He, Tianyu Qi, Jian Zhang, Weiping Fu, Jun Liu

机构 * University of Science and Technology of China(中国科学技术大学)

专题命中 领域大模型 :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 提出LMCD框架,通过知识扩散和语义-认知融合两阶段,利用大语言模型增强冷启动场景下的认知诊断性能。

Comments Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.12450 2026-06-09 cs.CY 版本更新 89%

Empirical Modeling of Therapist-Client Dynamics in Psychotherapy Using LLM-Based Assessments

基于LLM评估的心理治疗中治疗师-来访者动态的实证建模

Angela Chen, Siwei Jin, Canwen Wang, Holly Swartz, Tongshuang Wu, Robert E Kraut, Haiyi Zhu

专题命中 领域大模型 :LLM(title,title_cn);large language model(abstract);language model(abstract)

AI总结 本研究利用大语言模型测量治疗师行为、关系质量和来访者结果,结合结构方程模型分析约2000小时转录数据,发现共情和探索直接促进来访者自我表露和情绪变化,而融洽关系起调节作用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.05114 2026-06-09 cs.SE 89%

How Software Engineering Students Use LLMs to Write Research Papers: An Experience Report

软件工程学生如何使用LLM撰写研究论文:经验报告

Ronnie de Souza Santos, Maria Teresa Baldassarre, Cleyton Magalhaes, Italo Santos

专题命中 领域大模型 :LLM(title_cn,summary_cn);large language model(abstract);language model(abstract)

AI总结 本文报告了一项教育经验,通过分析146份学生披露声明,探讨学生在实证方法作业中如何整合LLM进行头脑风暴、方法澄清、结果组织和写作润色,并讨论其教育意义。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.09489 2026-06-09 cs.AI 新提交 89%

LLM-Orchestrated Conformance Checking in Stroke Care Without Computer-Interpretable Guidelines

LLM编排的卒中护理合规性检查无需计算机可解释指南

Giorgio Leonardi, Stefania Montani, Manuel Striani, Alessandro Canessa, Delfina Ferrandi

机构 * Computer Science Institute, DiSIT, University of Piemonte Orientale(皮埃蒙特东方大学计算机科学研究所) Integrated Laboratory of AI and Medical Informatics, DAIRI, SS. Antonio e Biagio e Cesare Arrigo Hospital(圣安东尼奥、比亚焦与切萨雷·阿里戈医院DAIRI人工智能与医学信息学综合实验室)

专题命中 领域大模型 :LLM(title,title_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出基于大语言模型编排的模块化框架,从非结构化临床文本和指南中自动提取患者轨迹、识别规范规则并计算合规性指标,在卒中护理领域验证了86%以上的轨迹合规。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07712 2026-06-09 cond-mat.mtrl-sci cs.AI 新提交 89%

MatMind: A Structure-Activity Knowledge-Driven Generative Foundation Model for Materials Science

MatMind:面向材料科学的结构-活性知识驱动生成基础模型

Zhan'ao Yao, Boxuan Zhang, Jingyuan Shu, Xiaoyu Wu, Rongyan Wang, Linjing Li, Dajun Zeng, Yudong Yao, Tingwei Chen, Youwei Wang, Xiaolin Zhao, Jiahui Shi, Jianjun Liu

机构 * State Key Laboratory of High Performance Ceramics(高性能陶瓷国家重点实验室) Shanghai Institute of Ceramics, Chinese Academy of Sciences(中国科学院上海陶瓷研究所) Center of Materials Science and Optoelectronics Engineering, University of Chinese Academy of Sciences(中国科学院大学材料科学与光电子工程中心) School of Chemistry and Materials Science, Hangzhou Institute for Advanced Study, University of Chinese Academy of Sciences(中国科学院大学杭州先进研究所化学与材料科学学院) State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,中国科学院自动化研究所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Beijing Wenge Technology Co., Ltd.(北京文格科技有限公司) College of Medicine and Biological Information Engineering, Northeastern University(东北大学医学与生物信息工程学院)

专题命中 领域大模型 :foundation model(title,abstract);LLM(abstract,abstract_cn);large language model(abstract);language model(abstract)

AI总结 提出MatMind,一种基于大语言模型的晶体材料生成基础模型,通过结构-活性知识注入、双头架构和物理信息强化学习,在性质预测、无条件生成和条件生成任务上超越专用模型。

Comments 29 pages, 5 figures, including references

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.09632 2026-06-09 cs.CL 新提交 88%

Civil Court Simulation with Large Language Models

基于大型语言模型的民事法庭模拟

Yifan Chen, Haitao Li, Kaiyuan Zhang, Yueyue Wu, Qingyao Ai, Yiqun Liu

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) Tsinghua University(清华大学)

专题命中 领域大模型 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

AI总结 提出多智能体民事法庭模拟框架,通过五阶段审判程序、记忆模块和法规检索实现可靠判决,在责任分配和多项裁决上表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07570 2026-06-09 cs.DL cs.LG 新提交 88%

Can LLMs extract scientific consensus? A case study in high-temperature superconductivity

LLMs能否提取科学共识?以高温超导为例

Mouyang Cheng, Wenhao He, Zhuotao Jin, Bowen Yu, Ju Li, Boris Kozinsky, Yao Wang, Pavel Volkov, Liangzi Deng, Ching-Wu Chu, Xiao-Gang Wen, Mingda Li

机构 * Center for Computational Science and Engineering, MIT(MIT计算科学与工程中心) Department of Materials Science and Engineering, MIT(MIT材料科学与工程系) Department of Physics, MIT(MIT物理系) Department of Nuclear Science and Engineering, MIT(MIT核科学与工程系) John A. Paulson School of Engineering and Applied Sciences, Harvard University(哈佛大学约翰·A·保罗森工程与应用科学学院) Department of Chemistry, Emory University(埃默里大学化学系) Department of Physics, University of Connecticut(康涅狄格大学物理系) Department of Physics and Texas Center for Superconductivity, University of Houston(休斯顿大学物理系和德克萨斯超导中心)

专题命中 领域大模型 :LLM(summary_cn,abstract);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 本研究以高温超导领域为测试平台,利用近18,000篇高被引文献构建知识图谱,发现LLM提取的表征能恢复出连贯且物理可解释的结构,表明LLM可作为解码竞争性科学知识的可扩展工具。

Comments 23 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.08969 2026-06-09 cs.CL cs.AI 新提交 87%

CARE: A Conformal Safety Layer for Medical Summarization

CARE:面向医学摘要的保形安全层

Suhana Bedi, Bridget Lin, Anson Y. Zhou, Chloe O. Stanwyck, Jenelle A. Jindal, Sanmi Koyejo, David Stutz, Nigam H. Shah

机构 * Stanford University(斯坦福大学) Google DeepMind(谷歌DeepMind)

专题命中 领域大模型 :LLM(summary_cn,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 提出CARE方法,通过保形风险控制为LLM医学摘要提供校准的遗漏和幻觉标记,在保证安全性的同时减少审查负担。

Comments 29 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.27441 2026-06-09 cs.IR cs.LG 版本更新 86%

A Unified Structured Query Understanding Framework for Industrial Semantic Search

面向工业语义搜索的统一结构化查询理解框架

Ping Liu, Qianqi Shen, Jianqiang Shen, Chunnan Yao, Kevin Kao, Rajat Arora, Dan Xu, Baofen Zheng, Yunxiang Ren, Benjamin Le, Ali Hooshmand, Igor Lapchuk, Juan Bottaro, Raghavan Muthuregunathan, Caleb Johnson, Liangjie Hong, Jingwei Wu, Wenjing Zhang

机构 * LinkedIn Corporation(领英公司)

专题命中 领域大模型 :SLM(summary_cn,abstract);language model(abstract);small language model(abstract);分类 cs.LG

AI总结 提出一个统一的结构化查询理解系统,将多个异构功能整合到单个小语言模型(SLM)中,并引入Query Illuminator框架用于自动标注和评估,在LinkedIn的职位搜索和人员搜索中验证了效果。

Comments Accepted by KDD-ADS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00462 2026-06-09 cs.CY 版本更新 86%

AI Self-preferencing in Algorithmic Hiring: Empirical Evidence and Insights

算法招聘中的AI自我偏好:实证证据与洞见

Jiannan Xu, Gujie Li, Jane Yi Jiang

专题命中 领域大模型 :LLM(summary_cn,abstract);large language model(abstract);language model(abstract)

AI总结 通过大规模简历实验和模拟,发现LLM在招聘中偏好自己生成的简历,导致使用同模型简历的候选人获得23%-60%的短名单优势,且可通过干预减少50%以上偏差。

Comments This paper has been accepted as a non-archival submission at EAAMO 2025 and AIES 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.12263 2026-06-09 cs.CL cs.AI cs.LG 版本更新 85%

Multimodal Generative Engine Optimization: Rank Manipulation for Vision-Language Model Rankers

多模态生成式引擎优化:针对视觉-语言模型排序器的排名操纵

Yixuan Du, Chenxiao Yu, Haoyan Xu, Ziyi Wang, Yue Zhao, Xiyang Hu

机构 * Georgetown University(乔治城大学) University of Southern California(南加州大学) University of Maryland, College Park(马里兰大学学院公园分校) Arizona State University(亚利桑那州立大学)

专题命中 领域大模型 :language model(title,abstract);foundation model(abstract,comments);分类 cs.CL、cs.AI、cs.LG

AI总结 提出多模态生成式引擎优化(MGEO)方法,通过联合优化图像扰动和文本后缀,利用视觉-语言模型内部跨模态知识耦合,实现对产品排名的有效操纵,揭示了多模态基础模型知识基础的脆弱性。

Comments Proceedings of the 4th Workshop on Towards Knowledgeable Foundation Models (KnowFM) at ACL 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07830 2026-06-09 cs.SE 新提交 85%

Academic Integrity and Emotional Responses to Inappropriate LLM Use in Software Engineering Education

学术诚信与软件工程教育中不当使用大语言模型的情感反应

Ronnie de Souza Santos, Italo Santos, Giuseppe Destefanis, Cleyton Magalhaes, Mairieli Wessel

专题命中 领域大模型 :LLM(title,abstract_cn);large language model(abstract);language model(abstract)

AI总结 本研究通过116名本科生的横截面调查,探讨软件工程学生在感知学术不当使用大语言模型后的情感反应,发现漠不关心最常见,内疚、焦虑、宽慰和满足感与不同情境相关。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.01869 2026-06-09 cs.AI 版本更新 84%

WorldCoder-Bench: Benchmarking Physically Grounded 3D World Synthesis

WorldCoder-Bench:物理接地3D世界合成基准

Shuo Lu, Yinuo Xu, Kecheng Yu, Siru Jiang, Yongcan Yu, Yubin Wang, Haitao Yang, Yuxiang Zhang, Bin Wang, Ran He, Jian Liang

机构 * NLPR & MAIS, CASIA(中国科学院自动化研究所与模式识别国家重点实验室) Huawei Noah’s Ark Lab(华为诺亚实验室)

专题命中 领域大模型 :LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 提出WorldCoder-Bench基准,通过StateProbe协议评估LLM生成Three.js 3D世界的物理正确性和交互可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.06469 2026-06-09 cs.CL 版本更新 84%

ClinicalBench: Can LLMs Beat Traditional ML Models in Clinical Prediction?

ClinicalBench: 大型语言模型能在临床预测中击败传统机器学习模型吗?

Canyu Chen, Jian Yu, Shan Chen, Che Liu, Zhongwei Wan, Shuang Zhou, Yuan Luo, Rui Zhang, Danielle Bitterman, Fei Wang, Kai Shu

机构 * Department of Computer Science Northwestern University Evanston USA(计算机科学系西北大学艾文斯顿美国) Department of Computer Science University of Texas at Austin Austin USA(计算机科学系德克萨斯大学奥斯汀美国) Boston Children's Hospital, Harvard Medical School Boston USA(波士顿儿童医院哈佛医学院波士顿美国) Department of Computer Science Imperial College London London UK(计算机科学系伦敦帝国学院伦敦英国) Department of Computer Science Ohio State University Columbus USA(计算机科学系俄亥俄州立大学哥伦布美国) Massachusetts General Hospital, Harvard Medical School Boston USA(麻省总医院哈佛医学院波士顿美国) Department of Preventive Medicine, Feinberg School of Medicine Northwestern University Chicago USA(预防医学系费因伯格医学院西北大学芝加哥美国) Division of Computational Health Sciences, Department of Surgery University of Minnesota Minneapolis USA(计算健康科学部外科部明尼苏达大学明尼阿波利斯美国) Department of Population Health Sciences, Weill Cornell Medicine Cornell University New York USA(流行病学与公共卫生系韦尔·科恩医学中心康奈尔大学纽约美国) Department of Computer Science Emory University Atlanta USA(计算机科学系埃默里大学亚特兰大美国) Northwestern University(西北大学) University of Texas at Austin(德克萨斯大学奥斯汀) Boston Children's Hospital, Harvard Medical School(波士顿儿童医院哈佛医学院) Imperial College London(伦敦帝国学院) Ohio State University(俄亥俄州立大学) Massachusetts General Hospital, Harvard Medical School(麻省总医院哈佛医学院) University of Minnesota(明尼苏达大学) Cornell University(康奈尔大学) Emory University(埃默里大学)

专题命中 领域大模型 :LLM(summary_cn);large language model(abstract);language model(abstract);prompting(abstract)

AI总结 构建ClinicalBench基准,通过三个临床预测任务比较14个通用和8个医学LLM与11个传统ML模型,发现LLM在临床预测上仍无法超越传统ML模型。

Comments Accepted to Proceedings of KDD 2026. The first two authors contributed equally. 12 pages for main paper, 62 pages including appendix. Project website: https://clinicalbench.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07853 2026-06-09 cs.CL cs.AI 新提交 84%

Beyond English benchmarks: clinical llm evaluation in Brazilian Portuguese

超越英语基准:巴西葡萄牙语临床大语言模型评估

Giordano de Pinho Souza, Glaucia Melo, Josefino Cabral Melo Lima, Daniel Schneider

机构 * Federal University of Rio de Janeiro(里约热内卢联邦大学) Toronto Metropolitan University(多伦多都会大学)

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 提出首个双语临床基准ClinicalBr,基于巴西病例报告构建,评估四个模型发现葡萄牙语-英语性能差距具有任务依赖性,诊断检索英语优势明显,其他任务差距消失。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.09030 2026-06-09 cs.LG cs.AI cs.CL 新提交 83%

TRIAGE: Dialectical Reasoning for Explainable Risk Prediction on Irregularly Sampled Medical Time Series with LLMs

TRIAGE: 基于辩证推理的不规则采样医学时间序列风险可解释预测方法

Hyeongwon Jang, Gyouk Chu, Changhun Kim, Joonhyung Park, Hangyul Yoon, Eunho Yang

机构 * KAIST(韩国科学技术院) AITRICS University of Wisconsin-Madison(威斯康星大学麦迪逊分校)

专题命中 领域大模型 :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 提出TRIAGE框架,利用大语言模型对竞争性临床结果生成辩证推理,缓解风险极化,实现连续风险评分与可解释推理,在三个基准上AUPRC提升3.3%,校准误差降低81%。

Comments Code is available at https://github.com/HyeongWon-Jang/TRIAGE

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07688 2026-06-09 cs.IR cs.AI cs.CL cs.LG 新提交 83%

TRACER: Token ReAssignment for Concept ERasure in Generative Recommendation

TRACER: 面向生成式推荐中概念擦除的令牌重分配

Ziheng Chen, Jiali Cheng, Zezhong Fan, Hadi Amiri, Diyuan Wu, Gabriele Tolomei, Yang Zhang

机构 * Stony Brook University(石英布鲁克大学) University of Massachusetts Lowell(马萨诸塞大学洛厄尔分校) Columbia University(哥伦比亚大学) Institute of Science and Technology Austria(奥地利科学技术研究院) Sapienza University of Rome(罗马大学 sapienza) National University of Singapore(新加坡国立大学)

专题命中 领域大模型 :LLM(abstract,abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 针对生成式推荐中概念遗忘与推荐效用冲突的问题,提出基于令牌重分配的概念遗忘框架TRACER,通过将概念相关物品重分配给替代令牌并引入一致性正则化,有效移除目标概念同时保持推荐效用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09506 2026-06-09 cs.CE 83%

Beyond Knowledge to Agency: Evaluating Expertise, Autonomy, and Integrity in Finance with CNFinBench

超越知识到能动性:用CNFinBench评估金融领域的专业知识、自主性和完整性

Jinru Ding, Chao Ding, Yidong Jiang, Wenrao Pang, Boyi Xiao, Zhiqiang Liu, Jiayuan Chen, Yun Zhong, Tiantian Yuan, Junming Guan, Dawei Cheng, Jie Xu

专题命中 领域大模型 :LLM(summary_cn,abstract_cn);large language model(abstract);language model(abstract)

AI总结 提出CNFinBench基准,通过专业知识、自主性和完整性三维度评估LLM在金融领域的代理能力,并引入HICS指标量化多轮对抗攻击下的安全退化。

Journal ref Proceedings of the 32nd ACM SIGKDD Conference on Knowledge Discovery and Data Mining (KDD 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.14883 2026-06-09 cs.CL cs.CY 83%

OATH-Frames: Characterizing Online Attitudes Towards Homelessness with LLM Assistants

OATH-Frames: 利用大语言模型助手分析在线对无家可归者的态度

Jaspreet Ranjit, Brihi Joshi, Rebecca Dorn, Laura Petry, Olga Koumoundouros, Jayne Bottarini, Peichen Liu, Eric Rice, Swabha Swayamdipta

机构 * Dept. of Computer Science, University of Southern California(计算机科学系,南加州大学) Suzanne-Dwork School of Social Work, University of Southern California(苏兹曼-道克社会工作学院,南加州大学)

专题命中 领域大模型 :LLM(title);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 本文提出OATH-Frames框架,通过大语言模型分析社交媒体上的无家可归者态度,提升大规模分析效率并揭示态度趋势。

Comments Project website: https://dill-lab.github.io/oath-frames/, EMNLP Main 2024

Journal ref In Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07704 2026-06-09 cs.LG cs.AI 新提交 82%

FunctionEvolve: Structure-Guided Symbolic Regression with LLMs

FunctionEvolve: 基于结构引导的符号回归与大型语言模型

Zeyu Xia, Jun Zhu, Dong Yan

机构 * Bosch Center for Artificial Intelligence(博世人工智能中心) Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系) Tsinghua-Bosch Joint Center for ML, Tsinghua University(清华大学-博世联合机器学习中心)

专题命中 领域大模型 :LLM(summary_cn,abstract);分类 cs.AI、cs.LG

AI总结 提出FunctionEvolve框架,利用表达式树组织符号回归搜索,通过结构摘要、局部树编辑和结构感知系数拟合,在LLM-SRBench合成子集上以Claude Opus 4.6实现82.9%的SA@50,较同基线提升4.5倍。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07951 2026-06-09 cs.CL cs.AI cs.LG 新提交 82%

From `May' to `Is': Certainty Distortion in Language Model Rewriting

从“可能”到“是”:语言模型改写中的确定性扭曲

Catarina G Belem, Shang Wu, Hongyu Yao, Mark Steyvers, Sameer Singh, Padhraic Smyth

机构 * University of California Irvine(加利福尼亚大学尔湾分校) Massachusetts Institute of Technology(麻省理工学院)

专题命中 领域大模型 :language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

AI总结 研究语言模型在改写任务中系统性增加表达确定性的偏差,提出基于人群判断的评估指标,发现高达75%的输出存在确定性扭曲,且模型更倾向于提高确定性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.08307 2026-06-09 cs.CL 新提交 81%

Understanding the Sociocultural Dimensions of Mental Health Discourse in Arabic-Language X Communities

理解阿拉伯语X社区中心理健康话语的社会文化维度

Amal Alqahtani, Rana Salama, Mona Diab

机构 * King Saud University(沙特国王大学) Cairo University(开罗大学) Carnegie Mellon University(卡内基梅隆大学)

专题命中 领域大模型 :LLM(summary_cn,abstract);分类 cs.CL

AI总结 通过GPT-4.1识别个人披露的推特用户,分析边缘型人格障碍、双相障碍和ADHD相关话语,发现不同病症的词汇模式差异,提出可复用的LLM辅助披露流程和文化关键词框架。

Comments Accepted to the SMM4H-HeaRD Workshop, co-located with the 64th Annual Meeting of the Association for Computational Linguistics (ACL 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.08847 2026-06-09 cs.CV cs.AI cs.LG 新提交 81%

BLM-SGAN: Bidirectional Language Modeling for Semantic-Spatial Text-to-Image Generation

BLM-SGAN: 用于语义-空间文本到图像生成的双向语言建模

Ahmed Abdelmoneim Mazrou, Haidy Maher El-Amir, Ali Hamdi

机构 * Faculty of Computer Science, MSA University, Egypt(MSA大学计算机科学学院,埃及)

专题命中 领域大模型 :language model(title,abstract);分类 cs.AI、cs.LG

AI总结 提出BLM-SGAN模型,利用BERT的双向注意力机制捕获长程依赖,解决GAN在文本到图像生成中的梯度消失和序列处理限制,在鸟类图像生成上达到SOTA。

Comments Published in ICACIn 2024. Appears in Advances on Intelligent Computing and Data Science II, Lecture Notes on Data Engineering and Communications Technologies, vol. 254, Springer, 2025

Journal ref Advances on Intelligent Computing and Data Science II (ICACIn 2024), Lecture Notes on Data Engineering and Communications Technologies, vol. 254, Springer, Cham, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.09637 2026-06-09 cs.SE 新提交 80%

Agentic Persona Generation with Critique-Refinement: An Industrial Evaluation

基于批评-改进的智能体角色生成:工业评估

Mohammad Hossein Amini, David Dewar, Shiva Nejati, Mehrdad Sabetzadeh

专题命中 领域大模型 :LLM(summary_cn,abstract)

AI总结 提出PerGent方法,通过迭代批评-改进循环利用LLM智能体生成角色,在工业部署中专家认可率达96.9%,优于单次生成方法。

Comments Accepted in the Industry Track of the Requirements Engineering (RE) 2026 Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.29475 2026-06-09 cs.CL cs.AI cs.CE cs.HC 版本更新 79%

MOOSE-Copilot: A Web-Based Interactive Assistant for Unified Exploratory and Fine-Grained Scientific Hypothesis Discovery

MOOSE-Copilot:一个基于网络的交互式助手,用于统一探索性和细粒度科学假设发现

Hongran An, Zonglin Yang

机构 * Central Conservatory of Music(中央音乐学院) Nanyang Technological University(南洋理工大学)

专题命中 领域大模型 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

AI总结 提出MOOSE-Copilot,通过形式化的人机交互协议,将发散性探索和收敛性细化统一,利用蓝图、路由和反馈三种信号引导生成,显著优于纯自主基线。

Comments Accepted to ACL 2026 (System Demonstrations)

详情

展开后加载摘要…

URL PDF HTML 收藏