arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7552 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7552 篇

2407.07061 2024-07-11 cs.CL 77%

Internet of Agents: Weaving a Web of Heterogeneous Agents for Collaborative Intelligence

Weize Chen, Ziming You, Ran Li, Yitong Guan, Chen Qian, Chenyang Zhao, Cheng Yang, Ruobing Xie, Zhiyuan Liu, Maosong Sun

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.03621 2024-07-08 cs.CL 77%

The Mysterious Case of Neuron 1512: Injectable Realignment Architectures Reveal Internal Characteristics of Meta's Llama 2 Model

Brenden Smith, Dallin Baker, Clayton Chase, Myles Barney, Kaden Parker, Makenna Allred, Peter Hu, Alex Evans, Nancy Fulda

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments 21 pages, 17 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.15214 2024-06-24 cs.CL 77%

Unsupervised Extraction of Dialogue Policies from Conversations

Makesh Narsimhan Sreedhar, Traian Rebedea, Christopher Parisien

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07913 2024-06-13 cs.CL cs.IR 77%

DeTriever: Decoder-representation-based Retriever for Improving NL2SQL In-Context Learning

Yuxi Feng, Raymond Li, Zhenan Fan, Giuseppe Carenini, Mohammadreza Pourreza, Weiwei Zhang, Yong Zhang

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.11905 2024-06-06 cs.CL 77%

Learning to Edit: Aligning LLMs with Knowledge Editing

Yuxin Jiang, Yufei Wang, Chuhan Wu, Wanjun Zhong, Xingshan Zeng, Jiahui Gao, Liangyou Li, Xin Jiang, Lifeng Shang, Ruiming Tang, Qun Liu, Wei Wang

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments 17 pages, 8 figures, 9 tables. ACL 2024 main camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.19660 2024-04-04 cs.CL 77%

Interpretable-by-Design Text Understanding with Iteratively Generated Concept Bottleneck

Josh Magnus Ludan, Qing Lyu, Yue Yang, Liam Dugan, Mark Yatskar, Chris Callison-Burch

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.09410 2024-03-15 cs.CV cs.AI 77%

XCoOp: Explainable Prompt Learning for Computer-Aided Diagnosis via Concept-guided Context Optimization

Yequan Bie, Luyang Luo, Zhixuan Chen, Hao Chen

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);foundation model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.04852 2024-03-12 cs.LG 77%

Multi-Patch Prediction: Adapting LLMs for Time Series Representation Learning

Yuxuan Bian, Xuan Ju, Jiangtong Li, Zhijian Xu, Dawei Cheng, Qiang Xu

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.12713 2024-02-26 cs.CL 77%

Generating Zero-shot Abstractive Explanations for Rumour Verification

Iman Munire Bilal, Preslav Nakov, Rob Procter, Maria Liakata

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments Revised version of the original

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.15194 2024-02-16 cs.CL 77%

PokeMQA: Programmable knowledge editing for Multi-hop Question Answering

Hengrui Gu, Kaixiong Zhou, Xiaotian Han, Ninghao Liu, Ruobing Wang, Xin Wang

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments Our code is available at https://github.com/Hengrui-Gu/PokeMQA

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.01780 2024-01-04 cs.CL cs.IR 77%

Navigating Uncertainty: Optimizing API Dependency for Hallucination Reduction in Closed-Book Question Answering

Pierre Erbacher, Louis Falissar, Vincent Guigue, Laure Soulier

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.13040 2023-12-21 cs.CL 77%

Retrieval-augmented Multilingual Knowledge Editing

Weixuan Wang, Barry Haddow, Alexandra Birch

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.11795 2023-12-20 cs.CL 77%

MELO: Enhancing Model Editing with Neuron-Indexed Dynamic LoRA

Lang Yu, Qin Chen, Jie Zhou, Liang He

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments In Proceedings of The 38th Annual AAAI Conference on Artificial Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.10961 2023-11-21 cs.CL 77%

Journey of Hallucination-minimized Generative AI Solutions for Financial Decision Makers

Sohini Roychowdhury

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments 4 pages, 2 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.07897 2023-11-15 cs.CL 77%

CPopQA: Ranking Cultural Concept Popularity by LLMs

Ming Jiang, Mansi Joshi

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.13136 2023-09-26 cs.CV cs.AI 77%

Contextual Emotion Estimation from Image Captions

Vera Yang, Archita Srivastava, Yasaman Etesam, Chuxuan Zhang, Angelica Lim

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

Comments Accepted to ACII 2023. Project page: http://rosielab.github.io/emotion-captions/

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.11478 2023-09-21 cs.AI 77%

Fictional Worlds, Real Connections: Developing Community Storytelling Social Chatbots through LLMs

Yuqian Sun, Hanyi Wang, Pok Man Chan, Morteza Tabibi, Yan Zhang, Huan Lu, Yuheng Chen, Chang Hee Lee, Ali Asadipour

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.08594 2023-09-18 cs.CL 77%

"Merge Conflicts!" Exploring the Impacts of External Distractors to Parametric Knowledge Graphs

Cheng Qian, Xinran Zhao, Sherry Tongshuang Wu

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.12562 2023-08-25 cs.LG stat.ML 77%

Variational Information Pursuit with Large Language and Multimodal Models for Interpretable Predictions

Kwan Ho Ryan Chan, Aditya Chattopadhyay, Benjamin David Haeffele, Rene Vidal

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.03987 2023-08-15 cs.CL 77%

A Stitch in Time Saves Nine: Detecting and Mitigating Hallucinations of LLMs by Validating Low-Confidence Generation

Neeraj Varshney, Wenlin Yao, Hongming Zhang, Jianshu Chen, Dong Yu

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments update to include additional experiments

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.13865 2023-06-27 cs.CL 77%

IERL: Interpretable Ensemble Representation Learning -- Combining CrowdSourced Knowledge and Distributed Semantic Representations

Yuxin Zi, Kaushik Roy, Vignesh Narayanan, Manas Gaur, Amit Sheth

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments Accepted for publication at the KDD workshop on Knowledge-infused Machine Learning, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.10838 2023-06-06 cs.LG cs.PL 77%

ProgSG: Cross-Modality Representation Learning for Programs in Electronic Design Automation

Yunsheng Bai, Atefeh Sohrabizadeh, Zongyue Qin, Ziniu Hu, Yizhou Sun, Jason Cong

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.LG

Comments Requires further polishing

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.02029 2023-02-07 cs.CL 77%

Towards Few-Shot Identification of Morality Frames using In-Context Learning

Shamik Roy, Nishanth Sridhar Nakshatri, Dan Goldwasser

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL

Comments Accepted to the 5th Workshop on NLP and CSS at EMNLP 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.13371 2022-12-29 cs.AI cs.HC econ.GN q-fin.EC 77%

Measuring an artificial intelligence agent's trust in humans using machine incentives

Tim Johnson, Nick Obradovich

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.13757 2022-06-29 cs.CL cs.CY 77%

Flexible text generation for counterfactual fairness probing

Zee Fryer, Vera Axelrod, Ben Packer, Alex Beutel, Jilin Chen, Kellie Webster

专题命中 知识编辑与模型理解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.12264 2025-05-22 cs.LG cs.AI cs.CL stat.ML 76%

Uncertainty quantification in fine-tuned LLMs using LoRA ensembles

Oleksandr Balabanov, Hampus Linander

机构 * Stockholm University Department of Physics(斯德哥尔摩大学物理系) Department of Mathematical Sciences(数学科学系) Chalmers university of technology & University of Gothenburg(楚德勒斯技术大学及哥德堡大学) VERSES Research Lab(VERSES研究实验室)

专题命中 知识编辑与模型理解 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG;foundation model(comments)

Comments Accepted for ICLR2025 Workshop "Quantify Uncertainty and Hallucination in Foundation Models: The Next Frontier in Reliable AI"

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.16259 2026-07-21 cs.LG cs.CL stat.ML 新提交 76%

Quantifying Ranking Uncertainty in LLM Benchmarks

量化语言模型基准测试中的排名不确定性

Bitya Neuhof, Yuval Benjamini

专题命中 知识编辑与模型理解 :LLM(title);分类 cs.CL、cs.LG

AI总结 研究量化语言模型基准测试中排名不确定性问题,通过汇总成对假设检验来实现,分析了知识评估基准MMLU的不确定性来源并展示如何修改假设检验,指出MMLU各主题排名变异性大,比较模型时应考虑。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.26783 2026-07-08 cs.LG cs.CL 新提交 76%

Reproducibility Study of "AlphaEdit: Null-Space Constrained Knowledge Editing for Language Models"

可重复性研究:"AlphaEdit: 面向语言模型的零空间约束知识编辑"

Ananth K Suresh, Arya Hariharan

机构 * Independent(独立研究者)

专题命中 知识编辑与模型理解 :language model(title);分类 cs.CL、cs.LG

AI总结 本研究复现了AlphaEdit知识编辑方法,发现其在原始设置下结果可复现,但扩展到新架构和大量顺序编辑时性能下降,表明其理论保证有边界。

Comments 21 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.19462 2026-05-20 cs.LG cs.AI 76%

Quantifying the Pre-training Dividend: Generative versus Latent Self-Supervised Learning for Time Series Foundation Models

量化预训练红利:生成与潜在自监督学习在时间序列基础模型中的应用

Noam Major, Kathy Razmadze, Yoli Shavit

机构 * Faculty of Engineering, Bar-Ilan University(巴伊兰大学工程学院)

专题命中 知识编辑与模型理解 :foundation model(title);分类 cs.AI、cs.LG

AI总结 本文研究了自监督学习在时间序列中的应用,比较了生成范式与潜在对齐架构,发现预训练红利在异常检测和分类任务中显著提升,但在预测任务中效果有限,同时表明表示质量与数据来源无关,且在适度的架构深度下趋于稳定。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.08645 2026-04-13 cs.CV cs.AI cs.LG cs.RO 76%

3D-VCD: Hallucination Mitigation in 3D-LLM Embodied Agents through Visual Contrastive Decoding

3D-VCD:通过视觉对比解码缓解3D-LLM具身代理中的幻觉

Makanjuola Ogunleye, Eman Abdelrahman, Ismini Lourentzou

机构 * Virginia Tech(弗吉尼亚理工大学) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

专题命中 知识编辑与模型理解 :LLM(title);分类 cs.AI、cs.LG

AI总结 本文提出3D-VCD,一种用于缓解3D具身代理幻觉的视觉对比解码框架,通过构造扭曲的3D场景图来提升 grounded 推理能力,实验表明其在3D-POPE和HEAL基准上有效。

Comments 8 pages, 6 figures, Accepted at IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏