arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7550 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7550 篇

2407.09434 2024-11-13 cs.LG cs.AI cs.CE cs.SY eess.SY 81%

Foundation Models for the Electric Power Grid

Hendrik F. Hamann, Thomas Brunschwiler, Blazhe Gjorgiev, Leonardo S. A. Martins, Alban Puech, Anna Varbella, Jonas Weiss, Juan Bernabe-Moreno, Alexandre Blondin Massé, Seong Choi, Ian Foster, Bri-Mathias Hodge, Rishabh Jain, Kibaek Kim, Vincent Mai, François Mirallès, Martin De Montigny, Octavio Ramos-Leaños, Hussein Suprême, Le Xie, El-Nasser S. Youssef, Arnaud Zinflou, Alexander J. Belyi, Ricardo J. Bessa, Bishnu Prasad Bhattarai, Johannes Schmude, Stanislav Sobolevsky

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.AI、cs.LG

Comments Major equal contributors: H.F.H., T.B., B.G., L.S.A.M., A.P., A.V., J.W.; Significant equal contributors: J.B., A.B.M., S.C., I.F., B.H., R.J., K.K., V.M., F.M., M.D.M., O.R., H.S., L.X., E.S.Y., A.Z.; Other equal contributors: A.J.B., R.J.B., B.P.B., J.S., S.S; Lead contact: H.F.H

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.11944 2024-11-08 cs.LG cs.CL 81%

Transcoders Find Interpretable LLM Feature Circuits

Jacob Dunefsky, Philippe Chlenski, Neel Nanda

专题命中 知识编辑与模型理解 :LLM(title);language model(abstract);分类 cs.CL、cs.LG

Comments 29 pages, 6 figures, 4 tables, 2 algorithms. NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.06192 2024-11-04 cs.CV cs.AI cs.CL 81%

Multi-Object Hallucination in Vision-Language Models

Xuweiyi Chen, Ziqiao Ma, Xuejun Zhang, Sihan Xu, Shengyi Qian, Jianing Yang, David F. Fouhey, Joyce Chai

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

Comments Accepted to NeurIPS 2024 | Project page: https://multi-object-hallucination.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.00123 2024-10-31 cs.CL cs.LG 81%

Comparing Template-based and Template-free Language Model Probing

Sagi Shaier, Kevin Bennett, Lawrence E Hunter, Katharina von der Wense

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.LG

Comments Accepted to EACL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.18819 2024-10-25 cs.CL cs.CY cs.LG 81%

From Imitation to Introspection: Probing Self-Consciousness in Language Models

Sirui Chen, Shu Yu, Shengjie Zhao, Chaochao Lu

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14746 2024-10-22 cs.CL cs.AI cs.HC 81%

Accounting for Sycophancy in Language Model Uncertainty Estimation

Anthony Sicilia, Mert Inan, Malihe Alikhani

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.11665 2024-10-16 cs.CV cs.AI cs.CL 81%

VisualRWKV-HD and UHD: Advancing High-Resolution Processing for Visual Language Models

Zihang Li, Haowen Hou

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.10455 2024-10-15 cs.IR cs.CL cs.LG 81%

Advancing Academic Knowledge Retrieval via LLM-enhanced Representation Similarity Fusion

Wei Dai, Peng Fu, Chunjing Gan

专题命中 知识编辑与模型理解 :LLM(title,abstract);分类 cs.CL、cs.LG

Comments The 2nd Place of KDD Cup 2024 OAG-Challenge AQA

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.09762 2024-09-02 cs.CL cs.AI 81%

Diversifying the Mixture-of-Experts Representation for Language Models with Orthogonal Optimizer

Boan Liu, Liang Ding, Li Shen, Keqin Peng, Yu Cao, Dazhao Cheng, Dacheng Tao

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

Comments ECAI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17700 2024-08-28 cs.CL cs.LG 81%

RAVEL: Evaluating Interpretability Methods on Disentangling Language Model Representations

Jing Huang, Zhengxuan Wu, Christopher Potts, Mor Geva, Atticus Geiger

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.LG

Comments Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (ACL 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.00569 2024-08-06 cs.CV cs.AI cs.CL 81%

Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models

Weihong Zhong, Xiaocheng Feng, Liang Zhao, Qiming Li, Lei Huang, Yuxuan Gu, Weitao Ma, Yuan Xu, Bing Qin

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

Comments Accepted to ACL 2024 Main Conference. 21 pages, 20 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.08176 2024-07-12 cs.SE cs.AI cs.LG 81%

Foundation Model Engineering: Engineering Foundation Models Just as Engineering Software

Dezhi Ran, Mengzhou Wu, Wei Yang, Tao Xie

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.AI、cs.LG

Comments Accepted by 2030 Software Engineering Workshop, co-located with FSE24; Invited to ACM TOSEM 2030 Roadmap for Software Engineering

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18765 2024-06-28 cs.LG cs.AI cs.CV 81%

WV-Net: A foundation model for SAR WV-mode satellite imagery trained using contrastive self-supervised learning on 10 million images

Yannik Glaser, Justin E. Stopa, Linnea M. Wolniewicz, Ralph Foster, Doug Vandemark, Alexis Mouche, Bertrand Chapron, Peter Sadowski

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.AI、cs.LG

Comments 20 pages, 9 figures, submitted to NeurIPS 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.12824 2024-06-19 cs.CL cs.AI 81%

From RAGs to rich parameters: Probing how language models utilize external knowledge over parametric information for factual queries

Hitesh Wadhwa, Rahul Seetharaman, Somyaa Aggarwal, Reshmi Ghosh, Samyadeep Basu, Soundararajan Srinivasan, Wenlong Zhao, Shreyas Chaudhari, Ehsan Aghazadeh

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18167 2024-06-19 cs.CL cs.AI 81%

Mechanistic Understanding and Mitigation of Language Model Non-Factual Hallucinations

Lei Yu, Meng Cao, Jackie Chi Kit Cheung, Yue Dong

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.08648 2024-06-18 cs.CL cs.AI 81%

Explore Spurious Correlations at the Concept Level in Language Models for Text Classification

Yuhang Zhou, Paiheng Xu, Xiaoyu Liu, Bang An, Wei Ai, Furong Huang

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

Comments 14 pages, 4 page appendix, Accepted by ACL 2024 Main

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06441 2024-06-11 cs.CL cs.AI 81%

Interpretability of Language Models via Task Spaces

Lucas Weber, Jaap Jumelet, Elia Bruni, Dieuwke Hupkes

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

Comments To be published at ACL 2024 (main)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.14381 2024-05-14 cs.CL cs.AI 81%

Editing Knowledge Representation of Language Model via Rephrased Prefix Prompts

Yuchen Cai, Ding Cao, Rongxi Guo, Yaqin Wen, Guiquan Liu, Enhong Chen

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

Comments 19pages,3figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.00253 2024-05-07 cs.CV cs.CL cs.LG 81%

A Survey on Hallucination in Large Vision-Language Models

Hanchao Liu, Wenyuan Xue, Yifei Chen, Dapeng Chen, Xiutian Zhao, Ke Wang, Liping Hou, Rongjun Li, Wei Peng

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.10967 2024-05-02 cs.CL cs.AI cs.IR 81%

Knowledge Graphs and Pre-trained Language Models enhanced Representation Learning for Conversational Recommender Systems

Zhangchi Qiu, Ye Tao, Shirui Pan, Alan Wee-Chung Liew

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

Comments Accepted by IEEE Transactions on Neural Networks and Learning Systems (TNNLS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.14906 2024-04-24 cs.CV cs.AI cs.LG 81%

Driver Activity Classification Using Generalizable Representations from Vision-Language Models

Ross Greer, Mathias Viborg Andersen, Andreas Møgelmose, Mohan Trivedi

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.12043 2024-04-16 cs.CV cs.CL cs.LG 81%

When are Lemons Purple? The Concept Association Bias of Vision-Language Models

Yutaro Yamada, Yingtian Tang, Yoyo Zhang, Ilker Yildirim

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.LG

Comments EMNLP 2023 main

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.07204 2024-04-11 cs.CV cs.AI cs.LG 81%

BRAVE: Broadening the visual encoding of vision-language models

Oğuzhan Fatih Kar, Alessio Tonioni, Petra Poklukar, Achin Kulshrestha, Amir Zamir, Federico Tombari

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI、cs.LG

Comments Project page at https://brave-vlms.epfl.ch/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.04538 2024-04-09 cs.AI cs.CL 81%

Soft-Prompting with Graph-of-Thought for Multi-modal Representation Learning

Juncheng Yang, Zuchao Li, Shuai Xie, Wei Yu, Shijun Li, Bo Du

专题命中 知识编辑与模型理解 :prompting(title,abstract);分类 cs.CL、cs.AI

Comments This paper is accepted to LREC-COLING 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.15875 2024-03-26 cs.AI cs.CL 81%

LAMPER: LanguAge Model and Prompt EngineeRing for zero-shot time series classification

Zhicheng Du, Zhaotian Xie, Yan Tong, Peiwu Qin

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

Comments Accepted as tiny paper in ICLR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.18248 2024-03-21 cs.CL cs.AI 81%

Do Language Models Know When They're Hallucinating References?

Ayush Agrawal, Mirac Suzgun, Lester Mackey, Adam Tauman Kalai

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17527 2024-03-19 cs.CL cs.AI 81%

Predict the Next Word: Humans exhibit uncertainty in this task and language models _____

Evgenia Ilia, Wilker Aziz

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

Comments 22 pages, EACL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.00754 2024-03-19 cs.LG cs.CL cs.CV 81%

Analyzing and Mitigating Object Hallucination in Large Vision-Language Models

Yiyang Zhou, Chenhang Cui, Jaehong Yoon, Linjun Zhang, Zhun Deng, Chelsea Finn, Mohit Bansal, Huaxiu Yao

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.LG

Comments Accepted by ICLR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.09714 2024-03-18 cs.CL cs.AI 81%

Linguistic Structure Induction from Language Models

Omar Momen

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

Comments Master's Thesis. Supervised by Laura Kallmeyer and David Arps

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.05189 2024-03-11 cs.CL cs.AI 81%

Tracing the Roots of Facts in Multilingual Language Models: Independent, Shared, and Transferred Knowledge

Xin Zhao, Naoki Yoshinaga, Daisuke Oba

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL、cs.AI

Comments EACL 2024 main conference

详情

展开后加载摘要…

URL PDF HTML 收藏