arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12157 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12157 篇

2410.01929 2024-10-04 cs.AI cs.LG 86%

LLM-Augmented Symbolic Reinforcement Learning with Landmark-Based Task Decomposition

Alireza Kheirandish, Duo Xu, Faramarz Fekri

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.01270 2024-10-01 cs.CL cs.AI 86%

The African Woman is Rhythmic and Soulful: An Investigation of Implicit Biases in LLM Open-ended Text Generation

Serene Lim, María Pérez-Ortiz

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16974 2024-09-26 cs.CL cs.AI 86%

Decoding Large-Language Models: A Systematic Overview of Socio-Technical Impacts, Constraints, and Emerging Questions

Zeyneb N. Kaya, Souvick Ghosh

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.AI

Comments 28 pages, 5 figures, preprint submitted to journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.15241 2024-09-24 cs.DC cs.AI cs.LG 86%

Domino: Eliminating Communication in LLM Training via Generic Tensor Slicing and Overlapping

Guanhua Wang, Chengming Zhang, Zheyu Shen, Ang Li, Olatunji Ruwase

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments 12 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.01676 2024-09-02 cs.CL cs.AI 86%

Language models align with human judgments on key grammatical constructions

Jennifer Hu, Kyle Mahowald, Gary Lupyan, Anna Ivanova, Roger Levy

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.AI

Comments Published in PNAS at https://www.pnas.org/doi/10.1073/pnas.2400917121 as response to Dentella et al. (2023)

Journal ref Proceedings of the National Academy of Sciences, 121(36), e2400917121 (2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.07377 2024-08-16 cs.CL cs.AI cs.CY 86%

Do GPT Language Models Suffer From Split Personality Disorder? The Advent Of Substrate-Free Psychometrics

Peter Romero, Stephen Fitz, Teruo Nakatsuma

专题命中 其他LLM :language model(title,abstract);large language model(abstract);foundation model(abstract);分类 cs.CL、cs.AI

Comments 37 pages, 7 figures, 3 tables, date v1: Mar 26 2023; replaced with new version; reason: removed journal logo from older version of article that is no longer valid

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.04664 2024-08-12 cs.CL cs.AI cs.CV 86%

Mitigating Hallucinations in Large Vision-Language Models (LVLMs) via Language-Contrastive Decoding (LCD)

Avshalom Manevich, Reut Tsarfaty

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.04216 2024-08-06 cs.CL cs.LG 86%

What Do Language Models Learn in Context? The Structured Task Hypothesis

Jiaoda Li, Yifan Hou, Mrinmaya Sachan, Ryan Cotterell

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.LG

Comments This work is published in ACL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.03181 2024-07-31 cs.AI cs.CL cs.IR 86%

C-RAG: Certified Generation Risks for Retrieval-Augmented Language Models

Mintong Kang, Nezihe Merve Gürel, Ning Yu, Dawn Song, Bo Li

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.AI

Comments Accepted to ICML 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.12016 2024-07-18 cs.CL cs.AI 86%

LLM-based Frameworks for API Argument Filling in Task-Oriented Conversational Systems

Jisoo Mok, Mohammad Kachuee, Shuyang Dai, Shayan Ray, Tara Taghavi, Sungroh Yoon

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.11090 2024-07-17 cs.CL cs.AI cs.SE 86%

A Reference Architecture for Designing Foundation Model based Systems

Qinghua Lu, Liming Zhu, Xiwei Xu, Zhenchang Xing, Jon Whittle

专题命中 其他LLM :foundation model(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.03460 2024-07-08 cs.CL cs.AI 86%

Collaborative Quest Completion with LLM-driven Non-Player Characters in Minecraft

Sudha Rao, Weijia Xu, Michael Xu, Jorge Leandro, Ken Lobb, Gabriel DesGarennes, Chris Brockett, Bill Dolan

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted at Wordplay workshop at ACL 2024

Journal ref ACL 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.03045 2024-07-04 cs.HC cs.CL cs.LG 86%

JailbreakHunter: A Visual Analytics Approach for Jailbreak Prompts Discovery from Large-Scale Human-LLM Conversational Datasets

Zhihua Jin, Shiyi Liu, Haotian Li, Xun Zhao, Huamin Qu

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments 18 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.05196 2024-07-02 cs.CL cs.CY cs.HC cs.LG 86%

Does Writing with Language Models Reduce Content Diversity?

Vishakh Padmakumar, He He

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.LG

Comments ICLR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.18239 2024-06-26 cs.LG cs.CL 86%

SOUL: Unlocking the Power of Second-Order Optimization for LLM Unlearning

Jinghan Jia, Yihua Zhang, Yimeng Zhang, Jiancheng Liu, Bharat Runwal, James Diffenderfer, Bhavya Kailkhura, Sijia Liu

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.05644 2024-06-14 cs.CL cs.AI cs.CR cs.CY 86%

How Alignment and Jailbreak Work: Explain LLM Safety through Intermediate Hidden States

Zhenhong Zhou, Haiyang Yu, Xinghua Zhang, Rongwu Xu, Fei Huang, Yongbin Li

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 27 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.11911 2024-06-13 cs.CL cs.AI 86%

Blinded by Generated Contexts: How Language Models Merge Generated and Retrieved Contexts When Knowledge Conflicts?

Hexiang Tan, Fei Sun, Wanli Yang, Yuanzhuo Wang, Qi Cao, Xueqi Cheng

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.AI

Comments Accepted at ACL 2024 Main, Homepage (https://tan-hexiang.github.io/Blinded_by_Generated_Contexts/)

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.12192 2024-06-06 cs.CL cs.AI cs.CR 86%

Text Embedding Inversion Security for Multilingual Language Models

Yiyi Chen, Heather Lent, Johannes Bjerva

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.AI

Comments 18 pages, 17 Tables, 6 Figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.17027 2024-06-05 cs.CL cs.AI 86%

Player-Driven Emergence in LLM-Driven Game Narrative

Xiangyu Peng, Jessica Quaye, Sudha Rao, Weijia Xu, Portia Botchway, Chris Brockett, Nebojsa Jojic, Gabriel DesGarennes, Ken Lobb, Michael Xu, Jorge Leandro, Claire Jin, Bill Dolan

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted at IEEE Conference on Games 2024

Journal ref IEEE Conference on Games 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.16640 2024-05-20 cs.CL cs.LG 86%

TeenyTinyLlama: open-source tiny language models trained in Brazilian Portuguese

Nicholas Kluge Corrêa, Sophia Falk, Shiza Fatimah, Aniket Sen, Nythamar de Oliveira

专题命中 其他LLM :language model(title,abstract);large language model(abstract);foundation model(abstract);分类 cs.CL、cs.LG

Comments 21 pages, 5 figures

Journal ref Machine Learning With Applications, 16, 100558

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.08605 2024-05-14 cs.CL cs.AI cs.CY cs.SI 86%

Exploring the Jungle of Bias: Political Bias Attribution in Language Models via Dependency Analysis

David F. Jenny, Yann Billeter, Mrinmaya Sachan, Bernhard Schölkopf, Zhijing Jin

专题命中 其他LLM :language model(title,abstract);LLM(abstract);large language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.08309 2024-04-15 cs.CR cs.AI cs.CL 86%

Subtoxic Questions: Dive Into Attitude Change of LLM's Response in Jailbreak Attempts

Tianyu Zhang, Zixuan Zhao, Jiaqi Huang, Jingyu Hua, Sheng Zhong

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 4 pages, 2 figures. This paper was submitted to The 7th Deep Learning Security and Privacy Workshop (DLSP 2024) and was accepted as extended abstract, see https://dlsp2024.ieee-security.org/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.02466 2024-04-04 cs.CL cs.AI cs.CE 86%

Prompting for Numerical Sequences: A Case Study on Market Comment Generation

Masayuki Kawarada, Tatsuya Ishigaki, Hiroya Takamura

专题命中 其他LLM :prompting(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted to LREC-COLING2024 long paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.07838 2024-03-28 cs.CL cs.AI cs.IR 86%

LLatrieval: LLM-Verified Retrieval for Verifiable Generation

Xiaonan Li, Changtai Zhu, Linyang Li, Zhangyue Yin, Tianxiang Sun, Xipeng Qiu

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted by NAACL 2024 (Main Conference)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.16345 2024-03-26 cs.CL cs.AI cs.IR 86%

Enhanced Facet Generation with LLM Editing

Joosung Lee, Jinhong Kim

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted at LREC-COLING 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.07283 2024-03-13 cs.CR cs.CL cs.LG 86%

A Framework for Cost-Effective and Self-Adaptive LLM Shaking and Recovery Mechanism

Zhiyu Chen, Yu Li, Suochao Zhang, Jingbo Zhou, Jiwen Zhou, Chenfu Bao, Dianhai Yu

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.LG

Comments 9 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.05880 2024-02-13 cs.CL cs.AI cs.HC 86%

Generative Echo Chamber? Effects of LLM-Powered Search Systems on Diverse Information Seeking

Nikhil Sharma, Q. Vera Liao, Ziang Xiao

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Accepted in CHI'24. Supplementary material will be available online with the official submission in CHI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.14698 2024-01-31 cs.CL cs.AI 86%

Under the Surface: Tracking the Artifactuality of LLM-Generated Data

Debarati Das, Karin De Langis, Anna Martin-Boyle, Jaehyung Kim, Minhwa Lee, Zae Myung Kim, Shirley Anugrah Hayati, Risako Owan, Bin Hu, Ritik Parkar, Ryan Koo, Jonginn Park, Aahan Tyagi, Libby Ferland, Sanjali Roy, Vincent Liu, Dongyeop Kang

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments Core Authors: Debarati Das, Karin De Langis, Anna Martin-Boyle, Jaehyung Kim, Minhwa Lee and Zae Myung Kim | Project lead : Debarati Das | PI : Dongyeop Kang

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.08642 2023-12-27 cs.CL cs.AI 86%

Metacognition-Enhanced Few-Shot Prompting With Positive Reinforcement

Yu Ji, Wen Wu, Yi Hu, Hong Zheng, Liang He

专题命中 其他LLM :prompting(title,abstract);large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

Comments 5 pages, 4 figures, 2 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.10029 2023-12-19 cs.LG cs.AI 86%

Challenges with unsupervised LLM knowledge discovery

Sebastian Farquhar, Vikrant Varma, Zachary Kenton, Johannes Gasteiger, Vladimir Mikulik, Rohin Shah

专题命中 其他LLM :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments 12 pages (38 including references and appendices). First three authors equal contribution, randomised order

详情

展开后加载摘要…

URL PDF HTML 收藏