arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12169 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12169 篇

2506.05739 2025-06-09 cs.CR cs.AI 79%

To Protect the LLM Agent Against the Prompt Injection Attack with Polymorphic Prompt

Zhilong Wang, Neha Nagaraja, Lan Zhang, Hayretdin Bahsi, Pawan Patil, Peng Liu

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

Comments To appear in the Industry Track of the 55th Annual IEEE/IFIP International Conference on Dependable Systems and Networks (DSN 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00725 2025-06-03 cond-mat.mtrl-sci cs.LG 79%

A Foundation Model for Non-Destructive Defect Identification from Vibrational Spectra

Mouyang Cheng, Chu-Liang Fu, Bowen Yu, Eunbi Rha, Abhijatmedhi Chotrattanapituk, Douglas L Abernathy, Yongqiang Cheng, Mingda Li

机构 * Quantum Measurement Group, MIT(麻省理工学院量子测量组) Center for Computational Science and Engineering, MIT(麻省理工学院计算科学与工程中心) Department of Materials Science and Engineering, MIT(麻省理工学院材料科学与工程系) Department of Nuclear Science and Engineering, MIT(麻省理工学院核科学与工程系) Department of Physics, MIT(麻省理工学院物理系) Department of Electrical Engineering and Computer Science, MIT(麻省理工学院电气工程与计算机科学系) Neutron Scattering Division, Oak Ridge National Laboratory(橡树岭国家实验室中子散射部)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.LG

Comments 14 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.16173 2025-06-03 cs.CL 79%

Mapping 1,000+ Language Models via the Log-Likelihood Vector

Momose Oyama, Hiroaki Yamagiwa, Yusuke Takase, Hidetoshi Shimodaira

机构 * Kyoto University(京都大学) RIKEN(理化学研究所)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.11940 2025-06-03 cs.CL 79%

The Impact of Token Granularity on the Predictive Power of Language Model Surprisal

Byung-Doh Oh, William Schuler

机构 * Center for Data Science(数据科学中心) New York University(纽约大学) Department of Linguistics(语言学系) The Ohio State University(俄亥俄州立大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments ACL 2025; results with Natural Stories alignment issue corrected (commit 4700daa)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.17820 2025-06-03 cs.CV cs.AI 79%

Don't Miss the Forest for the Trees: Attentional Vision Calibration for Large Vision Language Models

Sangmin Woo, Donguk Kim, Jaehyuk Jang, Yubin Choi, Changick Kim

机构 * KAIST(韩国科学技术院)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

Comments ACL 2025 Findings; Project: https://sangminwoo.github.io/AvisC/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.10304 2025-06-03 cs.SE cs.LG 79%

LLM-Powered Test Case Generation for Detecting Bugs in Plausible Programs

Kaibo Liu, Zhenpeng Chen, Yiyang Liu, Jie M. Zhang, Mark Harman, Yudong Han, Yun Ma, Yihong Dong, Ge Li, Gang Huang

机构 * Peking University(北京大学) Nanyang Technological University(南洋理工大学) King’s College London(伦敦国王学院) University College London(伦敦大学学院) National Key Laboratory of Data Space Technology and System(数据空间技术与系统国家重点实验室)

专题命中 其他LLM :LLM(title,abstract);分类 cs.LG

Comments Accepted by the 63rd Annual Meeting of the Association for Computational Linguistics (ACL 2025) Main Track

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.01135 2025-05-28 cs.SD cs.IR cs.LG eess.AS 79%

Music Foundation Model as Generic Booster for Music Downstream Tasks

WeiHsiang Liao, Yuhta Takida, Yukara Ikemiya, Zhi Zhong, Chieh-Hsin Lai, Giorgio Fabbro, Kazuki Shimada, Keisuke Toyama, Kinwai Cheuk, Marco A. Martínez-Ramírez, Shusuke Takahashi, Stefan Uhlich, Taketo Akama, Woosung Choi, Yuichiro Koyama, Yuki Mitsufuji

机构 * SonyAI(索尼人工智能) Sony Group Corporation(索尼集团) Sony Europe B.V.(索尼欧洲B.V.) Sony CSL(索尼计算机科学实验室)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.LG

Comments 41 pages with 14 figures

Journal ref Published in Transactions on Machine Learning Research (TMLR), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.10164 2025-05-27 cs.CL 79%

Analyzing FOMC Minutes: Accuracy and Constraints of Language Models

Wonseong Kim, Jan Frederic Spörer, Siegfried Handschuh

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments 15pages, 4 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.10681 2025-05-19 cs.CY cs.AI cs.HC 79%

Towards an LLM-powered Social Digital Twinning Platform

Önder Gürcan, Vanja Falck, Markus G. Rousseau, Larissa L. Lima

机构 * Center for Modeling Social Systems(社会科学建模中心) NORCE Norwegian Research Center AS(挪威NORCE研究机构) Kristiansand, Norway(挪威克里斯蒂安桑)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

Comments 13 pages, 3 figures, 23rd International Conference on Practical applications of Agents and Multi-Agent Systems (PAAMS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.10453 2025-05-16 cs.CV cs.AI 79%

Vision language models have difficulty recognizing virtual objects

Tyler Tran, Sangeet Khemlani, J. G. Trafton

机构 * US Naval Research Laboratory(美国海军研究实验室)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08215 2025-05-14 cs.AI cs.SD eess.AS 79%

Unveiling the Best Practices for Applying Speech Foundation Models to Speech Intelligibility Prediction for Hearing-Impaired People

Haoshuai Zhou, Boxuan Cao, Changgeng Mo, Linkai Li, Shan Xiang Wang

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06004 2025-05-12 cs.CL 79%

Exploring the Feasibility of Multilingual Grammatical Error Correction with a Single LLM up to 9B parameters: A Comparative Study of 17 Models

Dawid Wisniewski, Antoni Solarski, Artur Nowakowski

机构 * Poznan University of Technology(波兹南技术大学) Adam Mickiewicz University(亚当·密茨凯维奇大学)

专题命中 其他LLM :LLM(title);language model(abstract);分类 cs.CL

Comments Accepted at MTSummit 2025 (The 20th Machine Translation Summit)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05970 2025-05-12 cs.CL 79%

Towards Developmentally Plausible Rewards: Communicative Success as a Learning Signal for Interactive Language Models

Lennart Stöpler, Rufat Asadli, Mitja Nikolaus, Ryan Cotterell, Alex Warstadt

机构 * CerCo, CNRS(CerCo与法国国家科学研究中心) University of California San Diego(加州大学圣地亚哥分校)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.01900 2025-05-06 cs.CL 79%

CAMOUFLAGE: Exploiting Misinformation Detection Systems Through LLM-driven Adversarial Claim Transformation

Mazal Bethany, Nishant Vishwamitra, Cho-Yu Jason Chiang, Peyman Najafirad

机构 * University of Texas at San Antonio(德克萨斯大学圣安东尼奥分校)

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08985 2025-04-29 cs.HC cs.AI 79%

Learning from Elders: Making an LLM-powered Chatbot for Retirement Communities more Accessible through User-centered Design

Luna Xingyu Li, Ray-yuan Chung, Feng Chen, Wenyu Zeng, Yein Jeon, Oleg Zaslavsky

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI

Comments Accepted as Research talk for Considering Cultural and Linguistic Diversity in AI Applications workshop at CALD-AI@ASIS&T 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.20268 2025-04-29 cs.LG 79%

Centaur: a foundation model of human cognition

Marcel Binz, Elif Akata, Matthias Bethge, Franziska Brändle, Fred Callaway, Julian Coda-Forno, Peter Dayan, Can Demircan, Maria K. Eckstein, Noémi Éltető, Thomas L. Griffiths, Susanne Haridi, Akshay K. Jagadish, Li Ji-An, Alexander Kipnis, Sreejan Kumar, Tobias Ludwig, Marvin Mathony, Marcelo Mattar, Alireza Modirshanechi, Surabhi S. Nath, Joshua C. Peterson, Milena Rmus, Evan M. Russek, Tankred Saanum, Johannes A. Schubert, Luca M. Schulze Buschoff, Nishad Singhi, Xin Sui, Mirko Thalmann, Fabian Theis, Vuong Truong, Vishaal Udandarao, Konstantinos Voudouris, Robert Wilson, Kristin Witte, Shuchen Wu, Dirk Wulff, Huadong Xiong, Eric Schulz

机构 * Helmholtz Munich(慕尼黑海德堡医学研究院) University of Tuebingen(图宾根大学) University of Oxford(牛津大学) New York University(纽约大学) Max Planck Institute for Biological Cybernetics(生物控制研究所) Google DeepMind(谷歌DeepMind) Princeton University(普林斯顿大学) University of California San Diego(圣地亚哥大学) Boston University(波士顿大学) Georgia Institute of Technology(佐治亚理工学院) University of Basel(巴塞尔大学) Max Planck Institute for Human Development(人类发展研究所) Max Planck School of Cognition(认知研究所) TU Darmstadt(德累斯顿技术大学) University of Cambridge(剑桥大学)

专题命中 其他LLM :foundation model(title);language model(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07149 2025-04-29 cs.CV cs.LG 79%

Towards Interpreting Visual Information Processing in Vision-Language Models

Clement Neo, Luke Ong, Philip Torr, Mor Geva, David Krueger, Fazl Barez

机构 * Nanyang Technological University(南洋理工大学) University of Oxford(牛津大学) Tel Aviv University(特拉维夫大学) MILA(蒙特利尔人工智能研究院) ERA-Krueger AI Safety Lab(ERA-Krueger人工智能安全实验室)

专题命中 其他LLM :language model(title,abstract);分类 cs.LG

Comments Published at ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.17083 2025-04-25 cs.CL 79%

How Individual Traits and Language Styles Shape Preferences In Open-ended User-LLM Interaction: A Preliminary Study

Rendi Chevi, Kentaro Inui, Thamar Solorio, Alham Fikri Aji

机构 * MBZUAI Abu Dhabi UAE(阿布扎赫 MBZUAI)

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL

Comments Accepted at GenAICHI 2025 @ ACM CHI 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.14335 2025-04-22 cs.CV cs.AI 79%

Visual Prompting for One-shot Controllable Video Editing without Inversion

Zhengbo Zhang, Yuxi Zhou, Duo Peng, Joo-Hwee Lim, Zhigang Tu, De Wen Soh, Lin Geng Foo

机构 * Singapore University of Technology and Design(新加坡科技设计大学) Wuhan University(武汉大学) Institute for Infocomm Research, Agency for Science, Technology and Research, Singapore(新加坡资讯与通信研究院)

专题命中 其他LLM :prompting(title,abstract);分类 cs.AI

Comments accepted by cvpr2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.12767 2025-04-18 cs.CL 79%

Out of Sight Out of Mind, Out of Sight Out of Mind: Measuring Bias in Language Models Against Overlooked Marginalized Groups in Regional Contexts

Fatma Elsafoury, David Hartmann

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10391 2025-04-15 cs.CL 79%

LLM-driven Constrained Copy Generation through Iterative Refinement

Varun Vasudevan, Faezeh Akhavizadegan, Abhinav Prakash, Yokila Arora, Jason Cho, Tanya Mendiratta, Sushant Kumar, Kannan Achan

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL

Comments 10 pages, 2 figures, 7 Tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.04917 2025-04-15 cs.CV cs.AI 79%

Avoid Wasted Annotation Costs in Open-set Active Learning with Pre-trained Vision-Language Model

Jaehyuk Heo, Pilsung Kang

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19510 2025-03-26 cs.RO cs.AI cs.CV 79%

RoboFlamingo-Plus: Fusion of Depth and RGB Perception with Vision-Language Models for Enhanced Robotic Manipulation

Sheng Wang

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18556 2025-03-25 cs.CV cs.CL 79%

Instruction-Aligned Visual Attention for Mitigating Hallucinations in Large Vision-Language Models

Bin Li, Dehong Gao, Yeyuan Wang, Linbo Jin, Shanqing Yu, Xiaoyan Cai, Libin Yang

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments Accepted by ICME2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.07304 2025-03-25 cs.CL 79%

We're Calling an Intervention: Exploring Fundamental Hurdles in Adapting Language Models to Nonstandard Text

Aarohi Srivastava, David Chiang

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments Accepted for publication at W-NUT 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.10419 2025-03-14 cs.RO cs.AI 79%

HiFi-CS: Towards Open Vocabulary Visual Grounding For Robotic Grasping Using Vision-Language Models

Vineet Bhat, Prashanth Krishnamurthy, Ramesh Karri, Farshad Khorrami

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06137 2025-03-11 cs.CL 79%

Evaluating Discourse Cohesion in Pre-trained Language Models

Jie He, Wanqiu Long, Deyi Xiong

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.04996 2025-03-10 cs.CL 79%

HieroLM: Egyptian Hieroglyph Recovery with Next Word Prediction Language Model

Xuheng Cai, Erica Zhang

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

Comments Accepted at LaTeCH-CLfL 2025 @ NAACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.03474 2025-03-06 cs.CL 79%

Enhancing Spoken Discourse Modeling in Language Models Using Gestural Cues

Varsha Suresh, M. Hamza Mughal, Christian Theobalt, Vera Demberg

专题命中 其他LLM :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.16007 2025-03-06 cs.AI 79%

Language Model Probabilities are Not Calibrated in Numeric Contexts

Charles Lovering, Michael Krumdick, Viet Dac Lai, Seth Ebner, Nilesh Kumar, Varshini Reddy, Rik Koncel-Kedziorski, Chris Tanner

专题命中 其他LLM :language model(title,abstract);分类 cs.AI

Comments 8 pages (main), 39 pages (references and appendix), in submission

详情

展开后加载摘要…

URL PDF HTML 收藏