arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12169 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12169 篇

2509.21434 2026-01-12 hep-ph cs.AI cs.LG hep-ex physics.data-an 81%

Foundation models for high-energy physics

高能物理中的基础模型

Anna Hallin

机构 * Institute for Experimental Physics, University of Hamburg(汉堡大学实验物理研究所)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

AI总结 本文综述了高能物理中基础模型的应用现状,探讨了其在粒子物理数据中的潜在应用与研究进展。

Comments Submitted to SciPost Physics Proceedings (EuCAIFCon 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05630 2025-12-09 cs.CR cs.AI cs.LG 81%

How Not to Detect Prompt Injections with an LLM

如何不通过LLM检测提示注入

Sarthak Choudhary, Divyam Anshumaan, Nils Palumbo, Somesh Jha

机构 * University of Wisconsin-Madison(威斯康星大学麦迪逊分校)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI、cs.LG

AI总结 本文揭示了KAD方案的结构性漏洞,并提出DataFlip攻击方法,能够有效绕过KAD防御,实现低检测率和高恶意行为诱导率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.03442 2025-11-17 cs.CL cs.AI 81%

Are language models rational? The case of coherence norms and belief revision

Thomas Hofweber, Peter Hase, Elias Stengel-Eskin, Mohit Bansal

机构 * Department of Philosophy University of North Carolina at Chapel Hill(哲学系北卡罗来纳大学教堂山分校) Department of Computer Science University of North Carolina at Chapel Hill(计算机科学系北卡罗来纳大学教堂山分校) Department of Computer Science University of Texas at Austin(计算机科学系德克萨斯大学奥斯汀分校)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

Comments substantial expansions of sections 4 and 5, updated references, numerous smaller additions and clarifications

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08389 2025-11-12 eess.AS cs.AI cs.CL 81%

Unifying Model and Layer Fusion for Speech Foundation Models

Yi-Jen Shih, David Harwath

专题命中 其他LLM :foundation model(title,abstract);分类 cs.CL、cs.AI

Comments Accepted by IEEE ASRU 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.15758 2025-11-11 cs.LG cs.CL math.OC 81%

Distributional Surgery for Language Model Activations

Bao Nguyen, Binh Nguyen, Duy Nguyen, Viet Anh Nguyen

机构 * CUHK(香港中文大学) VinUniversity(文大学) UNC-Chapel Hill(北卡罗来纳大学教堂山分校)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.LG

Comments 3 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26697 2025-11-03 cs.CL cs.AI 81%

The End of Manual Decoding: Towards Truly End-to-End Language Models

Zhichao Wang, Dongyang Ma, Xinting Huang, Deng Cai, Tian Lan, Jiahao Xu, Haitao Mi, Xiaoying Tang, Yan Wang

机构 * Tencent AI Lab(腾讯AI实验室)

专题命中 其他LLM :language model(title);LLM(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00102 2025-10-27 cs.LG cs.AI 81%

ECG-Soup: Harnessing Multi-Layer Synergy for ECG Foundation Models

Phu X. Nguyen, Huy Phan, Hieu Pham, Christos Chatzichristos, Bert Vandenberk, Maarten De Vos

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11602 2025-10-14 cs.CL cs.LG 81%

Deconstructing Attention: Investigating Design Principles for Effective Language Modeling

Huiyin Xue, Nafise Sadat Moosavi, Nikolaos Aletras

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08632 2025-10-13 cs.CL cs.LG 81%

Next Semantic Scale Prediction via Hierarchical Diffusion Language Models

Cai Zhou, Chenyu Wang, Dinghuai Zhang, Shangyuan Tong, Yifei Wang, Stephen Bates, Tommi Jaakkola

机构 * Massachusetts Institute of Technology(麻省理工学院) Microsoft Research(微软研究院) Mila - Quebec AI Institute(魁北克AI研究所)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.LG

Comments Accepted to NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07024 2025-10-10 cs.CL cs.AI 81%

Mining the Mind: What 100M Beliefs Reveal About Frontier LLM Knowledge

Shrestha Ghosh, Luca Giordano, Yujia Hu, Tuan-Phong Nguyen, Simon Razniewski

机构 * University of Tübingen, Germany(图宾根大学) ScaDS.AI Dresden/Leipzig & TU Dresden, Germany(ScaDS.AI 德累斯顿/莱比希 & 德累斯顿技术大学) VNU University of Engineering and Technology, Hanoi, Vietnam(越南河内工程与技术大学)

专题命中 其他LLM :LLM(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.06266 2025-10-09 cs.CL cs.AI 81%

Language models for longitudinal analysis of abusive content in Billboard Music Charts

Rohitash Chandra, Yathin Suresh, Divyansh Raj Sinha, Sanchit Jindal

机构 * Transitional Artificial Intelligence Research Group, School of Mathematics and Statistics(过渡人工智能研究组、数学与统计学学院) Department of Electrical Engineering(电气工程系) Indian Institute of Technology Delhi(印度理工学院德里)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.04317 2025-10-07 cs.LG cs.AI 81%

FairAgent: Democratizing Fairness-Aware Machine Learning with LLM-Powered Agents

Yucong Dai, Lu Zhang, Feng Luo, Mashrur Chowdhury, Yongkai Wu

机构 * Clemson University USA(卡罗来纳州克莱姆森大学) University of Arkansas USA(阿肯色大学)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI、cs.LG

Comments Accepted by ICDM 2025 Demo Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03650 2025-10-07 cs.LG cs.AI cs.CE cs.NA cs.NE math.NA 81%

LLM-Guided Evolutionary Program Synthesis for Quasi-Monte Carlo Design

Amir Sadikov

机构 * Amir Sadikov(阿米尔·萨迪科夫)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13948 2025-10-07 cs.CL cs.AI 81%

Longitudinal Abuse and Sentiment Analysis of Hollywood Movie Dialogues using Language Models

Rohitash Chandra, Guoxiang Ren, Group-H

机构 * Transitional Artificial Intelligence Research Group, School of Mathematics and Statistics, UNSW Sydney(过渡人工智能研究组,数学与统计学学院) Centre for Artificial Intelligence and Innovation, Pingla Institute, Sydney, Australia(人工智能与创新中心,Pingla研究所)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21718 2025-10-06 cs.CL cs.AI 81%

Not a nuisance but a useful heuristic: Outlier dimensions favor frequent tokens in language models

Iuri Macocco, Nora Graichen, Gemma Boleda, Marco Baroni

机构 * Universitat Pompeu Fabra(庞培法布拉大学) ICREA

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

Comments Published as workshop paper at BlackBox NLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26643 2025-10-01 cs.CL cs.LG 81%

Convergence and Divergence of Language Models under Different Random Seeds

Finlay Fehlauer, Kyle Mahowald, Tiago Pimentel

机构 * ETH Zürich(苏黎世联邦理工学院) University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.LG

Comments Published at EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26224 2025-10-01 cs.CL cs.AI 81%

Type-Less yet Type-Aware Inductive Link Prediction with Pretrained Language Models

Alessandro De Bellis, Salvatore Bufi, Giovanni Servedio, Vito Walter Anelli, Tommaso Di Noia, Eugenio Di Sciascio

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

Comments Accepted and to appear in Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25268 2025-10-01 cs.LG cs.AI physics.ao-ph 81%

A Weather Foundation Model for the Power Grid

Cristian Bodnar, Raphaël Rousseau-Rizzi, Nikhil Shankar, James Merleau, Stylianos Flampouris, Guillem Candille, Slavica Antic, François Miralles, Jayesh K. Gupta

机构 * Silurian AI Hydro-Québec(水电公司)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

Comments 31 pages, 22 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.21673 2025-09-29 cs.LG cs.AI 81%

SlotFM: A Motion Foundation Model with Slot Attention for Diverse Downstream Tasks

Junyong Park, Oron Levy, Rebecca Adaimi, Asaf Liberman, Gierad Laput, Abdelkareem Bedri

机构 * Apple(苹果公司) KAIST(韩国科学技术院)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16355 2025-09-29 cs.LG cs.AI 81%

How Strategic Agents Respond: Comparing Analytical Models with LLM-Generated Responses in Strategic Classification

Tian Xie, Pavan Rauch, Xueru Zhang

机构 * The Ohio State University(俄亥俄州立大学)

专题命中 其他LLM :LLM(title,abstract);分类 cs.AI、cs.LG

Comments Add GPT 5 experiments

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.02783 2025-09-24 cs.LG cs.AI physics.geo-ph 81%

The Transparent Earth: A Multimodal Foundation Model for the Earth's Subsurface

Arnab Mazumder, Javier E. Santos, Noah Hobbs, Mohamed Mehana, Daniel O'Malley

机构 * Energy and Natural Resources Security Group(能源与自然资源安全组) Los Alamos National Laboratory(洛斯阿拉莫斯国家实验室)

专题命中 其他LLM :foundation model(title,abstract);分类 cs.AI、cs.LG

Comments Accepted at the Neurips 2025 AI4Science Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.17588 2025-09-23 cs.CV cs.AI cs.LG 81%

Interpreting Attention Heads for Image-to-Text Information Flow in Large Vision-Language Models

Jinyeong Kim, Seil Kang, Jiwoo Park, Junhyeok Kim, Seong Jae Hwang

专题命中 其他LLM :language model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16457 2025-09-23 cs.CL cs.AI cs.CY 81%

Implicit Behavioral Alignment of Language Agents in High-Stakes Crowd Simulations

Yunzhe Wang, Gale M. Lucas, Burcin Becerik-Gerber, Volkan Ustun

机构 * University of Southern California(南加州大学) USC Institute for Creative Technologies(USC创意技术研究所)

专题命中 其他LLM :language agent(title);LLM(abstract);分类 cs.CL、cs.AI

Comments Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (EMNLP 2025), Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16163 2025-09-22 cs.CV cs.AI cs.CL 81%

Robust Vision-Language Models via Tensor Decomposition: A Defense Against Adversarial Attacks

Het Patel, Muzammil Allie, Qian Zhang, Jia Chen, Evangelos E. Papalexakis

机构 * University of California, Riverside(加州大学河滨分校)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

Comments To be presented as a poster at the Workshop on Safe and Trustworthy Multimodal AI Systems (SafeMM-AI), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.14435 2025-09-16 cs.CV cs.AI cs.CY cs.LG 81%

Social Perception of Faces in a Vision-Language Model

Carina I. Hausladen, Manuel Knott, Colin F. Camerer, Pietro Perona

机构 * California Institute of Technology(加州理工学院) ETH Zurich(苏黎世联邦理工学院) Computational Social Science(计算社会科学) Swiss Data Science Center(瑞士数据科学中心) Empa(瑞士联邦材料科学与技术研究院)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI、cs.LG

Journal ref Published in the Proceedings of the 2025 ACM Conference on Fairness, Accountability, and Transparency (FAccT 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.03726 2025-09-09 cs.LG cs.AI q-bio.BM 81%

Diffusion on language model encodings for protein sequence generation

Viacheslav Meshchaninov, Pavel Strashnov, Andrey Shevtsov, Fedor Nikolaev, Nikita Ivanisenko, Olga Kardymon, Dmitry Vetrov

机构 * Constructor University, Bremen, Germany(Constructor大学,不来梅,德国)

专题命中 其他LLM :language model(title,abstract);分类 cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05140 2025-08-26 cs.CL cs.AI cs.SD eess.AS 81%

AudioLens: A Closer Look at Auditory Attribute Perception of Large Audio-Language Models

Chih-Kai Yang, Neo Ho, Yi-Jyun Lee, Hung-yi Lee

机构 * National Taiwan University(台湾大学)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

Comments Accepted to ASRU 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12863 2025-08-19 cs.CL cs.AI 81%

Word Meanings in Transformer Language Models

Jumbly Grindrod, Peter Grindrod

机构 * University of Reading, Department of Philosophy(reading大学哲学系) University of Oxford, Mathematical Institute(牛津大学数学研究所)

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.11170 2025-08-15 cs.LG cs.AI 81%

PromptTSS: A Prompting-Based Approach for Interactive Multi-Granularity Time Series Segmentation

Ching Chang, Ming-Chih Lo, Wen-Chih Peng, Tien-Fu Chen

机构 * Computer Science(计算机科学) National Yang Ming Chiao Tung University(国立阳明交通大学)

专题命中 其他LLM :prompting(title,abstract);分类 cs.AI、cs.LG

Comments Accepted at the 34th ACM International Conference on Information and Knowledge Management (CIKM 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09654 2025-08-14 cs.CL cs.LG 81%

Improving Diversity in Language Models: When Temperature Fails, Change the Loss

Alexandre Verine, Florian Le Bronnec, Kunhao Zheng, Alexandre Allauzen, Yann Chevaleyre, Benjamin Negrevergne

专题命中 其他LLM :language model(title,abstract);分类 cs.CL、cs.LG

Comments Forty-Second International Conference on Machine Learning, ICML2025

详情

展开后加载摘要…

URL PDF HTML 收藏