arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 7552 信号源:cs.CL, cs.AI, cs.LG

1. 知识编辑与模型理解 7552 篇

2506.19498 2025-06-25 cs.RO cs.AI 79%

T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models

Yiteng Chen, Wenbo Li, Shiyi Wang, Huiping Zhuang, Qingyao Wu

机构 * School of Software Engineering, South China University of Technology(软件工程学院,华南理工大学) School of Future Technology, South China University of Technology(未来技术学院,华南理工大学) Shien-Ming Wu School of Intelligent Engineering, South China University of Technology(智能工程学院,华南理工大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI

Comments submitted to NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03036 2025-06-13 cs.CL 79%

IPA-CHILDES & G2P+: Feature-Rich Resources for Cross-Lingual Phonology and Phonemic Language Modeling

Zébulon Goriely, Paula Buttery

机构 * Department of Computer Science & Technology, University of Cambridge, U.K.(计算机科学与技术系,剑桥大学) ALTA Institute, University of Cambridge, U.K.(ALTA研究所,剑桥大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Accepted to CoNLL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09593 2025-06-12 cs.LG 79%

Beyond Overconfidence: Foundation Models Redefine Calibration in Deep Neural Networks

Achim Hekler, Lukas Kuhn, Florian Buettner

机构 * Goethe University Frankfurt(弗赖堡歌德大学) German Cancer Consortium (DKTK)(德国癌症联盟(DKTK)) German Cancer Research Center (DKFZ)(德国癌症研究中心(DKFZ)) Frankfurt Cancer Institute(法兰克福癌症研究所)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06686 2025-06-10 cs.CL 79%

Learning Distribution-Wise Control in Representation Space for Language Models

Chunyuan Deng, Ruidi Chang, Hanjie Chen

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05136 2025-06-06 cs.CL 79%

Information Locality as an Inductive Bias for Neural Language Models

Taiga Someya, Anej Svete, Brian DuSell, Timothy J. O'Donnell, Mario Giulianelli, Ryan Cotterell

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12414 2025-06-06 cs.CL 79%

Lost in Transcription, Found in Distribution Shift: Demystifying Hallucination in Speech Foundation Models

Hanin Atwany, Abdul Waheed, Rita Singh, Monojit Choudhury, Bhiksha Raj

机构 * Carnegie Mellon University(卡内基梅隆大学) MBZUAI(穆扎布伊人工智能研究所)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.CL

Comments ACL2025 camera-ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03741 2025-06-05 cs.HC cs.CL 79%

PromptCanvas: Composable Prompting Workspaces Using Dynamic Widgets for Exploration and Iteration in Creative Writing

Rifat Mehreen Amin, Oliver Hans Kühle, Daniel Buschek, Andreas Butz

机构 * LMU Munich(慕尼黑莱布尼茨大学) University of Bayreuth(拜罗伊特大学)

专题命中 知识编辑与模型理解 :prompting(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24731 2025-06-02 cs.CL 79%

Circuit Stability Characterizes Language Model Generalization

Alan Sun

机构 * Carnegie Mellon University(卡内基梅隆大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments 16 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24649 2025-06-02 cs.CV cs.AI 79%

BIMA: Bijective Maximum Likelihood Learning Approach to Hallucination Prediction and Mitigation in Large Vision-Language Models

Huu-Thien Tran, Thanh-Dat Truong, Khoa Luu

机构 * CVIU Lab, University of Arkansas(CVIU实验室,阿肯色大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI

Comments CVPRW 2025, 8 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23556 2025-05-30 cs.CL 79%

Understanding Refusal in Language Models with Sparse Autoencoders

Wei Jie Yeo, Nirmalendu Prakash, Clement Neo, Roy Ka-Wei Lee, Erik Cambria, Ranjan Satapathy

机构 * Nanyang Technological University(南洋理工大学) Singapore University of Technology and Design(新加坡科技设计大学) Digital Trust Centre(数字信任中心) Institute of High Performance Computing (IHPC)(高性能计算研究所)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12821 2025-05-30 cs.CV cs.AI 79%

From Head to Tail: Towards Balanced Representation in Large Vision-Language Models through Adaptive Data Calibration

Mingyang Song, Xiaoye Qu, Jiawei Zhou, Yu Cheng

机构 * Fudan University(复旦大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Stony Brook University(石溪大学) The Chinese University of Hong Kong(香港中文大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI

Comments Accepted by CVPR 2025. Project Page: https://vlmlt.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21547 2025-05-29 cs.CV cs.AI 79%

Image Tokens Matter: Mitigating Hallucination in Discrete Tokenizer-based Large Vision-Language Models via Latent Editing

Weixing Wang, Zifeng Ding, Jindong Gu, Rui Cao, Christoph Meinel, Gerard de Melo, Haojin Yang

机构 * Hasso Plattner Institute(霍普夫纳研究所) University of Potsdam(波茨坦大学) University of Cambridge(剑桥大学) University of Oxford(牛津大学) German University of Digital Science(德国数字科学大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.09090 2025-05-27 cs.CL 79%

Social Bias Probing: Fairness Benchmarking for Language Models

Marta Marchiori Manerba, Karolina Stańczak, Riccardo Guidotti, Isabelle Augenstein

机构 * University of Pisa(比萨大学) University of Copenhagen(哥本哈根大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Journal ref EMNLP 2024: Proceedings of the 2024 Conference on Empirical Methods in Natural Language Processing, 2024, 14653-14671

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18575 2025-05-27 cs.AI 79%

Response Uncertainty and Probe Modeling: Two Sides of the Same Coin in LLM Interpretability?

Yongjie Wang, Yibo Wang, Xin Zhou, Zhiqi Shen

机构 * Nanyang Technological University(南洋理工大学)

专题命中 知识编辑与模型理解 :LLM(title,abstract);分类 cs.AI

Comments 18 Pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18154 2025-05-26 cs.CL cs.CY 79%

The Staircase of Ethics: Probing LLM Value Priorities through Multi-Step Induction to Complex Moral Dilemmas

Ya Wu, Qiang Sheng, Danding Wang, Guang Yang, Yifan Sun, Zhengjia Wang, Yuyan Bu, Juan Cao

专题命中 知识编辑与模型理解 :LLM(title,abstract);分类 cs.CL

Comments 25 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15682 2025-05-22 cs.CL 79%

The Representational Alignment between Humans and Language Models is implicitly driven by a Concreteness Effect

Cosimo Iaia, Bhavin Choksi, Emily Wiebers, Gemma Roig, Christian J. Fiebach

机构 * Goethe University Frankfurt(弗赖堡歌德大学) Center for Brains, Minds and Machines, MIT Hessian.AI(大脑、心智与机器中心,MIT 荷尔斯泰因人工智能) Brain Imaging Center(脑成像中心)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments 13 pages, 4 Figures, 1 Table

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.11516 2025-05-16 cs.CL 79%

TopoLM: brain-like spatio-functional organization in a topographic language model

Neil Rathi, Johannes Mehrer, Badr AlKhamissi, Taha Binhuraib, Nicholas M. Blauch, Martin Schrimpf

机构 * EPFL(瑞士联邦理工学院) Stanford University(斯坦福大学) Georgia Institute of Technology(佐治亚理工学院) Harvard University(哈佛大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02056 2025-05-06 cs.CV cs.LG 79%

Handling Imbalanced Pseudolabels for Vision-Language Models with Concept Alignment and Confusion-Aware Calibrated Margin

Yuchen Wang, Xuefeng Bai, Xiucheng Li, Weili Guan, Liqiang Nie, Xinyang Chen

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.LG

Comments Accepted to ICML 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.20839 2025-04-30 cs.CL quant-ph 79%

Universal language model with the intervention of quantum theory

D. -F. Qin

机构 * School of Physics and Electronic Science, East China Normal University(物理与电子科学学院,华东师范大学)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.18027 2025-04-28 cs.CV cs.AI cs.RO 79%

A Large Vision-Language Model based Environment Perception System for Visually Impaired People

Zezhou Chen, Zhaoxiang Liu, Kai Wang, Kohou Wang, Shiguo Lian

机构 * AI Innovation Center, China Unicom(中国unicom人工智能创新中心) Unicom Digital Technology, China Unicom(中国unicom数字技术)

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.AI

Comments Accepted by IROS2024(9 pages, 8 figures)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.16948 2025-04-25 cs.CY cs.AI cs.ET 79%

Intrinsic Barriers to Explaining Deep Foundation Models

Zhen Tan, Huan Liu

机构 * Arizona State University(亚利桑那州立大学)

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.07983 2025-04-17 cs.CV cs.LG 79%

Two Effects, One Trigger: On the Modality Gap, Object Bias, and Information Imbalance in Contrastive Vision-Language Models

Simon Schrodi, David T. Hoffmann, Max Argus, Volker Fischer, Thomas Brox

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.LG

Comments ICLR 2025 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10422 2025-04-15 cs.LG 79%

Foundation models for electronic health records: representation dynamics and transferability

Michael C. Burkhart, Bashar Ramadan, Zewei Liao, Kaveri Chhikara, Juan C. Rojas, William F. Parker, Brett K. Beaulieu-Jones

专题命中 知识编辑与模型理解 :foundation model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10049 2025-04-15 cs.CV cs.CL 79%

Summarization of Multimodal Presentations with Vision-Language Models: Study of the Effect of Modalities and Structure

Théo Gigant, Camille Guinaudeau, Frédéric Dufaux

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08001 2025-04-14 cs.CL 79%

Linguistic Interpretability of Transformer-based Language Models: a systematic review

Miguel López-Otal, Jorge Gracia, Jordi Bernad, Carlos Bobed, Lucía Pitarch-Ballesteros, Emma Anglés-Herrero

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

Comments Supplementary material: https://github.com/sid-unizar/ling-int-survey/blob/main/table.pdf

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06575 2025-04-11 cs.CR cs.CL 79%

Defending LLM Watermarking Against Spoofing Attacks with Contrastive Representation Learning

Li An, Yujian Liu, Yepeng Liu, Yang Zhang, Yuheng Bu, Shiyu Chang

专题命中 知识编辑与模型理解 :LLM(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04264 2025-04-08 cs.CL 79%

Lost in Multilinguality: Dissecting Cross-lingual Factual Inconsistency in Transformer Language Models

Mingyang Wang, Heike Adel, Lukas Lange, Yihong Liu, Ercong Nie, Jannik Strötgen, Hinrich Schütze

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02862 2025-04-08 cs.CV cs.LG 79%

Towards Understanding How Knowledge Evolves in Large Vision-Language Models

Sudong Wang, Yunjian Zhang, Yao Zhu, Jianing Li, Zizhe Wang, Yanwei Liu, Xiangyang Ji

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.LG

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20796 2025-03-28 cs.CR cs.AI 79%

EXPLICATE: Enhancing Phishing Detection through Explainable AI and LLM-Powered Interpretability

Bryan Lim, Roman Huerta, Alejandro Sotelo, Anthonie Quintela, Priyanka Kumar

专题命中 知识编辑与模型理解 :LLM(title,abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08497 2025-03-27 cs.LG cs.CV 79%

MMRL: Multi-Modal Representation Learning for Vision-Language Models

Yuncheng Guo, Xiaodong Gu

专题命中 知识编辑与模型理解 :language model(title,abstract);分类 cs.LG

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏