arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型对齐与安全

大模型对齐、安全、越狱、红队、提示注入和可信评测。

共收录 1738 信号源:cs.CL, cs.AI, cs.CY, cs.LG

1. 幻觉与事实性 1738 篇

2502.11649 2025-09-03 cs.AI cs.SI 57%

Competing LLM Agents in a Non-Cooperative Game of Opinion Polarisation

Amin Qasmi, Usman Naseem, Mehwish Nasim

机构 * The University of Western Australia(西澳大学) Lahore University of Management Sciences(拉合尔管理科学大学) Macquarie University(麦考瑞大学)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12067 2025-09-01 eess.AS cs.AI cs.SD 57%

Evaluating Logit-Based GOP Scores for Mispronunciation Detection

Aditya Kamlesh Parikh, Cristian Tejedor-Garcia, Catia Cucchiarini, Helmer Strik

机构 * Centre for Language Studies(语言研究学院)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.AI

Comments Accepted to Interspeech 2025. This publication is part of the project Responsible AI for Voice Diagnostics (RAIVD) with file number NGF.1607.22.013 of the research programme NGF AiNed Fellowship Grants which is financed by the Dutch Research Council (NWO)

Journal ref https://www.isca-archive.org/interspeech_2025/parikh25b_interspeech.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.19432 2025-08-28 cs.AI 57%

Quantized but Deceptive? A Multi-Dimensional Truthfulness Evaluation of Quantized LLMs

Yao Fu, Xianxuan Long, Runchao Li, Haotian Yu, Mu Sheng, Xiaotian Han, Yu Yin, Pan Li

机构 * Case Western Reserve University(凯斯西储大学) Hangzhou Dianzi University(杭州电子科技大学)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.AI

Comments Accepted to EMNLP2025 main conference (poster)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12964 2025-08-26 cs.CL 57%

Trust Me, I'm Wrong: LLMs Hallucinate with Certainty Despite Knowing the Answer

Adi Simhi, Itay Itzhak, Fazl Barez, Gabriel Stanovsky, Yonatan Belinkov

机构 * Technion – Israel Institute of Technology(技术ion-以色列理工学院) University of Oxford and WhiteBox(牛津大学和WhiteBox) School of Computer Science and Engineering, The Hebrew University of Jerusalem(耶路撒冷希伯来大学计算机科学与工程学院)

专题命中 幻觉与事实性 :safety(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.18001 2025-08-26 cs.LG stat.ML 57%

A Novel Framework for Uncertainty Quantification via Proper Scores for Classification and Beyond

Sebastian G. Gruber

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.LG

Comments PhD Thesis (cumulative, spanning 6 peer-reviewed publications)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01225 2025-08-25 cs.CV cs.AI 57%

Multi-Cache Enhanced Prototype Learning for Test-Time Generalization of Vision-Language Models

Xinyu Chen, Haotian Zhai, Can Zhang, Xiupeng Shi, Ruirui Li

机构 * Shanghai University(上海大学) Beijing University of Chemical Technology(北京化工大学) University of Minnesota(明尼苏达大学)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.AI

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.14266 2025-08-21 cs.CV cs.AI 57%

Effect of Data Augmentation on Conformal Prediction for Diabetic Retinopathy

Rizwan Ahamed, Annahita Amireskandari, Joel Palko, Carol Laxson, Binod Bhattarai, Prashnna Gyawali

机构 * West Virginia University(西弗吉尼亚大学) University of Aberdeen(阿伯丁大学)

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.AI

Comments 3rd Workshop in Data Engineering in Medical Imaging (DEMI), MICCAI-2025 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11257 2025-08-18 cs.SE cs.AI 57%

Hallucination in LLM-Based Code Generation: An Automotive Case Study

Marc Pavel, Nenad Petrovic, Lukasz Mazur, Vahid Zolfaghari, Fengjunjie Pan, Alois Knoll

机构 * Real-Time Systems Technical University of Munich(实时系统技术大学慕尼黑)

专题命中 幻觉与事实性 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10010 2025-08-15 cs.CL 57%

An Audit and Analysis of LLM-Assisted Health Misinformation Jailbreaks Against LLMs

Ayana Hussain, Patrick Zhao, Nicholas Vincent

专题命中 幻觉与事实性 :jailbreak(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09458 2025-08-15 cs.HC cs.AI cs.ET 57%

Hallucination vs interpretation: rethinking accuracy and precision in AI-assisted data extraction for knowledge synthesis

Xi Long, Christy Boscardin, Lauren A. Maggio, Joseph A. Costello, Ralph Gonzales, Rasmyah Hammoudeh, Ki Lai, Yoon Soo Park, Brian C. Gin

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09561 2025-08-14 cs.LG 57%

Edge General Intelligence Through World Models and Agentic AI: Fundamentals, Solutions, and Challenges

Changyuan Zhao, Guangyuan Liu, Ruichen Zhang, Yinqiu Liu, Jiacheng Wang, Jiawen Kang, Dusit Niyato, Zan Li, Xuemin, Shen, Zhu Han, Sumei Sun, Chau Yuen, Dong In Kim

机构 * College of Computing and Data Science, Nanyang Technological University(南洋理工大学计算与数据科学学院) School of Automation, Guangdong University of Technology(广东工业大学自动化学院) State Key Laboratory of Integrated Services Networks, Xidian University(西安电子科技大学集成服务网络国家重点实验室) Department of Electrical and Computer Engineering, University of Waterloo(滑铁卢大学电气与计算机工程系) Department of Computer Science and Engineering, Kyung Hee University(韩国庆熙大学计算机科学与工程系) Institute for Infocomm Research, Agency for Science, Technology and Research(科技研究局信息通信研究所) Department of Electrical and Computer Engineering, Sungkyunkwan University(庆熙大学电气与计算机工程系)

专题命中 幻觉与事实性 :safety(abstract);分类 cs.LG

Comments 21 pages. 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.06795 2025-08-13 cs.CL cs.CV 57%

From Pixels to Tokens: Revisiting Object Hallucinations in Large Vision-Language Models

Yuying Shang, Xinyi Zeng, Yutao Zhu, Xiao Yang, Zhengwei Fang, Jingyuan Zhang, Jiawei Chen, Zinan Liu, Yu Tian

机构 * University of Chinese Academy of Sciences(中国科学院大学) Dept. of Comp. Sci. and Tech., Institute for AI, Tsinghua University(计算机科学与技术系,人工智能研究院,清华大学) Gaoling School of Artificial Intelligence, Renmin University of China(人工智能学院,中国人民大学) Kuaishou Technology Inc.(快手科技有限公司) Shanghai Key Laboratory of Multi. Info. Processing, East China Normal University(多信息处理重点实验室,华东师范大学)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07223 2025-08-12 cs.IR cs.AI 57%

Selection and Exploitation of High-Quality Knowledge from Large Language Models for Recommendation

Guanchen Wang, Mingming Ha, Tianbao Ma, Linxun Chen, Zhaojie Liu, Guorui Zhou, Kun Gai

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04105 2025-08-07 cs.AI 57%

Towards Transparent AI Grading: Semantic Entropy as a Signal for Human-AI Disagreement

Karrtik Iyer, Manikandan Ravikiran, Prasanna Pendse, Shayan Mohanty

机构 * Thoughtworks AI Research Labs(Thoughtworks人工智能研究实验室)

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04005 2025-08-07 cs.LG 57%

Decoupled Contrastive Learning for Federated Learning

Hyungbin Kim, Incheol Baek, Yon Dohn Chung

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.12591 2025-08-06 cs.CV cs.CL 57%

CutPaste&Find: Efficient Multimodal Hallucination Detector with Visual-aid Knowledge Base

Cong-Duy Nguyen, Xiaobao Wu, Duc Anh Vu, Shuai Zhao, Thong Nguyen, Anh Tuan Luu

机构 * Nanyang Technological University, Singapore(南洋理工大学) National University of Singapore, Singapore(国立新加坡大学)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02419 2025-08-05 cs.CV cs.CL 57%

Modality Bias in LVLMs: Analyzing and Mitigating Object Hallucination via Attention Lens

Haohan Zheng, Zhenguo Zhang

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01678 2025-08-05 cs.CV cs.AI 57%

Cure or Poison? Embedding Instructions Visually Alters Hallucination in Vision-Language Models

Zhaochen Wang, Yiwei Wang, Yujun Cai

机构 * The University of Queensland(昆士兰大学) University of California, Merced(加州大学默塞德分校)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.AI

Comments Work in progress

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23736 2025-08-01 stat.ML cs.LG 57%

DICOM De-Identification via Hybrid AI and Rule-Based Framework for Scalable, Uncertainty-Aware Redaction

Kyle Naddeo, Nikolas Koutsoubis, Rahul Krish, Ghulam Rasool, Nidhal Bouaynaya, Tony OSullivan, Raj Krish

机构 * Rowan University(罗文大学) Moffitt Cancer Center(莫菲特癌症中心) University of South Florida(佛罗里达州立大学) Impact Business Information Solutions, Inc(Impact商务信息解决方案公司)

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.LG

Comments 15 pages, 6 figures,

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21406 2025-07-30 cs.AI 57%

Shapley Uncertainty in Natural Language Generation

Meilin Zhu, Gaojie Jin, Xiaowei Huang, Lijun Zhang

机构 * University of Exeter(埃克塞特大学) University of Liverpool(利物浦大学) ISCAS(国际信息科学协会)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16771 2025-07-23 cs.LG stat.AP stat.ML 57%

A Partitioned Sparse Variational Gaussian Process for Fast, Distributed Spatial Modeling

Michael Grosskopf, Kellin Rumsey, Ayan Biswas, Earl Lawrence

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19073 2025-07-22 cs.CL 57%

Towards Harmonized Uncertainty Estimation for Large Language Models

Rui Li, Jing Long, Muge Qi, Heming Xia, Lei Sha, Peiyi Wang, Zhifang Sui

机构 * Peking University(北京大学) The Hong Kong Polytechnic University(香港理工大学) Beihang University(北航)

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.CL

Comments ACL 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14782 2025-07-22 stat.ML cs.LG math-ph math.MP stat.CO 57%

Uncertainty Quantification for Machine Learning-Based Prediction: A Polynomial Chaos Expansion Approach for Joint Model and Input Uncertainty Propagation

Xiaoping Du

机构 * School of Mechanical Engineering(机械工程学院) Purdue University(普渡大学)

专题命中 幻觉与事实性 :trustworthy(abstract);分类 cs.LG

Comments This manuscript has been submitted to Multidisciplinary and Structural Optimization

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.14540 2025-07-18 cs.SE cs.AI cs.CR 57%

Risks of ignoring uncertainty propagation in AI-augmented security pipelines

Emanuele Mezzi, Aurora Papotti, Fabio Massacci, Katja Tuma

机构 * Vrije Universiteit Amsterdam(阿姆斯特丹自由大学) University of Trento(特伦托大学) Eindhoven University of Technology(埃因霍温理工大学)

专题命中 幻觉与事实性 :safety(abstract);分类 cs.AI

Comments Accepted for publication in Risk Analysis: An International Journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23146 2025-07-15 cs.CL 57%

Learning-to-Context Slope: Evaluating In-Context Learning Effectiveness Beyond Performance Illusions

Dingzriui Wang, Xuanliang Zhang, Keyan Xu, Qingfu Zhu, Wanxiang Che, Yang Deng

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06571 2025-07-10 cs.CL 57%

Enhancing Food-Domain Question Answering with a Multimodal Knowledge Graph: Hybrid QA Generation and Diversity Analysis

Srihari K B, Pushpak Bhattacharyya

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.12182 2025-07-09 cs.CL 57%

Truth Neurons

Haohang Li, Yupeng Cao, Yangyang Yu, Jordan W. Suchow, Zining Zhu

机构 * Stevens Institute of Technology(史蒂文斯理工学院)

专题命中 幻觉与事实性 :safety(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05163 2025-07-08 cs.CV cs.LG 57%

Probabilistic Embeddings for Frozen Vision-Language Models: Uncertainty Quantification with Gaussian Process Latent Variable Models

Aishwarya Venkataramanan, Paul Bodesheim, Joachim Denzler

机构 * Computer Vision Group, Friedrich Schiller University Jena(计算机视觉组,费迪里奇·施勒尔大学耶纳)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.LG

Comments UAI 2025, 22 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10435 2025-07-04 cs.CV cs.AI 57%

COEF-VQ: Cost-Efficient Video Quality Understanding through a Cascaded Multimodal LLM Framework

Xin Dong, Sen Jia, Ming Rui Wang, Yan Li, Zhenheng Yang, Bingfeng Deng, Hongyu Xiong

机构 * ByteDance Inc.(字节跳动公司)

专题命中 幻觉与事实性 :safety(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.17815 2025-07-03 cs.AI 57%

DREAMS: A python framework for Training Deep Learning Models on EEG Data with Model Card Reporting for Medical Applications

Rabindra Khadka, Pedro G Lind, Anis Yazidi, Asma Belhadi

机构 * Department of Computer Science, OsloMet -- Oslo Metropolitan University(计算机科学系,奥斯陆Met大学) Simula Research Laboratory, Numerical Analysis and Scientific Computing(Simula研究实验室,数值分析与科学计算)

专题命中 幻觉与事实性 :alignment(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏