机构
*
Institute of Big Data Science and Industry(大数据科学与产业研究院)
;
Key Laboratory of Evolutionary Science Intelligence of Shanxi Province(山西省进化智能科学重点实验室)
;
School of Artificial Intelligence, Shanxi University(山西大学人工智能学院)
专题命中
幻觉与鲁棒性
:multimodal large language model(abstract);分类 cs.CV、cs.LG
Finding the Correct Visual Evidence Without Forgetting: Mitigating Hallucination in LVLMs via Inter-Layer Visual Attention Discrepancy
在不遗忘的情况下寻找正确的视觉证据:通过层间视觉注意力差异减轻LVLMs中的幻觉
Yutong Xie, Zhenglin Hua, Ran Wang, Wing W. Y. Ng, Xizhao Wang, Yuheng Jia
机构
*
School of Computer Science and Engineering, Southeast University, Nanjing, China(东南大学计算机科学与工程学院)
;
School of Artificial Intelligence, Shenzhen University, Shenzhen, China(深圳大学人工智能学院)
;
College of Computer Science and Software Engineering, Shenzhen University, Shenzhen, China(深圳大学计算机科学与软件工程学院)
;
Engineering, South China University of Technology, Guangzhou, China(华南理工大学工程学院)
;
National Engineering Laboratory for Big Data Systems Computing Technology, Shenzhen University, Shenzhen, China(深圳大学大数据系统计算技术国家工程实验室)
;
Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University), Ministry of Education, China(新一代人工智能技术及其交叉应用重点实验室(东南大学),中华人民共和国教育部)
机构
*
State Key Lab. of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(中国科学院人工智能安全国家重点实验室,计算技术研究所)
;
School of Advanced Interdisciplinary Sciences, University of Chinese Academy of Sciences(中国科学院大学先进交叉科学学院)
;
School of Computer Science and Technology, University of Chinese Academy of Sciences(中国科学院大学计算机科学与技术学院)
Jailbreaking on Text-to-Video Models via Scene Splitting Strategy
通过场景分割策略对文本到视频模型进行劫持
Wonjun Lee, Haon Park, Doehyeon Lee, Bumsub Ham, Suhyun Kim
机构
*
Yonsei University(延世大学)
;
Korea Institute of Science and Technology(韩国科学技术院)
;
AIM Intelligence(AIM智能)
;
Seoul National University(首尔国立大学)
;
Kyung Hee University(庆熙大学)
When and Why Adversarial Training Improves PINNs: A Neural Tangent Kernel Perspective
何时以及为何对抗训练能提升PINNs:神经 tangent 核视角
Yuan-dong Cao, Chi Chiu SO, Jun-Min Wang, He Wang
机构
*
School of Mathematics and Statistics, Beijing Institute of Technology, China(北京理工大学数学与统计学院,中国)
;
School of Professional Education and Executive Development The Hong Kong Polytechnic University, China(香港理工大学专业教育学院及管理发展学院,中国)
;
Department of Computer Science & UCL AI Centre, University College London, UK(伦敦大学学院计算机科学系及UCL人工智能中心,英国)
Anchoring the Eigengap: Cross-Modal Spectral Stabilization for Sample-Efficient Representation Learning
锚定特征间隙:跨模态谱稳定化以实现样本高效表征学习
Nikhil J. Dhinagar, Vidhi Chhatbar, Chirag Jagad, Pavithra Senthilkumar, Sophia I. Thomopoulos, Mahir H. Khan, Sook-Lei Liew, the ENIGMA-Stroke Recovery Working Group, Paul M. Thompson
机构
*
Imaging Genetics Center, Mark & Mary Stevens Neuroimaging & Informatics Institute, Keck School of Medicine, University of Southern California(影像基因中心,马克与玛丽史蒂文斯神经影像与信息学研究所,凯克医学院,南加州大学)
;
Neuroscience Graduate Program, Mark & Mary Stevens Neuroimaging & Informatics Institute, Chan Division of Occupational Science & Occupational Therapy, Biomedical Engineering, University of Southern California(神经科学研究生项目,马克与玛丽史蒂文斯神经影像与信息学研究所,查恩职业科学与职业治疗 division,生物医学工程,南加州大学)
Retrieval-Guided Generation for Safer Histopathology Image Captioning
基于检索的生成用于更安全的病理科图像描述生成
Md. Enamul Hoq, Wataru Uegami, Saghir Alfasly, Ghazal Alabtah, Sahar Rahimi Malakshan, Armita Kazemi, Alex T. Schmitgen, Fred Prior, H. R. Tizhoosh
机构
*
Kimia Lab, Department of Artificial Intelligence \& Informatics, Mayo Clinic, Rochester, MN, USA
;
Department of Biomedical Informatics, University of Arkansas for Medical Sciences, Little Rock, AR, USA
;
Lane Department of Computer Science
;
Electrical Engineering, West Virginia University, Morgantown, WV, USA
;
Department of Computer Science
;
Engineering, Princeton University, Princeton, NJ, USA
;
Department of Computer Sciences, University of Wisconsin--Madison, Madison, WI, USA
Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models
穿越不确定性:音频感知大语言模型不确定性估计的实证研究
Chun-Yi Kuan, Wei-Ping Huang, Hung-yi Lee
机构
*
Graduate Institute of Communication Engineering, National Taiwan University, Taiwan(台湾大学通讯工程研究所)
;
Artificial Intelligence Center of Research Excellence (AI-CoRE), National Taiwan University, Taiwan(台湾大学人工智能卓越研究中心)
Suhas BN, Andrew M. Sherrill, Rosa I. Arriaga, Chris W. Wiese, Saeed Abdullah
机构
*
College of Information Sciences and Technology, Penn State University(宾夕法尼亚州立大学信息科学与技术学院)
;
Department of Psychiatry and Behavioral Sciences, Emory University(埃默里大学精神病学与行为科学系)
;
School of Interactive Computing, Georgia Tech(佐治亚理工学院交互计算学院)
;
School of Psychology, Georgia Tech(佐治亚理工学院心理学学院)
From Handwriting to Structured Data: Benchmarking AI Digitisation of Handwritten Forms
从手写到结构化数据:基准测试AI手写文档数字化
Nicholas Pather, Joshua Fouché, Sitwala Mundia, Karl-Günter Technau, Thokozile Malaba, Alex Welte, Ushma Mehta, Bruce A. Bassett
机构
*
CSAM, University of the Witwatersrand(沃特沙尔德大学计算机科学与数学系)
;
Grai Labs(Grai实验室)
;
Faculty of Health Sciences, University of the Witwatersrand(沃特沙尔德大学健康科学学院)
;
Empilweni Services and Research Unit, Department of Paediatrics and Child Health, University of the Witwatersrand(沃特沙尔德大学儿科与儿童健康部门Empilweni服务与研究单位)
;
Division of Epidemiology and Biostatistics, School of Public Health, Faculty of Health Sciences, University of Cape Town(开普敦大学公共卫生学院流行病学与生物统计学系)
;
Discipline of Public Health, School of Medicine, University of KwaZulu-Natal(夸祖鲁-纳塔尔大学医学院公共卫生系)
;
WITS MIND Institute and CSAM, University of the Witwatersrand, University of Cape Town & Grai Labs(沃特沙尔德大学WITS MIND研究所和CSAM、开普敦大学及Grai实验室)
专题命中
幻觉与鲁棒性
:multimodal large language model(abstract);分类 cs.CV、cs.LG
机构
*
Department of Electrical and Computer Engineering, Northeastern University(东北大学电气与计算机工程系)
;
Khoury College of Computer Science, Northeastern University(东北大学科赫里计算机科学学院)
Safety at Scale: A Comprehensive Survey of Large Model and Agent Safety
大规模安全:大型模型和智能体安全的全面综述
Xingjun Ma, Yifeng Gao, Yixu Wang, Ruofan Wang, Xin Wang, Ye Sun, Yifan Ding, Hengyuan Xu, Yunhao Chen, Yunhan Zhao, Hanxun Huang, Yige Li, Yutao Wu, Jiaming Zhang, Xiang Zheng, Yang Bai, Zuxuan Wu, Xipeng Qiu, Jingfeng Zhang, Yiming Li, Xudong Han, Haonan Li, Jun Sun, Cong Wang, Jindong Gu, Baoyuan Wu, Siheng Chen, Tianwei Zhang, Yang Liu, Mingming Gong, Tongliang Liu, Shirui Pan, Cihang Xie, Tianyu Pang, Yinpeng Dong, Ruoxi Jia, Yang Zhang, Shiqing Ma, Xiangyu Zhang, Neil Gong, Chaowei Xiao, Sarah Erfani, Tim Baldwin, Bo Li, Masashi Sugiyama, Dacheng Tao, James Bailey, Yu-Gang Jiang
机构
*
Fudan University(复旦大学)
;
The University of Melbourne(墨尔本大学)
;
Singapore Management University(新加坡国立大学)
;
Deakin University(德肯大学)
;
Hong Kong University of Science and Technology(香港科学与技术大学)
;
City University of Hong Kong(香港城市大学)
;
University of Oxford(牛津大学)
;
Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Shanghai Jiao Tong University(上海交通大学)
;
Nanyang Technological University(南洋理工大学)
;
The University of Sydney(悉尼大学)
;
Griffith University(格里菲斯大学)
;
University of California, Santa Cruz(加州大学圣克鲁兹分校)
;
Sea AI Lab(Sea AI实验室)
;
Tsinghua University(清华大学)
;
Virginia Tech(弗吉尼亚理工大学)
;
CISPA Helmholtz Center for Information Security(CISPA海德堡信息安全部)
;
University of Massachusetts Amherst(马萨诸塞大学阿姆赫斯特分校)
;
Purdue University(普渡大学)
;
Duke University(杜克大学)
;
University of Wisconsin - Madison(威斯康星大学麦迪逊分校)
;
RIKEN(理化学研究所)
;
The University of Tokyo(东京大学)
GIRL: Generative Imagination Reinforcement Learning via Information-Theoretic Hallucination Control
GIRL:通过信息论幻觉控制实现生成性想象强化学习
Prakul Sunil Hiremath
机构
*
Department of Computer Science and Engineering, Visvesvaraya Technological University (VTU), Belagavi, India(维斯瓦拉亚科技大学计算机科学与工程系,贝拉加维,印度)
;
Aliens on Earth (AoE) Autonomous Research Group, Belagavi, India(地球外星人自主研究组,贝拉加维,印度)
Comments20 pages, 6 figures, 6 tables. Introduces a 15k-sample representation-level hallucination dataset with full transformer hidden states and multi-signal weak supervision. Evaluates 5 probing architectures and demonstrates internal hallucination detection without external inference-time signals. Includes held-out test evaluation and deployment benchmarks