机构
*
University of Science and Technology of China(中国科学技术大学)
;
Ningbo Institute of Digital Twin(宁波数字孪生研究所)
;
Eastern Institute of Technology(东部技术研究所)
;
Department of Computing, The Hong Kong Polytechnic University(香港理工大学计算机系)
Zero-Shot Parkinson's Disease Detection from Speech: Comparing Large Audio and Language Models
零样本帕金森病语音检测:比较大型音频和语言模型
Muhammad Ashad Kabir, Sirajam Munira
机构
*
School of Computing, Mathematics and Engineering, Charles Sturt University(计算机科学与工程学院,查尔斯·斯图尔特大学)
;
Department of Computer Science, Rensselaer Polytechnic Institute(计算机科学系,伦塞拉尔理工学院)
机构
*
X-LANCE Lab(X-LANCE实验室)
;
School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学学院)
;
MoE Key Lab of Artificial Intelligence(人工智能MOE重点实验室)
;
Jiangsu Key Lab of Language Computing(江苏省语言计算重点实验室)
;
Beijing Key Laboratory of Applied Experimental Psychology(北京应用实验心理学重点实验室)
;
National Demonstration Center for Experimental Psychology Education, Faculty of Psychology, Beijing Normal University(北京师范大学实验心理学教育国家级示范中心,心理学学院)
ORACAL: A Robust and Explainable Multimodal Framework for Smart Contract Vulnerability Detection with Causal Graph Enrichment
ORACAL: 一种基于因果图增强的鲁棒且可解释的智能合约漏洞检测多模态框架
Tran Duong Minh Dai, Triet Huynh Minh Le, M. Ali Babar, Van-Hau Pham, Phan The Duy
机构
*
Information Security Lab, University of Information Technology(信息安全部,信息科技大学)
;
Vietnam National University(越南国家大学)
;
School of Computer Science and Information Technology, Adelaide University(计算机科学与信息技术学院,阿德莱德大学)
Unifying Speech Editing Detection and Content Localization via Prior-Enhanced Audio LLMs
通过先验增强的音频大语言模型统一语音编辑检测与内容定位
Jun Xue, Yi Chai, Yanzhen Ren, Jinshen He, Zhiqiang Tang, Zhuolin Yi, Yihuan Huang, Yuankun Xie, Yujie Chen
机构
*
Key Laboratory of Aerospace Information Security(航空信息安全与可信计算重点实验室)
;
School of Cyber Science and Engineering(网络安全工程学院)
;
Wuhan University(武汉大学)
;
Independent Researcher(独立研究员)
;
School of Computer Science and Technology(计算机科学与技术学院)
;
Anhui University(安徽大学)
;
Communication University of China(中国通信大学)
;
Beihang University(北京航空航天大学)
Teaching large language models to reason like expert diagnosticians
教会大型语言模型像专家诊断医生一样推理
Thomas A. Buckley, Riccardo Conci, Peter G. Brodeur, Jason Gusdorf, Sourik Beltrán, Bita Behrouzi, Byron Crowe, Jacob Dockterman, Muzzammil Muhammad, Sarah Ohnigian, Andrew Sanchez, James A. Diao, Aashna P. Shah, Daniel Restrepo, Eric S. Rosenberg, Andrew S. Lea, Emily Glanton, Kimberly LeBlanc, Undiagnosed Diseases Network, Marinka Zitnik, Scott H. Podolsky, Zahir Kanjee, Raja-Elie E. Abdulnour, Jacob M. Koshy, Adam Rodman, Arjun K. Manrai
机构
*
Department of Biomedical Informatics, Harvard Medical School(哈佛医学院生物医学信息学系)
;
Department of Medicine, Beth Israel Deaconess Medical Center(贝塞斯达医院内科部)
;
The Mongan Institute, Massachusetts General Hospital(麻省总医院蒙根研究所)
;
Division of Gastroenterology, Brigham and Women’s Hospital(布里洛妇女医院胃肠病科)
;
Department of Medicine, Brigham and Women’s Hospital(布里洛妇女医院内科部)
;
Department of Medicine, Massachusetts General Hospital(麻省总医院内科部)
;
Department of Pathology, Massachusetts General Hospital(麻省总医院病理学部)
;
Department of Health Humanities and Bioethics, University of Rochester School of Medicine and Dentistry(罗切斯特大学医学院和牙科学院健康人文与生物伦理学部)
;
Kempner Institute for the Study of Natural and Artificial Intelligence, Harvard University(哈佛大学凯普纳人工智能研究所)
;
Center for the History of Medicine, Countway Library of Medicine, Harvard Medical School(哈佛医学院医学史中心,考特维图书馆)
;
Department of Global Health and Social Medicine, Harvard Medical School(哈佛医学院全球健康与社会医学部)
;
Division of Pulmonary and Critical Care Medicine, Brigham and Women’s Hospital(布里洛妇女医院呼吸科和重症医学科)
专题命中
推理评测
:reasoning(abstract);分类 cs.AI
AI总结
提出 Dr. CaBot 代理 AI 系统,通过生成基于初始病例描述的幻灯片演示来模拟专家诊断推理,并在 NEJM CPC 和 NIH 未诊断疾病网络病例上取得优于前沿模型的表现,同时发布 CPC-Bench 基准以促进临床 AI 发展。
机构
*
The Ohio State University(俄亥俄州立大学)
;
The University of Chicago(芝加哥大学)
;
University College London(伦敦大学学院)
;
University of Michigan(密歇根大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Case Western Reserve University(凯斯西储大学)
;
Amazon(亚马逊)
Faithful or Fabricated? A Causal Framework for Rationalization Bias in LLM Judges
忠实还是捏造?LLM 评判中合理化偏差的因果框架
Riya Tapwal, Abhishek Kumar, Carsten Maple
机构
*
School of Computing and Electrical Engineering(计算与电子工程学院)
;
Indian Institute of Technology (IIT) Mandi(印度理工学院(IIT)曼迪)
;
London U.K.(伦敦英国)
;
Warwick Manufacturing Group U.K.(沃里克制造集团英国)
Multi-Persona Debate System for Automated Scientific Hypothesis Generation
用于自动科学假设生成的多角色辩论系统
Jaeha Oh, Byungchan Kim, Ju Li, Yang Jeong Park, Jin-Sung Park
机构
*
Department of Materials Science & Engineering, Ajou University(材料科学与工程系,阿乔大学)
;
Department of Energy Systems Research, Ajou University(能源系统研究系,阿乔大学)
;
Department of Nuclear Science and Engineering, Massachusetts Institute of Technology(核科学与工程系,麻省理工学院)
;
Department of Materials Science and Engineering, Massachusetts Institute of Technology(材料科学与工程系,麻省理工学院)
;
Department of Materials Science and Engineering, Ulsan National Institute of Science and Technology(材料科学与工程系,乌山国家科学与技术研究院)
;
Graduate School of Artificial Intelligence, Ulsan National Institute of Science and Technology(人工智能研究生院,乌山国家科学与技术研究院)
Is GPT-4o mini Blinded by its Own Safety Filters? Exposing the Multimodal-to-Unimodal Bottleneck in Hate Speech Detection
GPT-4o mini 是否被自身的安全过滤器蒙蔽?揭示多模态到单模态瓶颈在仇恨言论检测中的作用
Niruthiha Selvanayagam, Ted Kurti
专题命中
推理评测
:reasoning(abstract);分类 cs.LG
AI总结
本文通过 Hateful Memes Challenge 数据集系统分析 GPT-4o mini 在多模态仇恨言论检测中的安全架构,发现并实验验证了“单模态瓶颈”缺陷,即上下文无关的安全过滤器会优先阻断多模态推理,导致误报。
CommentsThis paper reports preliminary findings from a small-scale study whose sample size is insufficient to support the stated conclusions. The authors are withdrawing it to conduct a more comprehensive evaluation
机构
*
Tencent Youtu Lab(腾讯优图实验室)
;
Tsinghua University(清华大学)
;
The University of Hong Kong(香港大学)
;
University of Warwick(沃林汉大学)
;
Monash University(墨尔本大学)
;
The Hong Kong Polytechnic University(香港理工大学)