Comments21 pages, 4 figures, 5 tables. Substantially revised: title, framing and several v1 results changed. Adds a coverage sweep and a separability analysis; corrects the DPO configuration, the density-accuracy correlation and the qualitative examples. Code and data: https://huggingface.co/datasets/overthelex/citation-grounding-eval
Explaining, Verifying, and Aligning Semantic Hierarchies in Vision-Language Model Embeddings
解释、验证和对齐视觉语言模型嵌入中的语义层次结构
Gesina Schwalbe, Mert Keser, Moritz Bayerkuhnlein, Edgar Heinert, Annika Mütze, Marvin Keller, Sparsh Tiwari, Georgii Mikriukov, Diedrich Wolter, Jae Hee Lee, Matthias Rottmann
机构
*
University of Lübeck(吕贝克大学)
;
Technical University of Munich(慕尼黑工业大学)
;
AUMOVIO SE
;
Osnabrück University(奥斯纳布吕克大学)
;
Leibniz Institute for Agricultural Engineering and Bioeconomy(莱布尼茨农业工程与生物经济研究所)
;
University of Hamburg(汉堡大学)
Integrating RCTs, RWD, AI/ML and Statistics: Next-Generation Evidence Synthesis
整合RCTs、RWD、AI/ML和统计学:下一代证据合成
Shu Yang, Margaret Gamalo, Haoda Fu
机构
*
Department of Statistics, North Carolina State University(统计学系,北卡罗来纳州立大学)
;
VP and Statistics Head, Inflammation, Immunology & Specialty Care, Pfizer(副总裁及统计学负责人,炎症、免疫与专科医疗,辉瑞)
;
Head of Exploratory Biostatistics, Amgen(探索性生物统计学负责人,安进)
When Skills Meet Safety: Benchmarking and Characterizing the Adaptive Jailbreak Robustness of Skill-Merged LLMs
当技能与安全相遇:对技能合并大语言模型的自适应越狱鲁棒性进行基准测试与表征
Yu Ma, Hongli Shi, Jing Li, Xinran Xu, Weiwei Hou
机构
*
Google(谷歌公司)
;
University of New South Wales(新南威尔士大学)
;
University of Technology Sydney(悉尼科技大学)
;
Zhejiang University(浙江大学)
;
Australian National University(澳大利亚国立大学)
Tak Ho Alex Li, Kaijie Liu, Lik-Hang Lee, Kin Chung Ho, Ping Shum, Michael K. Ng
机构
*
Hong Kong Baptist University(香港浸会大学)
;
Guangdong Polytechnic Normal University(广东技术师范大学)
;
Guangdong Institute of Digital Industry(广东数字产业研究院)
;
The Hong Kong Polytechnic University(香港理工大学)
;
The Education University of Hong Kong(香港教育大学)
;
Southern University of Science and Technology(南方科技大学)
Illusion of Alignment: Detecting Hidden Disagreement in Collaborative Dialogue
对齐的错觉:检测协同对话中的隐藏分歧
Kaiming Liu, Fuwen Luo, Ziyue Wang, Jinrui Ju, Yuxuan Liu, Xuanyu Lei, Yunghwei Lai, Peng Li, Yang Liu
机构
*
College of AI, Tsinghua University(清华大学人工智能学院)
;
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)
;
Institute for AI, Tsinghua University(清华大学人工智能研究院)
AIVV: Neuro-Symbolic LLM Agent-Integrated Verification and Validation for Trustworthy Autonomous Systems
AIVV: 用于可信自主系统的神经符号LLM代理集成验证与验证
Jiyong Kwon, Ujin Jeon, Sooji Lee, Guang Lin
机构
*
School of Mechanical Engineering, Purdue University(普渡大学机械工程学院)
;
School of Electrical and Computer Engineering, Purdue University(普渡大学电气与计算机工程学院)
;
Department of Computer Science, Purdue University(普渡大学计算机科学系)
;
Department of Mathematics, Purdue University(普渡大学数学系)