Mitigating Hallucination Through Theory-Consistent Symmetric Multimodal Preference Optimization
通过理论一致的对称多模态偏好优化缓解幻觉
Wenqi Liu, Xuemeng Song, Jiaxi Li, Yinwei Wei, Na Zheng, Jianhua Yin, Liqiang Nie
机构
*
Shandong University(山东大学)
;
Southern University of Science and Technology(南方科技大学)
;
University of Georgia(佐治亚大学)
;
National University of Singapore(新加坡国立大学)
;
Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))
CheXPO-v2: Preference Optimization for Chest X-ray VLMs with Knowledge Graph Consistency
CheXPO-v2:基于知识图谱一致性的胸片视觉-语言模型偏好优化
Xiao Liang, Yuxuan An, Di Wang, Jiawei Hu, Zhicheng Jiao, Bin Jing, Quan Wang
机构
*
School of Computer Science and Technology, Xidian University(西安电子科技大学计算机科学与技术学院)
;
Warren Alpert Medical School of Brown University(布朗大学沃伦·阿尔珀特医学学院)
;
School of Biomedical Engineering, Capital Medical University(首都医科大学生物医学工程学院)
机构
*
University of California, Los Angeles(加州大学洛杉矶分校)
;
Sea AI Lab(Sea AI 实验室)
;
Stanford University(斯坦福大学)
;
University of Oxford(牛津大学)
;
Yale University(耶鲁大学)
;
NTU(南洋理工大学)
;
NUS(新加坡国立大学)
;
Boston University(波士顿大学)
CommentsUpon further review, we realized that the version submitted to arXiv was not the final draft and omits crucial results and discussion. To avoid confusion and ensure the integrity of the record, we request withdrawal and will resubmit once the complete work is ready
Uncertainty Quantification for Large Language Model Reward Learning under Heterogeneous Human Feedback
大语言模型奖励学习中异质人类反馈的不确定性量化
Pangpang Liu, Junwei Lu, Will Wei Sun
机构
*
Department of Biostatistics, Yale University(耶鲁大学生物统计学系)
;
Department of Biostatistics, Harvard University(哈佛大学生物统计学系)
;
Department of Quantitative Methods, Purdue University(普渡大学定量方法系)
Aligning Diffusion Models with Noise-Conditioned Perception
对扩散模型进行噪声条件感知对齐
Alexander Gambashidze, Anton Kulikov, Yuriy Sosnin, Ilya Makarov
机构
*
Artificial Intelligence Research Institute (AIRI)(人工智能研究 institute(AIRI))
;
Skolkovo Institute of Science and Technology(斯克尔科沃科学与技术研究所)
;
HSE University(俄罗斯高等经济学院)
;
Research Center for Trusted Artificial Intelligence, ISP RAS(可信人工智能研究中心,信息与通信技术研究院)
;
ISP RAS(信息与通信技术研究院)
TinyLLM: Evaluation and Optimization of Small Language Models for Agentic Tasks on Edge Devices
TinyLLM: 小型语言模型在边缘设备上用于代理任务的评估与优化
Mohd Ariful Haque, Fahad Rahman, Kishor Datta Gupta, Khalil Shujaee, Roy George
机构
*
Department of Cyber-Physical Systems, Clark Atlanta University, USA(计算机物理系统系,Clark Atlanta大学,美国)
;
Department of Computer Science and Engineering, United International University, Bangladesh(计算机科学与工程系,联合国际大学,孟加拉国)
CommentsPaper accepted to the 2nd Workshop on Aligning Reinforcement Learning Experimentalists and Theorists (ARLET 2025) at NeurIPS; the paper consists of 14 pages (including the appendix) and contains 3 figures