Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization
机构 * University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学)
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.CL、cs.AI
Comments BMVC 2025
AI 大模型
大模型对齐、安全、越狱、红队、提示注入和可信评测。
机构 * University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学)
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.CL、cs.AI
Comments BMVC 2025
机构 * Wharton AI & Analytics Initiative(沃顿人工智能与分析倡议) ; University of Pennsylvania(宾夕法尼亚大学) ; Department of Computer and Information Science(计算机与信息科学系) ; Department of Physics and Astronomy(物理学与天文学系)
专题命中 偏好对齐 :alignment(abstract);RLHF(abstract);分类 cs.AI、cs.LG
机构 * University at Buffalo(布法罗大学) ; Adobe Research(Adobe研究) ; Pennsylvania State University(宾夕法尼亚州立大学)
专题命中 偏好对齐 :alignment(abstract);safety(abstract);分类 cs.CL、cs.AI
Comments Accepted at ICCV 2025. Code available at https://github.com/sjz5202/LLaVA-Reward
机构 * Korea University(韩国大学)
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.CL、cs.AI
机构 * Stanford University(斯坦福大学)
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.CL、cs.AI
机构 * Imperial College London(伦敦帝国理工学院)
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.CL、cs.LG
机构 * Tongyi Lab, Alibaba Group(通义实验室,阿里巴巴集团)
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.CL、cs.AI
Comments ACL 2025
机构 * Massachusetts Institute of Technology(麻省理工学院)
专题命中 偏好对齐 :RLHF(abstract);safety(abstract);分类 cs.CL、cs.AI
Comments 2025 ICML Efficient Systems for Foundation Models Workshop
机构 * Department of Electrical and Computer Engineering, University of Central Florida(中央佛罗里达大学电气与计算机工程系) ; School of Electrical Engineering and Computer Science, Oregon State University(俄勒冈州立大学电气工程与计算机科学学院) ; Department of Electrical and Computer Engineering, University of Minnesota(明尼苏达大学电气与计算机工程系)
专题命中 偏好对齐 :RLHF(abstract);DPO(abstract);分类 cs.AI、cs.LG
机构 * University of Zurich(苏黎世大学) ; EPFL(瑞士联邦理工学院) ; Zurich University of Applied Sciences(苏黎世应用科学大学)
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.CL、cs.AI
机构 * Department of Computer Science and Engineering, The Ohio State University(计算机科学与工程系,俄亥俄州立大学) ; Department of Computer Science(计算机科学系) ; Engineering, The Ohio State University(工程系,俄亥俄州立大学)
专题命中 偏好对齐 :alignment(abstract);RLHF(abstract);分类 cs.AI、cs.LG
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.AI、cs.LG
Comments 10pages
机构 * Department of Electronic Engineering, BRNist, Tsinghua University, Beijing, China(电子工程系,BRNist,清华大学,北京,中国)
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.CL、cs.AI
机构 * Wuhan University(武汉大学) ; Nanyang Technological University(南洋理工大学)
专题命中 偏好对齐 :RLHF(abstract);DPO(abstract);分类 cs.AI、cs.LG
Comments Under review
机构 * Rice University(里士大学) ; University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) ; University of Cambridge(剑桥大学) ; Columbia University(哥伦比亚大学)
专题命中 偏好对齐 :alignment(abstract);safety(abstract);分类 cs.CL、cs.AI
Comments 14 pages
专题命中 偏好对齐 :alignment(abstract);RLHF(abstract);分类 cs.AI、cs.LG
Comments Revision: paper accepted by the ICML2025 main conference
机构 * Peking University(北京大学)
专题命中 偏好对齐 :alignment(abstract);RLHF(abstract);分类 cs.AI、cs.LG
Comments Revision: The paper was accepted by Transactions of Machine Learning Research (TMLR)
机构 * Department of Computer Science, University of Pisa(比萨大学计算机科学系)
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.CL、cs.AI
Comments Accepted at Findings of ACL 2025
机构 * Key Laboratory of Computational Intelligence and Chinese Information Processing of Ministry of Education, School of Computer and Information Technology, Shanxi University(教育部计算智能与中文信息处理重点实验室,计算机与信息学院,山西大学) ; Institute for AI, Peking University(人工智能研究院,北京大学) ; University College London(伦敦大学学院)
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.AI、cs.LG
机构 * University of Science and Technology of China(中国科学技术大学)
专题命中 偏好对齐 :RLHF(abstract);DPO(abstract);分类 cs.CL、cs.LG
机构 * University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校) ; Nanyang Technological University(南洋理工大学) ; National University of Singapore(国立新加坡大学) ; Microsoft Research(微软研究院)
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.CL、cs.LG
Comments To appear at ICLR 2025
专题命中 偏好对齐 :alignment(abstract);RLHF(abstract);分类 cs.CL、cs.AI
专题命中 偏好对齐 :RLHF(abstract);DPO(abstract);分类 cs.AI、cs.LG
专题命中 偏好对齐 :RLHF(abstract);DPO(abstract);分类 cs.AI、cs.LG
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.AI、cs.LG
Comments ICLR 2025; Project Page available at : https://sprain02.github.io/FiFA/
专题命中 偏好对齐 :alignment(abstract);RLHF(abstract);分类 cs.AI、cs.LG
Comments 25 pages, 11 figures
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.CL、cs.LG
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.CL、cs.LG
Comments ICLR 2025
专题命中 偏好对齐 :alignment(abstract);DPO(abstract);分类 cs.CL、cs.AI
专题命中 偏好对齐 :alignment(abstract);RLHF(abstract);分类 cs.CL、cs.AI
Comments To appear in the proceedings of the first Workshop on the Scaling Behavior of Large Language Models (EACL 2024)
Journal ref https://aclanthology.org/2024.scalellm-1.5/