机构
*
Department of Computer Science, Stanford University, Stanford, CA, USA(计算机科学系,斯坦福大学,斯坦福,加州,美国)
;
Department of Psychiatry and Behavioral Sciences, Stanford University, Stanford, CA, USA(精神病学与行为科学系,斯坦福大学,斯坦福,加州,美国)
;
Department of Biomedical Data Science, Stanford University, Stanford, CA, USA(生物医学数据科学系,斯坦福大学,斯坦福,加州,美国)
;
University of California, Berkeley, Berkeley, CA, USA(加州大学伯克利分校,伯克利,加州,美国)
;
The University of Hong Kong, Hong Kong(香港大学,香港)
UrbanAlign: Post-hoc Semantic Calibration for VLM-Human Preference Alignment
UrbanAlign: 域内任务中VLM与人类偏好对齐的后处理语义校准
Yecheng Zhang, Rong Zhao, Zhizhou Sha, Yong Li, Lei Wang, Ce Hou, Wen Ji, Hao Huang, Yunshan Wan, Jian Yu, Junhao Xia, Yuru Zhang, Chunlei Shi
机构
*
Tsinghua University(清华大学)
;
University College London(伦敦大学学院)
;
Hong Kong University of Science and Technology(香港科学与技术大学)
;
Peking University(北京大学)
;
Southwest Jiaotong University(西南交通大学)
;
Zhejiang University(浙江大学)
;
University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
Renmin University of China(中国人民大学)
;
Southeast University(东南大学)
Alignment through Meta-Weighted Online Sampling: Bridging the Gap between Data Generation and Preference Optimization
通过元权重在线采样对齐:弥合数据生成与偏好优化之间的差距
Junming Yang, Ning Xu, Biao Liu, Shiqi Qiao, Xin Geng
机构
*
School of Computer Science and Engineering, Southeast University, Nanjing, China(东南大学计算机科学与工程学院)
;
Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University), Ministry of Education, China(新一代人工智能技术及其交叉应用国家重点实验室)
Less is More: Improving LLM Alignment via Preference Data Selection
少即是多:通过偏好数据选择改进大语言模型对齐
Xun Deng, Han Zhong, Rui Ai, Fuli Feng, Zheng Wang, Xiangnan He
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Peking University(北京大学)
;
Massachusetts Institute of Technology(麻省理工学院)
;
Alibaba Cloud Computing(阿里云计算)
;
MoE Key Lab of BIPC, University of Science and Technology of China(中国科学技术大学MoE关键实验室)
MM-SCALE: Grounded Multimodal Moral Reasoning via Scalar Judgment and Listwise Alignment
MM-SCALE: 基于标量判断和列表对齐的 grounded 多模态道德推理
Eunkyu Park, Wesley Hanwen Deng, Cheyon Jin, Matheus Kunzler Maldaner, Jordan Wheeler, Jason I. Hong, Hong Shen, Adam Perer, Ken Holstein, Motahhare Eslami, Gunhee Kim
机构
*
Seoul National University(首尔国立大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
University of Florida(佛罗里达大学)
;
Epic Games
MARE: Multimodal Alignment and Reinforcement for Explainable Deepfake Detection via Vision-Language Models
MARE: 多模态对齐与强化学习用于通过视觉-语言模型的可解释深度伪造检测
Wenbo Xu, Wei Lu, Xiangyang Luo, Jiantao Zhou
机构
*
School of Computer Science and Engineering, MoE Key Laboratory of Information Technology, Guangdong Province Key Laboratory of Information Security Technology, Sun Yat-sen University, Guangzhou 510006, China(计算机科学与工程学院,信息技术MOE实验室,广东省信息安全技术重点实验室,中山大学,广州510006,中国)
;
State Key Laboratory of Mathematical Engineering and Advanced Computing(数学工程与先进计算国家重点实验室)
;
Department of Computer and Information Science, University of Macau.(计算机与信息科学系,澳门大学)
机构
*
University of Waterloo(滑铁卢大学)
;
University of Melbourne(墨尔本大学)
;
University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
MBZUAI(马克斯·普朗克智能系统研究所)
;
Macquarie University(麦考瑞大学)
机构
*
UCL Centre for Artificial Intelligence(伦敦大学学院人工智能中心)
;
Department of Computer Science(计算机科学系)
;
University College London(伦敦大学学院)
;
UK AI Security Institute(英国人工智能安全研究所)
;
Department of Computer Science and Engineering(计算机科学与工程系)
;
University of Bologna(博洛尼亚大学)
Sem-DPO: Mitigating Semantic Inconsistency in Preference Optimization for Prompt Engineering
Anas Mohamed, Azal Ahmad Khan, Xinran Wang, Ahmad Faraz Khan, Shuwen Ge, Saman Bahzad Khan, Ayaan Ahmad, Ali Anwar
机构
*
University of Minnesota(明尼苏达大学)
;
Virginia Tech(弗吉尼亚理工大学)
;
Xi’an University of Technology(西安理工大学)
;
Lahore University of Management Sciences(拉合尔管理科学大学)
;
University of California, Santa Cruz(加州大学圣克鲁兹分校)
AVC-DPO: Aligned Video Captioning via Direct Preference Optimization
Jiyang Tang, Hengyi Li, Yifan Du, Wayne Xin Zhao
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学人工智能学院)
;
College of Artificial Intelligence, Nankai University(南开大学人工智能学院)
;
School of Computer Science and Technology, Beijing Institute of Technology(北京理工大学计算机科学与技术学院)
机构
*
State Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室)
;
School of Artificial Intelligence, Nanjing University(南京大学人工智能学院)
;
University of Chicago(芝加哥大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
East China Normal University(华东师范大学)