VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models
VALUEFLOW:迈向大语言模型中多元化和可引导的基于价值的对齐
Woojin Kim, Sieun Hyeon, Jusang Oh, Jaeyoung Do
机构
*
Department of Electrical and Computer Engineering, Seoul National University(首尔国立大学电气与计算机工程系)
;
Interdisciplinary Program in Artificial Intelligence, Seoul National University(首尔国立大学人工智能交叉学科项目)
Differentially Private Preference Data Synthesis for Large Language Model Alignment
面向大语言模型对齐的差分隐私偏好数据合成
Fengyu Gao, Jing Yang
机构
*
Department of Computer Science, University of Virginia, Charlottesville, Virginia, USA(弗吉尼亚大学计算机科学系)
;
Department of Electrical and Computer Engineering, University of Virginia, Charlottesville, Virginia, USA(弗吉尼亚大学电气与计算机工程系)
Bridging the Stability-Expressivity Gap: Synthetic Data Scaling and Preference Alignment for Low-Resource Spoken Language Models
弥合稳定性与表现力之间的差距:低资源口语语言模型的合成数据扩展与偏好对齐
Yizhong Geng, Yanliang Li, Jinghan Yang, Tianhan Jiang, Boxun An, Ya Li, Xiaoyu Shen
机构
*
Beijing University of Posts(北京邮电大学)
;
University of California, USA(美国加州大学)
;
Northwestern University, USA(美国西北大学)
;
Eastern Institute of Technology, Ningbo, China(宁波工程技术学院)
机构
*
Department of Computer Science, National University of Singapore(新加坡国立大学计算机科学系)
;
Singapore-MIT Alliance for Research and Technology Centre(新加坡-麻省理工联盟研究技术中心)
;
The Chinese University of Hong Kong, Shenzhen, China(香港中文大学(深圳))
;
CSAIL, Massachusetts Institute of Technology(麻省理工学院计算机科学与人工智能实验室)
;
Institute of Data Science, National University of Singapore(新加坡国立大学数据科学研究院)
机构
*
Tsinghua University(清华大学)
;
International Digital Economy Academy(国际数字经济学院)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
Nanyang Technological University(南洋理工大学)
Turning Drift into Constraint: Robust Reasoning Alignment in Non-Stationary Multi-Stream Environments
将漂移转化为约束:非稳态多流环境中的鲁棒推理对齐
Xiaoyu Yang, En Yu, Wei Duan, Jie Lu
机构
*
Australian Artificial Intelligence Institute (AAII)(澳大利亚人工智能研究所)
;
Faulty of Engineering and Information Technology(工程与信息技术学院)
;
University of Technology Sydney(悉尼技术大学)
;
Australia(澳大利亚)
机构
*
The Hong Kong Polytechnic University(香港理工大学)
;
Huawei Technologies Ltd.(华为技术有限公司)
;
National University of Singapore(新加坡国立大学)
;
University of Science and Technology of China(中国科学技术大学)
LVRPO: Language-Visual Alignment with GRPO for Multimodal Understanding and Generation
LVRPO:基于GRPO的语言-视觉对齐用于多模态理解和生成
Shentong Mo, Sukmin Yun
机构
*
Department of Machine Learning, CMU, USA(卡内基梅隆大学机器学习系,美国)
;
Department of Machine Learning, MBZUAI, UAE(穆罕默德·本·扎耶德人工智能大学机器学习系,阿联酋)
;
Department of Artificial Intelligence, Hanyang University ERICA, South Korea(汉阳大学ERICA校区人工智能系,韩国)
机构
*
Hong Kong JC STEM Lab of Smart City(香港JC STEM实验室)
;
Department of Computer Science, City University of Hong Kong(香港城市大学计算机科学系)
;
School of Computer Science and Technology, Xidian University(西安电子科技大学计算机科学与技术学院)
;
Department of Electrical and Computer Engineering, The University of Hong Kong(香港大学电子与计算机工程系)