VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models
VALUEFLOW:迈向大语言模型中多元化和可引导的基于价值的对齐
Woojin Kim, Sieun Hyeon, Jusang Oh, Jaeyoung Do
机构
*
Department of Electrical and Computer Engineering, Seoul National University(首尔国立大学电气与计算机工程系)
;
Interdisciplinary Program in Artificial Intelligence, Seoul National University(首尔国立大学人工智能交叉学科项目)
机构
*
The Grainger College of Engineering, Nuclear, Plasma & Radiological Engineering, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校格雷格学院工程学院、核等工程学院)
;
Department of Nuclear Engineering, Hanyang University(汉阳大学核工程系)
;
University of Texas - El Paso(德克萨斯大学埃尔帕索分校)
;
National Center for Supercomputing Applications(国家超级计算应用中心)
;
Department of Applied Mechanics, Indian Institute of Technology Delhi(印度德里理工学院应用力学系)
;
Yardi School of Artificial Intelligence, Indian Institute of Technology Delhi(印度德里理工学院亚里人工智能学院)
InvEvolve: Evolving White-Box Inventory Policies via Large Language Models with Performance Guarantees
InvEvolve:通过具有性能保证的大语言模型进化白盒库存策略
Chenyu Huang, Jianghao Lin, Zhengyang Tang, Bo Jiang, Ruoqing Jiang, Benyou Wang, Lai Wei
机构
*
Shanghai University of Finance and Economics(上海财经大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Tsinghua University(清华大学)
;
Boston College(波士顿大学)
Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models
自回归与扩散语言模型中的逐步拒绝动态
Eliron Rahimi, Elad Hirshel, Rom Himelstein, Amit LeVi, Avi Mendelson, Chaim Baskin
机构
*
Department of Computer Science, Technion – Israel Institute of Technology(技术学院计算机科学系,以色列技术学院)
;
INSIGHT Lab, School of Electrical and Computer Engineering, Ben-Gurion University of the Negev, Israel(内斯坦实验室,贝内-加隆大学内加尔分校,以色列)
;
Computer Science Department, University of Haifa, Haifa, Israel(海法大学计算机科学系,海法,以色列)
Robust Driving Control for Autonomous Vehicles: An Intelligent General-sum Constrained Adversarial Reinforcement Learning Approach
自动驾驶鲁棒控制:一种智能一般和约束对抗强化学习方法
Junchao Fan, Qi Wei, Ruichen Zhang, Yang Lu, Jianhua Wang, Xiaolin Chang, Bo Ai
机构
*
Beijing Key Laboratory of Security and Privacy in Intelligent Transportation(北京智能交通安全与隐私重点实验室)
;
Beijing Jiaotong University(北京交通大学)
;
College of Computing and Data Science(计算与数据科学学院)
;
Nanyang Technological University(南洋理工大学)
;
School of Computer Science and Technology(计算机科学与技术学院)
;
Taiyuan University of Technology(太原科技大学)
;
School of Electronics and Information Engineering(电子与信息工程学院)
Teaching the Way, Not the Answer: Privileged Tutoring Distillation for Multimodal Policy Optimization
教方法而非答案:用于多模态策略优化的特权辅导蒸馏
Shizhe Xiang, Ke An, Wenlong Yu, Yue Liu, Jian Luan, Pei Fu, Qilong Wang
机构
*
Tianjin University(天津大学)
;
Beijing Institute of Technology(北京理工大学)
;
Singapore Management University(新加坡国立大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Xiaomi Inc(小米公司)
Zhanhao Hu, Xiao Huang, Patrick Mendoza, Emad A. Alghamdi, Basel Alomair, Raluca Ada Popa, David Wagner
机构
*
University of California, Berkeley(加州大学伯克利分校)
;
HUMAIN
;
King Abdulaziz City for Science and Technology(国王阿卜杜勒阿齐兹科学与技术城)
;
University of Washington, Seattle(华盛顿大学(西雅图))
机构
*
School of Cyber Science and Engineering, Wuhan University(武汉大学计算机科学与工程学院)
;
School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学学院)
;
Independent Researcher(独立研究者)
Causality Is Key to Understand and Balance Multiple Goals in Trustworthy ML and Foundation Models
因果关系是理解和平衡可信机器学习与基础模型中多个目标的关键
Ruta Binkyte, Ivaxi Sheth, Zhijing Jin, Mohammad Havaei, Bernhard Schölkopf, Mario Fritz
机构
*
CISPA Helmholtz Center for Information Security(CISPA海德堡信息安全中心)
;
Max Planck Institute for Intelligent Systems, Tübingen(马克斯·普朗克智能系统研究所(图宾根))
;
Google Research(谷歌研究)
;
ETH Zürich(苏黎世联邦理工学院)
;
University of Toronto(多伦多大学)