Learning social norms enhances compatibility in dynamic human-AI coordination
学习社会规范可增强动态人机协作中的兼容性
Yi Yang, Siyuan Liu, Xin Gao, Huamu Sun, Chao Liu, Qing Zhou, Bingbing Nie
机构
*
School of Vehicle and Mobility, Tsinghua University, Beijing, China(清华大学车辆与移动系统学院)
;
State Key Laboratory of Intelligent Green Vehicle and Mobility, Tsinghua University, Beijing 100084, China(清华大学智能绿色车辆与移动系统国家重点实验室)
;
State Key Laboratory of Cognitive Neuroscience and Learning & IDG/McGovern Institute for Brain Research, Beijing Normal University, Beijing, China(北京师范大学认知神经科学与学习国家重点实验室)
;
Beijing Key Laboratory of Safe AI and Superalignment, Beijing, China(北京安全人工智能与超对齐关键实验室)
;
Beijing Institute of AI Safety and Governance, Beijing, China(北京人工智能安全与治理研究院)
Strategic Bargaining in Multi-Buyer Markets: Reinforcement Learning from Verifiable Rewards for LLM Negotiations
多买家市场中的战略谈判:基于可验证奖励的强化学习用于大语言模型谈判
Shuze Daniel Liu, Claire Chen, Jiabao Sean Xiao, Xin Chen, David Simchi-Levi
机构
*
Institute for Data, Systems, and Society, Massachusetts Institute of Technology(数据、系统与社会研究所,麻省理工学院)
;
Mitch Daniels School of Business, Purdue University(米奇·丹尼尔斯商学院,普渡大学)
;
The Division of Physics, Mathematics and Astronomy, California Institute of Technology(物理、数学与天文学部,加州理工学院)
;
Department of Computing and Mathematical Sciences, California Institute of Technology(计算与数学科学系,加州理工学院)
;
H. Milton Stewart School of Industrial and Systems Engineering, Georgia Institute of Technology(H. 米尔顿·斯图尔特工业与系统工程学院,佐治亚理工学院)
;
Department of Civil and Environmental Engineering, Operations Research Center, Massachusetts Institute of Technology(土木与环境工程系、运筹学中心,麻省理工学院)