机构
*
Speech Technology Lab, University of Groningen, The Netherlands(格罗宁根大学语音技术实验室,荷兰)
;
Center for Language and Cognition, University of Groningen, The Netherlands(格罗宁根大学语言与认知中心,荷兰)
机构
*
EPFL(苏黎世联邦理工学院)
;
Amazon AGI(亚马逊人工智能实验室)
;
University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Cornell University(康奈尔大学)
;
Qualcomm AI Research(高通人工智能研究)
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
The Grainger College of Engineering, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校格拉inger工程学院)
;
CAS Center for Excellence in Brain Science and Intelligence Technology(中国科学院脑科学与智能技术卓越创新中心)
;
Joint Laboratory of Intelligence Science and Technology, Institute of Systems Engineering, Macau University of Science and Technology(澳门科技大学系统工程学院智能科学与技术联合实验室)
机构
*
University of Alberta(阿尔伯塔大学)
;
University of Toronto(多伦多大学)
;
Zhejiang University(浙江大学)
;
School of Computer Science McGill University(麦吉尔大学计算机学院)
;
Alibaba Group (Ant Group)(阿里巴巴集团(蚂蚁集团))
;
Chinese Academy of Sciences(中国科学院)
机构
*
National Engineering Research Center for Mobile Network Technologies, Beijing University of Posts and Telecommunications(移动网络技术国家工程研究中心,北京邮电大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Singapore University of Technology and Design(新加坡科技设计大学)
;
Department of Electronic Engineering, Kyung Hee University(韩国庆熙大学电子工程系)
专题命中
后训练与偏好优化
:LLM(title);large language model(abstract);language model(abstract);分类 cs.AI
CommentsAfter further internal discussion, our author team has decided to withdraw this submission due to the need for several important refinements to the manuscript. All co-authors have been informed and agree with this decision
Human-assisted Robotic Policy Refinement via Action Preference Optimization
Wenke Xia, Yichu Yang, Hongtao Wu, Xiao Ma, Tao Kong, Di Hu
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China, Beijing(中国人民大学北京校区人工智能学院)
;
Engineering Research Center of Next-Generation Intelligent Search(下一代智能搜索与推荐工程研究中心)
;
Beijing Key Laboratory of Research on Large Models(北京大型模型研究重点实验室)
AdaRewriter: Unleashing the Power of Prompting-based Conversational Query Reformulation via Test-Time Adaptation
Yilong Lai, Jialong Wu, Zhenglin Wang, Deyu Zhou
机构
*
School of Computer Science and Engineering, Key Laboratory of Computer Network and Information Integration, Ministry of Education, Southeast University(计算机科学与工程学院、计算机网络与信息集成重点实验室、教育部、东南大学)
SPOGW: a Score-based Preference Optimization method via Group-Wise comparison for workflows
Yitong Cui, Liu Liu, Baosheng Yu, Jiayan Qiu, Xikai Zhang, Likang Xiao, Yixing Liu, Quan Chen
机构
*
Hangzhou International Innovation Institute and , Beihang University(杭州国际创新研究院和 北航)
;
School of Artificial Intelligence, Beihang University(人工智能学院,北航)
;
Nanyang Technological University(南洋理工大学)
;
University of Leicester(莱斯特大学)
;
China Mobile Communications Company Limited Research Institute(中国移动通信有限公司研究院)
专题命中
后训练与偏好优化
:preference optimization(title);large language model(abstract);language model(abstract);分类 cs.AI