Evolution Strategies at Scale: LLM Fine-Tuning Beyond Reinforcement Learning
大规模进化策略:超越强化学习的LLM微调
Xin Qiu, Yulu Gan, Conor F. Hayes, Qiyao Liang, Yinggan Xu, Roberto Dailey, Elliot Meyerson, Babak Hodjat, Risto Miikkulainen
机构
*
University of California, Los Angeles, Los Angeles, CA, USA(加州大学洛杉矶分校)
;
Cognizant AI Lab, San Francisco, CA, USA(Cognizant AI实验室)
;
The University of Texas at Austin, Austin, TX, USA(德克萨斯大学奥斯汀分校)
;
Massachusetts Institute of Technology, Cambridge, MA, USA(麻省理工学院)
专题命中
指令微调
:LLM(title,title_cn);large language model(abstract);language model(abstract);post-training(abstract)
机构
*
City University of Hong Kong(香港城市大学)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
The Chinese University of Hong Kong(香港中文大学)
机构
*
School of Biomedical Engineering, Tsinghua University(清华大学生物医学工程学院)
;
School of Biomedical Engineering, Shanghai Jiao Tong University(上海交通大学生物医学工程学院)
;
DAMO Academy, Alibaba Group(阿里云达摩院)
;
Hupan Laboratory(壶辰实验室)
;
Department of Biomedical Engineering, National University of Singapore(新加坡国立大学生物医学工程系)
;
Department of Radiology, Guizhou Provincial People’s Hospital(贵州省级人民医院放射科)
;
Department of Radiology, The First Affiliated Hospital, Zhejiang University School of Medicine(浙江大学医学院附属第一医院放射科)
;
Department of Radiology, Shanghai Sixth People’s Hospital Affiliated to Shanghai Jiao Tong University School of Medicine(上海交通大学医学院附属第六人民医院放射科)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
Auditable Context-Aware HFMD Forecasting with Structured LLM Agents
基于结构化大语言模型代理的可审计上下文感知手足口病预测
Joongwon Chae, Runming Wang, Chen Xiong, Gong Yunhan, Lian Zhang, Ji Jiansong, Dongmei Yu, Peiwu Qin
机构
*
Institute of Biomedicine and Health Engineering, Tsinghua Shenzhen International Graduate School (SIGS), Tsinghua University, Shenzhen, China(清华大学深圳国际研究生院生物医学与健康工程学院)
;
The Fifth Affiliated Hospital of Wenzhou Medical University, Lishui 323000, China(温州医科大学第五附属医院)
;
Hengqin Laboratory, Zhuhai, China(横琴实验室)
;
The First Hospital of Hebei Medical University, Shijiazhuang, China(河北医科大学第一医院)
A Neurosymbolic Approach to Natural Language Formalization and Verification
一种用于自然语言形式化和验证的神经符号方法
Chenyang An, Sam Bayless, Stefano Buliani, Darion Cassel, Byron Cook, Duncan Clough, Rémi Delmas, Nafi Diallo, Ferhat Erata, Nick Feng, Dimitra Giannakopoulou, Aman Goel, Aditya Gokhale, Joe Hendrix, Victor Heorhiadi, Marc Hudak, Dejan Jovanović, Andrew M. Kent, Benjamin Kiesl-Reiter, Jeffrey J. Kuna, Nadia Labai, Joseph Lilien, Divya Raghunathan, Zvonimir Rakamarić, Niloofar Razavi, Michael Tautschnig, Ali Torkamani, Nathaniel Weir, Michael W. Whalen, Jianan Yao
机构
*
Amazon Web Services(亚马逊网络服务)
;
University College London(伦敦大学学院)
;
University of Toronto(多伦多大学)
;
Queen Mary University of London(伦敦大学女王学院)
专题命中
推理与问题求解
:large language model(abstract);language model(abstract);分类 cs.CL、cs.AI、cs.LG
When and Why Does Multi-Agent Debate Fail and Does It Really Underperform?
多智能体辩论何时以及为何失败,它真的表现不佳吗?
Yongqiang Chen, Gang Niu, James Cheng, Bo Han, Masashi Sugiyama
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
RIKEN Center for Advanced Intelligence Project(日本理化学研究院高级智能项目中心)
;
Hong Kong Baptist University(香港 Baptist大学)
;
The University of Tokyo(东京大学)
专题命中
推理与问题求解
:large language model(abstract);language model(abstract);分类 cs.LG