机构
*
National Taiwan University(国立台湾大学)
;
Max Planck Institute for Psycholinguistics(马克斯·普朗克心理语言学研究所)
;
Radboud University(拉德堡德大学)
;
Institut Jean Nicod(让·尼科研究所)
Rethinking Data Curation in LLM Training: Online Reweighting Offers Better Generalization than Offline Methods
重新思考大语言模型训练中的数据整理:在线重加权优于离线方法
Wanru Zhao, Yihong Chen, Yuzhi Tang, Wentao Ma, Shengchao Hu, Shell Xu Hu, Alex Iacob, Abhinav Mehrotra, Nicholas D. Lane
机构
*
University of Cambridge(剑桥大学)
;
OATML, University of Oxford(牛津大学OATML实验室)
;
University of Toronto(多伦多大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Samsung AI Center(三星人工智能中心)
专题命中
预训练与数据
:LLM(title,abstract);large language model(abstract);language model(abstract);instruction tuning(abstract)
Efficacy of Large Language Models in Systematic Reviews
Aaditya Shah, Shridhar Mehendale, Siddha Kanthi
专题命中
预训练与数据
:large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.LG
CommentsBoth Shah and Mehendale contributed equally to this work; order of authorship is random. This paper will be published in the proceedings of The 2nd International Conference on Foundation and Large Language Models (FLLM2024) in IEEE Xplore