EstLLM: Enhancing Estonian Capabilities in Multilingual LLMs via Continued Pretraining and Post-Training
EstLLM:通过持续预训练和后训练增强多语言大语言模型中的爱沙尼亚能力
Aleksei Dorkin, Taido Purason, Emil Kalbaliyev, Hele-Andra Kuulmets, Marii Ojastu, Mark Fišel, Tanel Alumäe, Eleri Aedmaa, Krister Kruusmaa, Kairit Sirts
机构
*
Institute of Computer Science, University of Tartu(塔尔图大学计算机科学研究院)
;
Department of Software Science, Tallinn University of Technology(塔林技术大学软件科学系)
;
Institute of the Estonian Language, Tallinn, Estonia(爱沙尼亚语言研究院)
;
School of Humanities, Tallinn University(塔林大学人文学院)
According to Me: Long-Term Personalized Referential Memory QA
根据我:长期个性化参照记忆问答
Jingbiao Mei, Jinghong Chen, Guangyu Yang, Xinyu Hou, Margaret Li, Bill Byrne
机构
*
Department of Engineering, University of Cambridge, United Kingdom(剑桥大学工程系)
;
Department of Physics, University of Cambridge, United Kingdom(剑桥大学物理系)
;
Independent Researcher(独立研究者)
PanCanBench: A Comprehensive Benchmark for Evaluating Large Language Models in Pancreatic Oncology
PanCanBench: 用于评估大型语言模型在胰腺肿瘤学中的综合基准
Yimin Zhao, Sheela R. Damle, Simone E. Dekker, Scott Geng, Karly Williams Silva, Jesse J Hubbard, Manuel F Fernandez, Fatima Zelada-Arenas, Alejandra Alvarez, Brianne Flores, Alexis Rodriguez, Stephen Salerno, Carrie Wright, Zihao Wang, Pang Wei Koh, Jeffrey T. Leek
机构
*
Department of Biostatistics, University of Washington(华盛顿大学生物统计学系)
;
Clinical Research Division, Fred Hutch Cancer Center(Fred Hutch癌症中心临床研究部)
;
Division of Hematology and Oncology, Department of Medicine, University of Washington(华盛顿大学医学系血液学与肿瘤学分会)
;
Allen Institute for AI(Allen人工智能研究所)
;
Department of Computer Science and Engineering, University of Washington(华盛顿大学计算机科学与工程系)
;
Public Health Sciences, Biostatistics, Fred Hutchinson Cancer Center(Fred Hutchinson癌症中心公共卫生科学与生物统计学)
CMT-Benchmark: A Benchmark for Condensed Matter Theory Built by Expert Researchers
CMT-Benchmark:由专家研究人员构建的凝聚态理论基准
Haining Pan, James V. Roggeveen, Erez Berg, Juan Carrasquilla, Debanjan Chowdhury, Surya Ganguli, Federico Ghimenti, Juraj Hasik, Henry Hunt, Hong-Chen Jiang, Mason Kamb, Ying-Jer Kao, Ehsan Khatami, Michael J. Lawler, Di Luo, Titus Neupert, Xiaoliang Qi, Michael P. Brenner, Eun-Ah Kim
机构
*
Rutgers University(罗格斯大学)
;
Harvard University(哈佛大学)
;
Weizmann Institute of Science(魏茨曼科学研究所)
;
ETH Zürich(苏黎世联邦理工学院)
;
Cornell University(康奈尔大学)
;
Stanford University(斯坦福大学)
;
University of Zürich(苏黎世大学)
;
Stanford Institute for Materials and Energy Sciences(斯坦福材料与能源科学研究所)
;
SLAC National Accelerator Laboratory(斯坦福直线加速器实验室)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
National Taiwan University(台湾大学)
;
San José State University(圣何塞州立大学)
CommentsWe have decided to withdraw this manuscript because we believe it requires further revision and substantial improvement before it is suitable for dissemination to the academic community
机构
*
Zhejiang University(浙江大学)
;
Binjiang Institute of Zhejiang University(浙江大学滨江学院)
;
Om AI Research(奥姆人工智能研究)
;
Stanford University(斯坦福大学)
;
ETH Zürich(苏黎世联邦理工学院)
机构
*
UNC at Chapel Hill(北卡罗来纳大学教堂山分校)
;
University of Science and Technology of China(中国科学技术大学)
;
George Mason University(乔治·马歇尔大学)
;
Yale University(耶鲁大学)
;
City University of Hong Kong(香港城市大学)
;
Rice University(得克萨斯大学奥斯汀分校)
Decoupling Strategy and Execution in Task-Focused Dialogue via Goal-Oriented Preference Optimization
通过目标导向的偏好优化实现任务导向对话中的解耦策略与执行
Jingyi Xu, Xingyu Ren, Zhoupeng Shou, Yumeng Zhang, Zhiqiang You
机构
*
School of Information Engineering, China Jiliang University, Hangzhou, China(信息工程学院,中国浙江大学,杭州,中国)
;
College of Chemical and Biological Engineering, Zhejiang University, Hangzhou, China(化学与生物工程学院,浙江大学,杭州,中国)
VerifyBench: Benchmarking Reference-based Reward Systems for Large Language Models
VerifyBench: 大型语言模型参考式奖励系统的基准测试
Yuchen Yan, Jin Jiang, Zhenbang Ren, Yijun Li, Xudong Cai, Yang Liu, Xin Xu, Mengdi Zhang, Jian Shao, Yongliang Shen, Jun Xiao, Yueting Zhuang
机构
*
Zhejiang University(浙江大学)
;
Meituan Group(美团集团)
;
Peking University(北京大学)
;
University of Electronic Science and Technology of China(电子科技大学)
;
The Hong Kong University of Science and Technology(香港科技大学)
Naeimeh Nourmohammadi, Md Meem Hossain, The Anh Han, Safina Showkat Ara, Zia Ush Shamszaman
机构
*
Department of Computing
;
Games, Teesside University, Middlesbrough, United Kingdom Centre for Digital Innovation, Teesside University, Middlesbrough, United Kingdom Faculty of Business \& Technology, University of Sunderland, Sunderland, United Kingdom