Illusions of the Gold Standard: A Large-scale Analysis of Human Evaluation Protocols for Long-form Text Generation
黄金标准的幻觉:长文本生成中人类评估协议的大规模分析
Katelyn Xiaoying Mei, Yi-Li Hsu, Minjoon Choi, Zongwan Cao, Chenjun Xu, Bingbing Wen, Su Lin Blodgett, Lucy Lu Wang
机构
*
University of Washington(华盛顿大学)
;
National Tsing Hua University(国立清华大学)
;
Seoul National University(首尔大学)
;
Mila - Québec AI Institute(米拉-魁北克人工智能研究所)
;
Allen Institute for AI(艾伦人工智能研究所)
机构
*
Department of Economics, University of Washington(华盛顿大学经济系)
;
Paul G. Allen School of Computer Science & Engineering, University of Washington(华盛顿大学保罗·G·艾伦计算机科学与工程学院)
Learning to Keep a Promise: Scaling Language Model Decoding Parallelism with Learned Asynchronous Decoding
学习承诺:通过学习异步解码扩展语言模型解码并行性
Tian Jin, Ellie Y. Cheng, Zack Ankner, Nikunj Saunshi, Blake M. Elias, Amir Yazdanbakhsh, Jonathan Ragan-Kelley, Suvinay Subramanian, Michael Carbin
机构
*
DeepMind, London, UK(深度思维公司,伦敦,英国)
;
Google Research, New York, NY, USA(谷歌研究院,纽约,纽约州,美国)
;
Stanford University, Stanford, CA, USA(斯坦福大学,斯坦福,加利福尼亚州,美国)
;
University of Toronto, Toronto, Ontario, Canada(多伦多大学,多伦多,安大略省,加拿大)
;
University of Washington, Seattle, WA, USA(华盛顿大学,西雅图,华盛顿州,美国)