CoopEval: Benchmarking Cooperation-Sustaining Mechanisms and LLM Agents in Social Dilemmas
CoopEval:社会困境中合作维持机制与LLM代理的基准测试
Emanuel Tewolde, Xiao Zhang, David Guzman Piedrahita, Vincent Conitzer, Zhijing Jin
机构
*
Carnegie Mellon University
;
Foundations of Cooperative AI Lab (FOCAL)
;
Jinesis Lab, University of Toronto \& Vector Institute
;
ETH Z\" u rich
;
Max Planck Institute for Intelligent Systems, T\" u bingen, Germany
TokSuite: Measuring the Impact of Tokenizer Choice on Language Model Behavior
TokSuite:衡量分词器选择对语言模型行为的影响
Gül Sena Altıntaş, Malikeh Ehghaghi, Brian Lester, Fengyuan Liu, Wanru Zhao, Marco Ciccone, Colin Raffel
机构
*
University of Toronto(多伦多大学)
;
Vector Institute(向量研究所)
;
Google DeepMind(谷歌DeepMind)
;
McGill University(麦吉尔大学)
;
University of Cambridge(剑桥大学)
;
Hugging Face