Welfare, Improvability, and Variance: A Principal-Agent Approach to Optimal Benchmark Item Aggregation
福利、可改进性与方差:最优基准测试项聚合的主-代理方法
机构 * Department of Economics & Computer Science(经济与计算机科学系) ; Institute for Computational and Mathematical Engineering(计算与数学工程研究所) ; Department of Computer Science(计算机科学系) ; Department of Aeronautics & Astronautics(航空与航天系)
专题命中 Agent评测 :agent(title,abstract);分类 cs.LG
AI总结 提出将基准测试建模为多任务主-代理博弈,通过福利、可改进性和方差三个维度评估项目,并应用于OLMES数据集识别帕累托劣势项目。