Try Again, Don't Look Back: Blind Resampling Outperforms Self-Repair in Small Code Models
再试一次,不要回头:小型代码模型中盲重采样优于自我修复
专题命中 代码评测 :code model(title);分类 cs.SE、cs.AI、cs.LG
AI总结 该研究针对MBPP+数据集在三种规模代码模型上发现,盲重采样在7B以下表现最优,成本远低于其他重试方法,而基于自身失败尝试的自我修复存在锚定效应导致的性能损失。
Comments Code, pre-registrations and run traces: https://github.com/vermayuvraj/self-improving-agent