LLM Self-Correction with DeCRIM: Decompose, Critique, and Refine for Enhanced Following of Instructions with Multiple Constraints
基于DeCRIM的大语言模型自校正:分解、批判与优化,提升多约束指令遵循能力
专题命中 评测与基准 :LLM(title,summary_cn);分类 cs.CL、cs.AI、cs.LG
AI总结 针对LLM处理多约束指令的不足,研究推出RealInstruct基准,提出DeCRIM自校正流程,可提升开源模型性能,结合强反馈时其表现能超越GPT-4。
Comments EMNLP 2024, see https://aclanthology.org/2024.findings-emnlp.458/