Large Reasoning Models Learn Better Alignment from Flawed Thinking
大规模推理模型通过 flawed thinking 学习更好的对齐
机构 * Meta Superintelligence Labs(Meta超级智能实验室) ; IBM Research(IBM研究院)
专题命中 复杂问题求解 :reasoning(title,abstract);chain-of-thought(abstract);CoT(abstract);分类 cs.LG
AI总结 RECAP通过强化学习提升模型安全对齐能力,减少偏见并增强鲁棒性,同时保持推理效率。