Hit-RAG: Learning to Reason with Long Contexts via Preference Alignment
通过偏好对齐学习长上下文的推理
机构 * Tongji University(同济大学) ; The City University of New York(纽约城市大学) ; University of Technology Sydney(悉尼大学) ; Huazhong University of Science and Technology(华中科技大学) ; Shenzhen University of Advanced Technology(深圳先进技术大学)
专题命中 偏好对齐 :alignment(title,abstract);分类 cs.CL、cs.AI
AI总结 Hit-RAG通过多阶段偏好对齐框架解决长上下文推理中的注意力稀释和幻觉问题,提升模型在长上下文场景下的推理能力。
Comments 21 pages, 2 figures, 6 tables