Escaping the Context Bottleneck: Active Context Curation for LLM Agents via Reinforcement Learning
突破上下文瓶颈:通过强化学习实现LLM代理的主动上下文精修
机构 * Tongji University(同济大学) ; Stanford University(斯坦福大学) ; CurrentsAI Research(CurrentsAI 研究院)
专题命中 长上下文与记忆 :LLM(title);large language model(abstract);language model(abstract);foundation model(abstract)
AI总结 本文提出一种解耦上下文管理和任务执行的框架,通过强化学习训练轻量级策略模型ContextCurator,有效减少工作内存中的信息熵,提升LLM在长周期任务中的表现。