Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?
Terminus-4B:一个较小的模型能否在智能体执行任务中取代前沿大语言模型?
机构 * Microsoft USA(微软公司)
专题命中 代码评测 :coding agent(abstract);分类 cs.SE、cs.AI
AI总结 研究智能体终端执行任务中较小模型能否取代前沿大语言模型。提出经监督微调及强化学习训练的Terminus-4B模型,评估发现其能减少主智能体令牌使用量约30%,不影响性能,还缩小与前沿模型差距甚至超越其性能。
Comments The current article involves some product IP issues and needs to be withdrawn and re-approved