arXivDaily arXiv每日学术速递 周一至周五更新

大厂专区

Intel(英特尔)

2026-08-04 至 2026-08-04 共收录 1
2608.01078 2026-08-04 cs.CL cs.AI 新提交

Attend to Your Own Thoughts: Breaking the Barrier for Post-Training Quantization of Reasoning LLMs through the Lens of 1.58-Bit Quantization

关注自身思维:通过1.58比特量化视角打破推理型大语言模型的后训练量化障碍

Shigeng Wang, Chao Li, Yangyuxuan Kang, Jiawei Fan, Anbang Yao

机构 * Intel Labs China(英特尔中国实验室)

AI总结 该研究提出ScaleQ-1.58三值后训练量化框架,通过集成AYOT校准方法,提升推理型LLM量化性能,该框架可扩展且泛化性强,仅需少量校准token即可实现优异效果。

Comments This research work was completed and submitted for publication in early May 2026. The project page: https://github.com/IntelChina-AI/BitTern

详情

展开后加载摘要…

URL PDF HTML 收藏