Breaking the Tokenizer Barrier: On-Policy Distillation across Model Families
打破分词器壁垒:跨模型系列的在线策略蒸馏
Yifan Niu, Han Xiao, Dongyi Liu, Zelong Wang, Dihong Gong, Yasheng Wang, Jia Li
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Tencent(腾讯)
;
The Hong Kong University of Science and Technology(香港科技大学)
专题命中
效率与部署
:SFT(abstract,abstract_cn);LLM(abstract_cn);large language model(abstract);language model(abstract)
机构
*
Ningbo Institute of Digital Twin, Eastern Institute of Technology, Ningbo(宁波数字孪生研究院,东方理工大学(宁波))
;
Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东省人工智能与数字经济实验室(深圳))
专题命中
效率与部署
:LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG
Data-Efficient Autoregressive-to-Diffusion Language Models via On-Policy Distillation
数据高效的自回归到扩散语言模型通过策略内蒸馏
Xingyu Su, Jacob Helwig, Shubham Parashar, Atharv Chagi, Lakshmi Jotsna, Degui Zhi, James Caverlee, Dileep Kalathil, Shuiwang Ji
机构
*
Department of Computer Science and Engineering, Texas A&M University(德克萨斯大学阿马尔科分校计算机科学与工程系)
;
Department of Bioinformatics and Systems Medicine, University of Texas Health Science Center at Houston(德克萨斯大学健康科学中心休斯顿分校生物信息学与系统医学系)
;
Department of Electrical and Computer Engineering, Texas A&M University(德克萨斯大学阿马尔科分校电气与计算机工程系)
CommentsAccepted to IEEE International Symposium on Mixed and Augmented Reality (ISMAR) 2026, to appear in IEEE Transactions on Visualization and Computer Graphics (TVCG). 11 pages
LightSTAR: Efficient Visual Document Retrieval via Lightweight Selection with Vision-Adaptive Refinement
LightSTAR: 通过视觉自适应精炼的轻量级选择实现高效视觉文档检索
Tongkun Guan, Haocheng Wang, Wei Shen, Xiaokang Yang
机构
*
MoE Key Lab of Artificial Intelligence, AI Institute, School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学与工程学院人工智能研究院教育部人工智能重点实验室)
专题命中
效率与部署
:LLM(summary_cn,abstract);large language model(abstract);language model(abstract)