Information Suppression in Large Language Models: Auditing, Quantifying, and Characterizing Censorship in DeepSeek
大语言模型中的信息压制:对DeepSeek的信息审查、量化与特征化
Peiran Qiu, Siyi Zhou, Emilio Ferrara
机构
*
Thomas Lord Department of Computer Science, University of Southern California, USA(汤姆斯·劳德计算机科学系,南加州大学)
;
Information Sciences Institute, University of Southern California, USA(信息科学研究所,南加州大学)
;
Annenberg School of Communication, University of Southern California, USA(安纳伯格传播学院,南加州大学)
Distilled Reinforcement Learning for LLM Post-training
用于大语言模型训练后处理的蒸馏强化学习
Chen Wang, Zhaochun Li, Jionghao Bai, Yining Zhang, Hexuan Deng, Ge Lan, Yue Wang
机构
*
College of Elite Engineers, Nankai University(南开大学精英工程师学院)
;
Zhongguancun Academy(中关村学院)
;
Beijing Institute of Technology(北京理工大学)
;
Zhejiang University(浙江大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Harbin Institute of Technology(哈尔滨工业大学)
;
College of Software, Nankai University(南开大学软件学院)
Token-Wise Latent Streaming from Slow Reasoners to Fast Planners for Dynamic Vision Language Navigation
用于动态视觉语言导航的从慢速推理器到快速规划器的逐令牌潜在流
Tianshuai Hu, Yangyi Zhong, Zeying Gong, Lingdong Kong, Xiaodong Mei, Guoyang Zhao, Xiaolu Liu, Song Wang, Rong Li, Junwei Liang
机构
*
The Hong Kong University of Science and Technology(香港科技大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
National University of Singapore(新加坡国立大学)
;
Zhejiang University(浙江大学)