D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models
D-OPSD:用于连续调优步蒸馏扩散模型的在线自蒸馏方法
Dengyang Jiang, Xin Jin, Dongyang Liu, Zanyi Wang, Mingzhe Zheng, Ruoyi Du, Xiangpeng Yang, Qilong Wu, Zhen Li, Peng Gao, Harry Yang, Steven Hoi
机构
*
The Hong Kong University of Science and Technology(香港科技大学)
;
Z-Image Team, Alibaba Group(阿里集团Z-Image团队)
;
University of California, San Diego(加州大学圣地亚哥分校)
;
The Chinese University of Hong Kong(香港中文大学)
Beyond Text Prompts: Visual-to-Visual Generation as A Unified Paradigm
超越文本提示:视觉到视觉生成作为统一范式
Yaofang Liu, Kangning Cui, Meng Chu, Zhaoqing Li, Suiyun Zhang, Jean-Michel Morel, Xiaodong Cun, Haoxuan Che, Rui Liu, Raymond H. Chan
机构
*
City University of Hong Kong(香港城市大学)
;
City University of Hong Kong (Dongguan)(香港城市大学(东莞))
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Celia Research HK(Celia研究香港)
;
Great Bay University(大湾大学)
;
Lingnan University(岭南大学)
机构
*
MMLab, The Chinese University of Hong Kong, Hong Kong SAR(香港理工大学MMLab,香港中文大学,香港特别行政区)
;
The Hong Kong University of Science(香港理工大学)
;
The University of Hong Kong(香港大学)
;
Tsinghua University(清华大学)
机构
*
School of Statistics, East China Normal University(东华大学统计学院)
;
Department of Statistics and Data Science, University of California, Los Angeles(加州大学洛杉矶分校统计与数据科学系)
;
Department of Industrial and Systems Engineering, University of Minnesota(明尼苏达大学工业与系统工程系)
;
School of Data Science, The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)数据科学学院)
机构
*
Multimedia Laboratory, The Chinese University of Hong Kong(香港中文大学多媒体实验室)
;
Beijing University of Posts(北京邮电大学)
;
Hong Kong University of Science(香港理工大学)
Pusa V1.0: Unlocking Temporal Control in Pretrained Video Diffusion Models via Vectorized Timestep Adaptation
Pusa V1.0: 通过向量化时间步长适应解锁预训练视频扩散模型中的时间控制
Yaofang Liu, Yumeng Ren, Aitor Artola, Yuxuan Hu, Xiaodong Cun, Xiaotong Zhao, Alan Zhao, Raymond H. Chan, Suiyun Zhang, Rui Liu, Dandan Tu, Jean-Michel Morel
机构
*
City University of Hong Kong(香港城市大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Huawei Research(华为研究)
;
Great Bay University(大湾大学)
;
AI Technology Center, Tencent PCG(腾讯AI技术中心)
;
Lingnan University(岭南大学)
;
Hong Kong Centre for Cerebro-Cardiovascular Health Engineering(香港脑心血管健康工程中心)
Universal Activation Verbalizer: A Unified Framework for Cross-Model Activation Explanation
通用激活词化器:跨模型激活解释的统一框架
Haiyan Zhao, Zirui He, Guanchu Wang, Ali Payani, Yingcong Li, Mengnan Du
机构
*
New Jersey Institute of Technology(新泽西理工学院)
;
University of North Carolina at Charlotte(北卡罗来纳大学夏洛特分校)
;
Cisco Research(思科研究)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
OMGTex: One-stage Multi-style Facial Texture Reconstruction without Geometry Guidance
OMGTex: 无需几何引导的一阶段多风格面部纹理重建
Zitong Xiao, Yuda Qiu, Zisheng Ye, Xiaoguang Han
机构
*
School of Science and Engineering, The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)科学与工程学院)
;
Guangdong Provincial Key Laboratory of Future Networks of Intelligence(广东省未来网络智能化重点实验室)
;
FNii-Shenzhen(FNii-深圳)
CRPO: Character-centric Group Relative Policy Optimization for Role-aware Reasoning in Role-playing Agents
CRPO:以角色为中心的群体相对策略优化用于角色扮演代理中的角色感知推理
Yihong Tang, Kehai Chen, Liang Yue, Benyou Wang, Min Zhang
机构
*
Institute of Computing and Intelligence(计算与智能研究院)
;
Harbin Institute of Technology(哈尔滨工业大学)
;
Shenzhen Loop Area Institute (SLAI)(深圳Loop区研究院)
;
The Chinese University of Hong Kong(香港中文大学)
机构
*
Department of Statistics and Data Science, Southern University of Science and Technology(统计与数据科学系,南方科技大学)
;
Data Science, Southern University of Science(数据科学,南方科技大学)
;
School of Data Science, The Chinese University of Hong Kong, Shenzhen, China(数据科学学院,香港中文大学(深圳))
机构
*
Fudan University(复旦大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
King’s College London(伦敦国王学院)
;
The Alan Turing Institute(艾伦·图灵研究所)
;
Shanghai Innovation Institute(上海创新研究院)
ViroBench: Benchmarking Nucleotide Foundation Models on Viral Genomics Tasks
ViroBench:病毒基因组学任务中的核苷酸基础模型基准测试
Dongxin Ye, Fang Hu, Han Hu, Shu Hu, Yang Tan, Wanli Ouyang, Stan Z. Li, Jie Cui, Nanqing Dong
机构
*
Shanghai Innovation Institute Shanghai China(深圳河套学院)
;
University of Electronic Science
;
Fudan University Shanghai China
;
Shanghai Artificial Intelligence Laboratory Shanghai China
;
Institute of Infection
;
Health Fudan University Shanghai China
;
Shanghai Sci-Tech Inno Center for Infection \& Immunity Shanghai China
;
Shanghai Jiao Tong University Shanghai China
;
Shenzhen Loop Area Institute Shenzhen China
;
Chinese University of Hong Kong Hong Kong China
;
Westlake University Hangzhou China
;
Shanghai Innovation Institute
;
Fudan University
;
Shanghai Artificial Intelligence Laboratory
;
Shanghai Sci-Tech Inno Center for Infection \& Immunity
;
Shanghai Jiao Tong University
;
Shenzhen Loop Area Institute
;
Chinese University of Hong Kong
;
Westlake University
机构
*
Peking University(北京大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Tsinghua University(清华大学)
;
Beihang University(北京航空航天大学)
;
Xiaohongshu Inc.(小红书公司)
Spurious Stationarity and Hardness Results for Bregman Proximal-Type Algorithms
Bregman近端类型算法的伪平稳性和困难结果
He Chen, Jiajin Li, Anthony Man-Cho So
机构
*
Department of Systems Engineering and Engineering Management, The Chinese University of Hong Kong(香港中文大学系统工程与工程管理系)
;
Sauder School of Business, University of British Columbia(不列颠哥伦比亚大学萨德勒商学院)
InvariantCloud: A Globally Invariant, Uniquely Indexed Point Cloud Framework for Robust 6-DoF Tactile Pose Tracking
InvariantCloud:一种全局不变、唯一索引的点云框架,用于鲁棒的6自由度触觉姿态跟踪
Pengfei Ye, Yuxiang Ma, Yi Zhou, Wei Chen, Wenzhen Dong, Molong Duan
机构
*
Department of Mechanical and Aerospace Engineering, The Hong Kong University of Science and Technology(香港科学与技术大学机械与航空航天工程系)
;
Department of Mechanical Engineering, Massachusetts Institute of Technology(麻省理工学院机械工程系)
;
Department of Mechanical and Automation Engineering, The Chinese University of Hong Kong(香港中文大学机械与自动化工程系)
Explainable Multi-Task Retinal Imaging Reveals Microvascular Signals for Systemic Risk Stratification in Type 2 Diabetes: A Pilot Study
可解释多任务视网膜成像揭示2型糖尿病系统性风险分层的微血管信号:一项初步研究
Mini Han Wang, Liting Huang, Wei Hong, Boonthawan Wingwon
机构
*
Faculty of Computer Science and Artificial Intelligence, Shenzhen University of Advanced Technology(深圳先进技术大学计算机科学与人工智能学院)
;
Frontier Science Computing Center, Zhuhai Institute of Advanced Technology Chinese Academy of Sciences(中国科学院珠海先进技术研究院前沿科学计算中心)
;
Chinese University of Hong Kong(香港中文大学)
;
Zhuhai People's Hospital (The Affiliated Hospital of Beijing Institute of Technology, Zhuhai Clinical Medical College of Jinan University)(珠海人民医院(北京理工大学珠海临床医学院附属医院))
;
Lampang Inter-Tech College, Lampang Thailand(泰国 Lampang 职业技术学院)
Explainable Retinal Imaging for Prediction of Multi-Organ Dysfunction in Type 2 Diabetes
可解释的视网膜成像用于预测2型糖尿病多器官功能障碍
Mini Han Wang, Liting Huang, Wei Hong, Boonthawan Wingwon
机构
*
Faculty of Computer Science and Artificial Intelligence(计算机科学与人工智能学院)
;
Frontier Science Computing Center(前沿科学计算中心)
;
Chinese Academy of Sciences(中国科学院)
;
Chinese University of Hong Kong(香港中文大学)
;
Zhuhai People's Hospital(珠海人民医院)
;
Beijing Institute of Technology(北京理工大学)
;
Jinan University(暨南大学)
;
Lampang Inter-Tech College
机构
*
University of Electronic Science and Technology of China(电子科技大学)
;
Singapore Management University(新加坡国立大学)
;
Michigan State University(密歇根州立大学)
;
The Chinese University of Hong Kong(香港中文大学)
HiMed: Incentivizing Hindi Reasoning in Medical LLMs
HiMed: 激励医疗大语言模型中的印地语推理
Dingfeng Jiang, Han Yan, Chenze Ma, Amit Kumar Jaiswal, Ang Li, Yunxiang Jiang, Xinlei Xiong, Juhao Liang, Hongru Xiao, Xiang Li, Fan Bu, Jiale Han, Ruchir Gupta, Prayag Tiwari, Benyou Wang
机构
*
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Indian Institute of Technology (Banaras Hindu University) Varanasi(印度理工学院(班加罗尔 Hindu 大学)瓦拉纳西分校)
;
Tongji University(同济大学)
;
Shenzhen Research Institute of Big Data(深圳大数据研究院)
;
Shenzhen Loop Area Institute(深圳河套学院)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Halmstad University(哈尔姆斯塔德大学)