OSDM-MReg: Multimodal Image Registration based One Step Diffusion Model
OSDM-MReg: 基于一步扩散模型的多模态图像配准
Xiaochen Wei, Weiwei Guo, Wenxian Yu, Feiming Wei, Dongying Li
机构
*
Shanghai Key Laboratory of Intelligent Sensing and Recognition(上海智能感知与识别重点实验室)
;
Shanghai Jiao Tong University(上海交通大学)
;
Center of Digital Innovation(数字创新中心)
UltraViCo: Breaking Extrapolation Limits in Video Diffusion Transformers
UltraViCo: 突破视频扩散变换器的 extrapolation 限制
Min Zhao, Hongzhou Zhu, Yingze Wang, Bokai Yan, Jintao Zhang, Guande He, Ling Yang, Chongxuan Li, Jun Zhu
机构
*
Dept. of Comp. Sci. & Tech., BNRist Center, THU-Bosch ML Center, Tsinghua University(清华大学计算机科学与技术系,BNRist中心,THU-Bosch机器学习中心,清华大学)
;
ShengShu(盛书)
;
Gaoling School of Artificial Intelligence, Renmin University of China(北京理工大学人工智能学院,中国人民大学)
;
The University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
Princeton University(普林斯顿大学)
机构
*
Fudan University(复旦大学)
;
Tongyi Lab, Alibaba Group(阿里云实验室,阿里巴巴集团)
;
Zhejiang University(浙江大学)
;
The University of Hong Kong(香港大学)
;
MMLab
;
Institute of Science and Technology for Brain-inspired Intelligence, MOE Frontiers Center for Brain Science, Key Laboratory of Computational Neuroscience and Brain-Inspired Intelligence, and State Key Laboratory of Brain Function and Disorders, Fudan University(脑启发智能科学技术研究院、MOE前沿脑科学中心、计算神经科学与脑启发智能重点实验室、脑功能与疾病国家重点实验室,复旦大学)
机构
*
School of Artificial Intelligence, Beijing Normal University, Beijing 100875, China(人工智能学院,北京师范大学,北京)
;
Institute of Artificial Intelligence and Future Networks, Beijing Normal University, Zhuhai 519087, China(人工智能与未来网络研究院,北京师范大学,珠海)
SKeDA: A Generative Watermarking Framework for Text-to-video Diffusion Models
SKeDA:一种面向文本到视频扩散模型的生成水印框架
Yang Yang, Xinze Zou, Zehua Ma, Han Fang, Weiming Zhang
机构
*
School of Electronic and Information Engineering, Anhui University(安徽大学电子与信息工程学院)
;
Anhui Province Key Laboratory of Digital Security and the CAS Key Laboratory of Electromagnetic Space Information, University of Science and Technology of China(安徽省数字安全重点实验室和中国科学院电磁空间信息重点实验室,中国科学技术大学)
;
School of Computing, National University of Singapore(新加坡国立大学计算机学院)
ViTex: Visual Texture Control for Multi-Track Symbolic Music Generation via Discrete Diffusion Models
ViTex: 通过离散扩散模型实现多轨符号音乐生成的视觉纹理控制
Xiaoyu Yi, Qi He, Gus Xia, Ziyu Wang
机构
*
School of EECS, Peking University(电子工程系,北京大学)
;
Music X Lab, MBZUAI(音乐X实验室,MBZUAI)
;
Courant Institute of Mathematical Sciences, New York University(数学科学学院,纽约大学)