Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation
重新思考位置嵌入作为多参考和多镜头视频生成的上下文控制器
Binyuan Huang, Yuning Lu, Weinan Jia, Hualiang Wang, Mu Liu, Daiqing Yang
机构
*
Wuhan University(武汉大学)
;
University of Science and Technology of China(中国科学技术大学)
;
Hong Kong University of Science and Technology(香港科技大学)
;
Tsinghua University(清华大学)
HVG-3D: Bridging Real and Simulation Domains for 3D-Conditional Hand-Object Interaction Video Synthesis
HVG-3D: 联接真实与仿真领域用于3D条件手-物体交互视频合成
Mingjin Chen, Junhao Chen, Zhaoxin Fan, Yujian Lee, Zichen Dang, Lili Wang, Yawen Cui, Lap-Pui Chau, Yi Wang
机构
*
Dept. of EEE, The Hong Kong Polytechnic University(香港理工大学电机及电子工程学系)
;
Beijing Advanced Innovation Center for Future Blockchain and Privacy Computing, School of Artificial Intelligence, Beihang University(北京航空航天大学人工智能学院未来区块链与隐私计算北京高精尖创新中心)
;
Tsinghua University(清华大学)
;
Beijing Normal-Hong Kong Baptist University(北京师范大学-香港浸会大学联合国际学院)
;
State Key Laboratory of Virtual Reality Technology and Systems, School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院虚拟现实技术与系统国家重点实验室)