MMTalker: Multiresolution 3D Talking Head Synthesis with Multimodal Feature Fusion
MMTalker: 多分辨率3D说话头合成与多模态特征融合
机构 * IEEE Publication Technology Group(IEEE出版技术组) ; Piscataway, NJ(新泽西州皮萨卡威)
专题命中 通用Image Fusion :multimodal fusion(abstract);分类 cs.CV
AI总结 提出一种基于多分辨率表示和多模态特征融合的3D语音驱动面部动画合成方法MMTalker,通过网格参数化、非均匀可微采样、残差图卷积网络和双交叉注意力机制,实现高唇同步精度和逼真面部表情。
Comments This article presents only the preliminary research results, which are not yet complete and lack necessary supplementary experiments. The author has decided to withdraw it to improve the research work, and will submit a more complete version in the future