CreatiParser: Generative Image Parsing of Raster Graphic Designs into Editable Layers
CreatiParser: 从位图图形设计生成可编辑的图层
机构 * School of Information Science and Technology, University of Science and Technology of China(科学技术大学信息科学与技术学院) ; ByteDance Intelligent Creation(字节跳动智能创作) ; School of Computer Science and Technology, Harbin Institute of Technology (Weihai)(哈尔滨工业大学(威海)计算机科学与技术学院) ; Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合国家科学中心人工智能研究院)
专题命中 文档图表理解 :vision-language model(abstract);分类 cs.CV
AI总结 本文提出CreatiParser框架,将位图图形设计分解为可编辑的文本、背景和贴纸图层,结合视觉语言模型和多分支扩散架构,提升生成质量与编辑灵活性,实验显示在Parser-40K和Crello数据集上性能优于现有方法。