Dual Caption Preference Optimization for Diffusion Models
机构 * School of Computing and Augmented Intelligence(计算与增强智能学院) ; Arizona State University(亚利桑那州立大学)
专题命中 图像生成评测 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
视觉与机器人
图像生成、文生图、图像编辑、扩散模型和可控生成。
机构 * School of Computing and Augmented Intelligence(计算与增强智能学院) ; Arizona State University(亚利桑那州立大学)
专题命中 图像生成评测 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
机构 * Google DeepMind(谷歌DeepMind) ; KAIST(韩国科学技术院) ; Google(谷歌) ; Google Research(谷歌研究) ; Georgia Institute of Technology(佐治亚理工学院)
专题命中 图像生成评测 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
Comments CVPR 2025, Project page: https://kyungmnlee.github.io/capo.github.io/
机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) ; Sun Yat-sen University(中山大学) ; CUHK MMLab(香港中文大学多模态实验室) ; Peking University(北京大学)
专题命中 图像生成评测 :image generation(title,abstract);text-to-image(abstract);分类 cs.CV
Comments 19 pages, 8 figures
专题命中 图像生成评测 :image generation(title,abstract);text-to-image(abstract);分类 cs.CV
Comments Project page: https://kitten-project.github.io/
机构 * Gaoling School of Artificial Intelligence(中关村人工智能学院) ; Renmin University of China(中国人民大学) ; Beijing Key Laboratory of Research on Large Models and Intelligent Governance(北京大模型与智能治理重点实验室) ; Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(下一代智能搜索与推荐工程研究中心,教育部) ; iN2X
专题命中 图像生成评测 :image generation(title,abstract);text-to-image(abstract);分类 cs.CV
Comments Accepted to ICCV 2025
机构 * The Chinese University of Hong Kong(香港中文大学) ; Peking University(北京大学) ; Shanghai AI Laboratory(上海人工智能实验室) ; University of Electronic Science and Technology of China(电子科技大学)
专题命中 图像生成评测 :image generation(title,abstract);diffusion(abstract);分类 cs.CV
Comments Journal Version. Code and models are released at https://github.com/ZiyuGuo99/Image-Generation-CoT
机构 * SJTU(上海交通大学) ; StepFun
专题命中 图像生成评测 :image generation(title,abstract);text-to-image(abstract);分类 cs.CV
机构 * Tulane University(Tulane大学)
专题命中 图像生成评测 :image generation(title,abstract);image synthesis(abstract);分类 cs.CV
Comments Accepted at the 34th USENIX Security Symposium (USENIX Security '25). 21 pages, plus a 6-page appendix
专题命中 图像生成评测 :text-to-image(title,abstract);diffusion(abstract);分类 cs.CV
Comments 59 pages, 10 figures
专题命中 图像生成评测 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
Comments Accepted by ICML 2025
机构 * State Key Laboratory of Complex and Critical Software Environment(复杂与关键软件环境国家重点实验室) ; School of Computer Science and Engineering(计算机科学与工程学院) ; Beihang University(北京航空航天大学)
专题命中 图像生成评测 :image synthesis(title,abstract);diffusion(abstract);分类 cs.CV
机构 * University of Macau China(澳门大学) ; East China Normal University China(华东师范大学) ; Auckland University of Technology New Zealand(技术大学)
专题命中 图像生成评测 :image generation(title,abstract);text-to-image(abstract);分类 cs.CV
机构 * University of Rochester(罗切斯特大学) ; Purdue University(普渡大学) ; NVIDIA(英伟达)
专题命中 图像生成评测 :image generation(title,abstract);text-to-image(abstract);分类 cs.CV
专题命中 图像生成评测 :diffusion(title,abstract);image generation(abstract);分类 cs.CV
Comments arXiv admin note: text overlap with arXiv:2412.14167 by other authors
专题命中 图像生成评测 :image generation(title,abstract);text-to-image(abstract);分类 cs.CV
专题命中 图像生成评测 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
Comments Accepted by CVPR2025
专题命中 图像生成评测 :image generation(title,abstract);text-to-image(abstract);分类 cs.CV
Comments Labeled data and generated image Wordnet are published at https://huggingface.co/collections/VityaVitalich/generated-image-wordnet-67d2c868ff1414ec2f8e0d3d
专题命中 图像生成评测 :inpainting(title);diffusion(abstract);image editing(abstract);分类 cs.CV
专题命中 图像生成评测 :image generation(title,abstract);image editing(abstract);分类 cs.CV
Comments Published as a conference paper at NeurIPS 2024 Datasets and Benchmarks Track https://openreview.net/forum?id=0T8xRFrScB Project page: https://gulnazaki.github.io/counterfactual-benchmark
Journal ref The Thirty-eight Conference on Neural Information Processing Systems Datasets and Benchmarks Track (2024)
专题命中 图像生成评测 :image generation(title,abstract);diffusion(abstract);分类 cs.CV
专题命中 图像生成评测 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
专题命中 图像生成评测 :image generation(title,abstract);text-to-image(abstract);分类 cs.CV
Comments Code is available at: https://github.com/google-research/google-research/tree/master/cmmd
专题命中 图像生成评测 :image generation(title,abstract);diffusion(abstract);分类 cs.CV
专题命中 图像生成评测 :image generation(title,abstract);text-to-image(abstract);分类 cs.CV
Journal ref IEEE Trans. Image Process., vol. 32, pp. 5737-5750, 2023
专题命中 图像生成评测 :text-to-image(title,abstract);diffusion(abstract);分类 cs.CV
专题命中 图像生成评测 :text-to-image(title,abstract);diffusion(abstract);分类 cs.CV
Comments Accepted by ICCV 2023
专题命中 图像生成评测 :text-to-image(title,abstract);diffusion(abstract);分类 cs.CV
ICG: 通过基于MLLM的提示和个性化偏好对齐改进封面图像生成
机构 * Huazhong University of Science and Technology(华中科技大学) ; Huawei Noah’s Ark Lab(华为诺亚实验室) ; Hong Kong Polytechnic University(香港理工大学) ; Zhejiang University(浙江大学)
专题命中 图像生成评测 :image generation(title,abstract);diffusion(abstract)
AI总结 提出ICG框架,利用多模态大语言模型和扩散模型,通过元标记提取语义特征、用户嵌入个性化对齐及多奖励学习策略,实现高质量、个性化封面图像生成。
Comments Published in Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing, pages 12268-12278, EMNLP 2025. Official version: https://doi.org/10.18653/v1/2025.emnlp-main.617
Journal ref Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing (Main Track) EMNLP 2025 12268-12278
机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) ; Beihang University(北航) ; Harbin Institute of Technology(哈尔滨工业大学) ; Chinese University of Hong Kong (Shenzhen)(香港中文大学(深圳)) ; The University of Hong Kong(香港大学)
专题命中 图像生成评测 :image generation(title);text-to-image(abstract);diffusion(abstract)
Comments Accepted at CVPR 2025
机构 * Shanghai AI Laboratory(上海人工智能实验室) ; Shanghai Innovation Institute(上海创新研究院) ; Nanjing University(南京大学) ; The Chinese University of Hong Kong(香港中文大学) ; Shanghai Jiao Tong University(上海交通大学) ; Zhejiang University of Technology(浙江工业大学)
专题命中 图像生成评测 :image generation(abstract);text-to-image(abstract);diffusion(abstract);image editing(abstract)
Comments Tech Report, 23 pages, 11 figures, 7 tables