arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2026-04-27 至 2026-04-27 共收录 3 信号源:cs.CV, cs.GR, cs.MM

1. 文生图 3 篇

2604.22302 2026-04-27 cs.CV 88%

Knowledge Visualization: A Benchmark and Method for Knowledge-Intensive Text-to-Image Generation

知识可视化:一个用于知识密集型文本到图像生成的基准和方法

Ran Zhao, Sheng Jin, Size Wu, Kang Liao, Zerui Gong, Zujin Guo, Yang Xiao, Wei Li

机构 * Huazhong University of Science and Technology(华中科技大学) The University of Hong Kong(香港大学) S-Lab, Nanyang Technological University(南洋理工大学S实验室)

专题命中 文生图 :image generation(title,abstract);text-to-image(title,abstract);分类 cs.CV

AI总结 本文提出KVBench基准,评估知识密集型文本到图像生成的可靠性,揭示现有模型在逻辑推理和符号精度上的不足,并提出KE-Check框架提升科学准确性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05707 2026-04-27 cs.CR 78%

Evaluating Concept Filtering Defenses against Child Sexual Abuse Material Generation by Text-to-Image Models

评估文本到图像模型中概念过滤防御对生成儿童性虐待材料的效力

Ana-Maria Cretu, Klim Kireev, Amro Abdalla, Wisdom Obinna, Raphael Meier, Sarah Adel Bargal, Elissa M. Redmiles, Carmela Troncoso

专题命中 文生图 :text-to-image(title,abstract)

AI总结 本文评估了通过过滤训练数据中的儿童图像来防止文本到图像模型生成儿童性虐待材料的有效性,发现现有检测方法无法完全去除儿童图像,且即使过滤后仍可通过少量额外查询生成儿童图像。

Comments Extended version of the paper with the name published in the Proceedings of the 47th IEEE Symposium on Security & Privacy (IEEE S&P 2026). Please cite accordingly

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.05414 2026-04-27 cs.CL cs.AI stat.ML 50%

Large Language Models Are Bad Dice Players: LLMs Struggle to Generate Random Numbers from Statistical Distributions

大语言模型是糟糕的骰子玩家:LLM在生成统计分布的随机数时表现不佳

Minda Zhao, Yilun Du, Mengyu Wang

机构 * Harvard University(哈佛大学)

专题命中 文生图 :text-to-image(abstract)

AI总结 研究发现大语言模型在生成随机数时存在显著缺陷,其采样能力随分布复杂度和采样范围增加而下降,导致下游任务出现系统性偏差。

Comments Accepted to ACL 2026 (Main Conference)

详情

展开后加载摘要…

URL PDF HTML 收藏