Synthesizing Images of Humans in Unseen Poses
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
Comments CVPR 2018
视觉与机器人
图像生成、文生图、图像编辑、扩散模型和可控生成。
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
Comments CVPR 2018
专题命中 文生图 :text-to-image(abstract);分类 cs.CV
Comments 9 pages, Part of this work is accepted by IEEE International Conference on Multimedia Expo 2018
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
Comments accepted to AAAI 2018 conference
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
Comments IEEE International Symposium on Biomedical Imaging (ISBI) 2018
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
Comments Submitted to ICLR 2018
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
Comments The project web page at http://vcl.itn.liu.se/publications/2017/TKWU17/ contains a version of the paper with high-resolution images as well as additional material
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
Comments 14 pages, ICCV 2017
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
Comments 9 pages, 2 figures
专题命中 文生图 :image synthesis(abstract);分类 cs.MM
Comments 9 pages
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
专题命中 文生图 :image synthesis(abstract);分类 cs.GR
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
专题命中 文生图 :text-to-image(abstract);分类 cs.CV
专题命中 文生图 :image synthesis(abstract);分类 cs.CV
FrED:通过领域知识图谱基础进行外部数据影响估计
机构 * National Centre for Scientific Research “Demokritos”(国家科学研究中心“德谟克利特”) ; University of Glasgow(格拉斯哥大学)
专题命中 文生图 :image synthesis(abstract)
AI总结 针对生成式AI训练数据归因问题,提出在黑盒设置下运行的概率框架,融合连续特征相似度与领域特定知识图谱,在艺术图像合成和天气预报等领域评估,证明其有效性,为外部数据影响分析提供高效可解释机制。
当文本到图像合成数据适得其反时:真实-合成混合训练中的隐私风险放大
专题命中 文生图 :text-to-image(abstract)
AI总结 研究真实-合成混合训练中合成数据会放大真实训练样本隐私泄露问题,建立理论框架,提出RSMixLeak通过成员推理攻击评估风险,有两个变体,还给出轻量级泄露倾向指标以识别高风险数据集。
UNIBROWSE:用于多模态浏览比较的数据到智能体框架
机构 * Key Laboratory of Computational Linguistics, MOE, Peking University(教育部计算语言学重点实验室,北京大学) ; School of Software and Microelectronics, Peking University(北京大学软件与微电子学院) ; School of Computer Science, Peking University(北京大学计算机科学学院)
专题命中 文生图 :text-to-image(abstract)
AI总结 研究多模态浏览比较任务,提出UNIBROWSE统一数据管道,同时生成三种模式训练数据,增强知识图谱,引入探索度度量。经此训练的35B规模智能体在多模态浏览比较基准测试中性能领先,超多个闭源智能体工作流程。
Comments 17 pages, 5 figures
一框架适用于所有:生成模型的跨模态成员推理
专题命中 文生图 :text-to-image(abstract)
AI总结 研究生成模型的跨模态成员推理问题,基于生成模型输出分布可近似训练数据分布的特性,在共享嵌入空间建模,通过似然比测试推理,贡献统一框架,性能优于现有方法。
Arena-T2I Hard:基于依赖感知清单的忠实性基准测试与改进
机构 * Arena Intelligence Inc ; UCLA(加州大学洛杉矶分校)
专题命中 文生图 :text-to-image(abstract)
AI总结 针对现有忠实性基准在复杂指令上的不足,提出310条压力测试提示集Arena-T2I Hard,并设计依赖感知清单奖励结合美学奖励,显著提升文本到图像模型的忠实性-美学权衡。
SKA的法拉第层析成像:宇宙磁学研究的新时代
专题命中 文生图 :image synthesis(abstract)
AI总结 本文综述了SKA通过法拉第旋转和法拉第测量合成技术研究宇宙磁场的能力,重点介绍了AA4阶段的配置及其在星系、星系团和宇宙网中高分辨率法拉第层析成像的应用。
Comments Published in Advancing Astrophysics with the SKA II (AASKAII), 2026 (arXiv:2606.20366). Report-no:AASKAII/Carcamo01. Advancing Astrophysics with the SKA II (AASKAII) outlines the transformative scientific advances that will be enabled by the SKA telescopes
哪个波导网络实现指定的传输轮廓?一种精确的正向构造方法
专题命中 文生图 :image synthesis(abstract)
AI总结 提出一种基于规范平移亥姆霍兹算子的周期性波导网络散射特性的可解析反演框架,通过傅里叶展开将晶格电抗映射到图架构,直接从目标传输系数确定物理结构,实现从一维角滤波到二维图像合成的精确波前构造。
Comments 12 pages, 4 figures
从提示到偏好:生成式AI增强联合分析的开源平台
专题命中 文生图 :text-to-image(abstract)
AI总结 提出一个开源、自托管的联合分析调查平台,利用生成式AI(大语言模型和文本到图像模型)创建集成刺激格式,降低研究门槛,并通过概念验证研究展示其有效性。
ImageAuditor: 针对基于图像的检索增强生成系统的成员推理攻击
专题命中 文生图 :text-to-image(abstract)
AI总结 提出首个针对基于图像的检索增强生成(IRAG)的成员推理攻击方法ImageAuditor,通过奖励引导策略优化解决跨模态检索和判别信号提取两大挑战,在仅需四次查询时AUROC超过80%。
专业知识提升AI使用:比较普通人和专业艺术家的实验证据
机构 * Center for Humans and Machines, Max Planck Institute for Human Development(人类与机器中心,马克斯·普朗克人类发展研究所) ; Tallinn University(塔林大学) ; Estonian Business School(爱沙尼亚商学院) ; Academy of Media Arts Cologne(科隆媒体艺术学院)
专题命中 文生图 :text-to-image(abstract)
AI总结 通过实验比较50位专业艺术家和普通人使用生成式AI进行图像复制和创意生成的表现,发现艺术家的专业技能迁移到AI使用中,在复制准确性和发散思维上均优于普通人,而GPT-4o在创意任务上平均略优于艺术家但未超越最佳人类。
Comments Eisenmann and Karjus contributed equally to this work and share first authorship
Journal ref International Journal of Human-Computer Interaction, 2026, pp 1-22
Talking Slide Avatars: 面向教学的开源多模态通信方法
机构 * School of Mathematics and Computer Science, Kentucky State University(肯塔基州立大学数学与计算机科学学院)
专题命中 文生图 :image synthesis(abstract)
AI总结 提出一种集成OpenVoice和Ditto-TalkingHead的开源工作流,用于创建可说话的幻灯片头像,以增强在线教学中的教师存在感和叙事连续性。
Comments 15 pages