Example-Based Sampling with Diffusion Models
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV、cs.GR
视觉与机器人
图像生成、文生图、图像编辑、扩散模型和可控生成。
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV、cs.GR
专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV、cs.GR
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV、cs.GR
Comments Project page: https://jryanshue.com/nfd
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV、cs.GR
Comments 17pages, 12 figures
通过温度采样和方差校正时间偏移实现扩散多样化
机构 * ETH Zurich(苏黎世联邦理工学院) ; Google(谷歌)
专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV
AI总结 研究针对扩散模型难以达到罕见模式的问题,提出方差校正时间偏移方法,并结合温度采样,无需重新训练就能提升样本多样性,在多个模型中以低成本实现质量和保真度提升,还能实现从粗到细的控制。
Comments Webpage: https://peizhuoli.github.io/diversify-diffusion
全模态扩散:基于掩码离散扩散的统一多模态理解与生成
机构 * Nanjing University(南京大学) ; Tencent Youtu Lab(腾讯云科技实验室) ; CASIA(中国科学院自动化研究所)
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV
AI总结 受离散扩散模型在多领域成功应用启发,提出全模态扩散模型,基于掩码离散扩散模型构建,统一文本、语音和图像的理解与生成,用统一模型直接捕捉离散多模态令牌联合分布,性能优于现有多模态系统。
Comments Accepted to ICML 2026. This version updates the ICML submission with an optimized model checkpoint. Project page: https://omni-diffusion.github.io
弱扩散先验仍能实现强逆问题性能
机构 * University of California, Berkeley(加州大学伯克利分校) ; Stanford University(斯坦福大学)
专题命中 扩散模型 :diffusion(title,abstract);分类 cs.CV
AI总结 研究弱扩散先验在逆问题中的鲁棒性,通过贝叶斯一致性和局部相关性分析揭示其在信息丰富测量下仍有效的原因。
Comments 37 pages, ICML 2026 spotlight. Code: https://github.com/jjia131/weak-diffusion-priors-inverse-problem, Project Page: https://jjia131.github.io/weak-diffusion-priors-inverse-problem/
对图像条件扩散模型进行微调比你想象的更容易
机构 * RWTH Aachen University(亚琛工业大学) ; Eindhoven University of Technology(埃因霍温理工大学)
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV
AI总结 本文发现对图像条件扩散模型进行微调比预期更简单,并展示了其在深度和法线估计任务中的优越性能。
Comments WACV 2025 Oral. Project page at https://vision.rwth-aachen.de/diffusion-e2e-ft
基于点图的扩散模型用于一致的新型视角合成
机构 * Noah’s Ark, Huawei Paris Research Center(Noah’s Ark,华为巴黎研究中心) ; COSYS, Gustave Eiffel University(COSYS,格拉夫-埃菲尔大学) ; LASTIG, IGN-ENSG, Gustave Eiffel University(LASTIG,IGN-ENSG,格拉夫-埃菲尔大学)
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV
AI总结 PointmapDiff通过利用点图作为条件信号,结合预训练的2D扩散模型,实现城市驾驶场景中一致的新型视角合成,并能生成高质量结果。
Comments WACV 2026. Project page: https://ntaquan0125.github.io/pointmap-conditioned-diffusion
通过校准和正则化增强扩散模型引导
机构 * UC San Diego(加州大学圣迭戈分校) ; Politecnico di Milano(米兰理工学院)
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV
AI总结 本文通过校准和正则化方法提升扩散模型引导效果,改进分类器校准和采样策略,提升图像生成质量。
Comments Accepted from NeurIPS 2025 Workshop on Structured Probabilistic Inference & Generative Modeling. Code available at https://github.com/ajavid34/guided-info-diffusion
机构 * Peking University(北京大学) ; OpenAI ; University of California, Los Angeles(加州大学洛杉矶分校) ; Carnegie Mellon University(卡内基梅隆大学) ; University of California at Merced(加州大学默塞德分校)
专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV
Comments 59 pages, 19 figures, citing 396 (up-to-date) papers, project: https://github.com/YangLing0818/Diffusion-Models-Papers-Survey-Taxonomy, accepted by ACM Computing Surveys
机构 * Shanghai Jiao Tong University(上海交通大学) ; University of Electronic Science and Technology of China(电子科技大学) ; Shandong University(山东大学)
专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
Comments 11 pages, 11 figures; Accepted by ACM MM2025; Mainly focus on feature caching for diffusion transformers acceleration
机构 * Nanjing University(南京大学) ; ByteDance Seed(字节跳动种子) ; National University of Singapore(新加坡国立大学)
专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
Comments a single-scale, single-stage, efficient, end-to-end pixel space diffusion model
机构 * The Spin Group Research Institute(Spin Group研究 institute)
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV
Comments 39 pages, 21 figures. Associated code: https://github.com/arandono/Cloud-Diffusion
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV
Comments Accepted at MICCAI 2024 International Workshop on Computational Diffusion MRI. Zijian Chen and Jueqi Wang contributed equally to this work
专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
Comments ICML 2024. Code available at https://github.com/ruocwang/dpo-diffusion
Journal ref Proceedings of the 41st International Conference on Machine Learning (ICML 2024)
专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV
Comments CVPR 2024. Project link: https://lidar-diffusion.github.io
专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
Comments CVPR 2024. Our codes and benchmarks are available at https://github.com/cure-lab/MMA-Diffusion
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV
Comments Accepted by IEEE TGRS, we first present an iterative diffusion process for cloud removal, the code is available at: https://github.com/SongYxing/IDF-CR
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV
Comments Project page: https://lukemelas.github.io/fixed-point-diffusion-models
专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
Comments ICCV 2023; Github link: https://github.com/SHI-Labs/Versatile-Diffusion
专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
Comments A prompt re-description strategy is proposed for stabilizing the diffusion model in image-to-image translation. Code and dataset page: https://mirrordiffusion.github.io/
专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV
Comments Work in progress. Project Page: https://horseee.github.io/Diffusion_DeepCache/
专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
Comments GitHub: https://github.com/SHI-Labs/Smooth-Diffusion
专题命中 扩散模型 :diffusion(title,abstract);image generation(abstract);分类 cs.CV
Comments Project page: https://sage-diffusion.github.io/
专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
Comments Accepted to NeurIPS 2023. Our project page: https://dataset-diffusion.github.io/
专题命中 扩散模型 :diffusion(title,abstract);image synthesis(abstract);分类 cs.CV
Comments A short version of this paper is accepted in the NeurIPS 2023 Workshop on Diffusion Models: https://nips.cc/virtual/2023/74894
气候情景下的大地艺术预测
专题命中 扩散模型 :diffusion(summary_cn,abstract);image synthesis(abstract)
AI总结 本文基于《螺旋防波堤》的遥感数据,构建两阶段预测流程,结合IPCC气候情景与Stable Diffusion XL模型,预测该大地艺术的图像复杂性及暴露状态,并探讨文化遗产伦理问题。
Comments 15 pages, 4 figures, 1 table
CFG-OEC: 带正交误差校正的无分类器引导
机构 * School of Electrical Engineering, Korea Advanced Institute of Science(韩国科学技术院电子工程学院)
专题命中 扩散模型 :diffusion(summary_cn,abstract);image generation(abstract)
AI总结 针对扩散模型中无分类器引导的采样规则与训练目标不匹配导致的误差,提出正交误差校正方法(CFG-OEC)通过减少条件与无条件预测误差的交互项来提升采样质量,并在Stable Diffusion上验证了FID和CLIP分数的改进。
PixRestore:基于像素扩散Transformer的统一图像修复模型
机构 * The Hong Kong Polytechnic University(香港理工大学) ; OPPO Research Institute(OPPO研究院)
专题命中 扩散模型 :diffusion(title,abstract);text-to-image(abstract);分类 cs.CV
AI总结 本文提出无VAE的像素空间DiT模型PixRestore,通过流匹配与DINO特征可靠性预测实现UIR,仅50M参数且单步推理,在效率与修复性能上优于同类模型。