arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2026-03-18 至 2026-03-18 共收录 86 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 62 篇

2502.05175 2026-03-18 cs.CV cs.GR 73%

Fillerbuster: Unified Generative Scene Completion Model for Casual Captures

Fillerbuster:用于随意捕获的统一生成场景补全模型

Ethan Weber, Norman Müller, Yash Kant, Vasu Agrawal, Michael Zollhöfer, Angjoo Kanazawa, Christian Richardt

机构 * Meta Reality Labs UC Berkeley(加州大学伯克利分校) University of Toronto(多伦多大学)

专题命中 扩散模型 :diffusion(abstract);inpainting(abstract);分类 cs.CV、cs.GR

AI总结 本文提出Fillerbuster,一种统一的生成模型,用于补全3D场景中未知区域。针对随意捕获中稀疏且缺失物体后方或上方内容的问题,模型通过多视图潜在扩散变换器处理大量输入帧,生成未知目标视图并恢复图像姿态。

Comments Project page at https://ethanweber.me/fillerbuster/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16144 2026-03-18 physics.optics 71%

High-Bandwidth 940 nm VCSEL with Zn-diffusion for Optical Communications

高带宽940 nm VCSEL采用Zn扩散用于光通信

Fu-He Hsiao, Yu-Jie Lin, Chia-Jung Tsai, Chia-Chen Li, Yun-Han Chang, Chih-Ting Chang, Jr-Hau He, Chun-Liang Lin, Yu-Heng Hong, Hao-Chung Kuo

专题命中 扩散模型 :diffusion(title)

AI总结 本文提出一种结合仿真与实验验证的设计方法,优化940 nm VCSEL结构,实现35 GHz以上调制带宽和100 Gbit/s PAM-4传输。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16392 2026-03-18 cs.CV 70%

DermaFlux: Synthetic Skin Lesion Generation with Rectified Flows for Enhanced Image Classification

DermaFlux:基于修正流的合成皮肤病变生成用于增强图像分类

Stathis Galanakis, Alexandros Koliousis, Stefanos Zafeiriou

机构 * Imperial College London, UK(伦敦帝国理工学院) Northeastern University London, UK(伦敦东北大学)

专题命中 扩散模型 :text-to-image(abstract);diffusion(abstract);分类 cs.CV

AI总结 DermaFlux通过生成临床相关的皮肤病变图像,提升二分类性能,实验表明其在小数据集上提升6%,在合成数据上提升9%。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16566 2026-03-18 cs.CV cs.GR 62%

VideoMatGen: PBR Materials through Joint Generative Modeling

VideoMatGen:基于联合生成模型的PBR材质

Jon Hasselgren, Zheng Zeng, Milos Hasan, Jacob Munkberg

机构 * NVIDIA

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV、cs.GR

AI总结 本文提出基于视频扩散变压器架构的方法,通过输入几何和文本描述联合生成物理合理的PBR材质,包含基础颜色、粗糙度、金属度和高度图等属性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.11237 2026-03-18 cs.CV cs.GR 62%

WildCap: Facial Albedo Capture in the Wild via Hybrid Inverse Rendering

WildCap:通过混合反向渲染实现野外面部铝含量捕获

Yuxuan Han, Xin Ming, Tianxiao Li, Zhuofan Shen, Qixuan Zhang, Lan Xu, Feng Xu

机构 * School of Software and BNRist, Tsinghua University(软件学院和BNRist,清华大学) ShanghaiTech University(上海科技大学) Deemos Technology(德摩斯技术)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV、cs.GR

AI总结 本文提出WildCap,一种通过混合反向渲染在野外从手机视频中实现高质量面部铝含量捕获的方法,解决了复杂光照下的铝含量分离问题。

Comments CVPR 2026. project page: https://yxuhan.github.io/WildCap/index.html; code: https://github.com/yxuhan/WildCap

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16792 2026-03-18 cs.CV cs.AI 57%

V-Co: A Closer Look at Visual Representation Alignment via Co-Denoising

V-Co:通过去噪更深入地审视视觉表示对齐

Han Lin, Xichen Pan, Zun Wang, Yue Zhang, Chu Wang, Jaemin Cho, Mohit Bansal

机构 * UNC Chapel Hill(北卡罗来纳大学教堂山分校) NYU(纽约大学) Meta AI2(人工智能研究院)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 本文通过统一的JiT框架系统研究视觉去噪,揭示了四个关键要素:双流架构、结构化无条件预测、感知漂移混合损失和特征重缩放,从而在ImageNet-256上实现更高效的生成模型。

Comments code: https://github.com/HL-hanlin/V-Co

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16211 2026-03-18 cs.CV 57%

Leveling3D: Leveling Up 3D Reconstruction with Feed-Forward 3D Gaussian Splatting and Geometry-Aware Generation

Leveling3D: 通过前馈3D高斯点划法与几何感知生成提升3D重建

Yiming Huang, Baixiang Huang, Beilei Cui, Chi Kit Ng, Long Bai, Hongliang Ren

机构 * The Chinese University of Hong Kong(香港中文大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 Leveling3D结合前馈3D重建与几何一致生成,解决传统方法在 extrapolated view 中的缺失区域问题,通过几何感知适配器提升3D重建质量,实现生成与重建的同步优化。

Comments 26 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16160 2026-03-18 cs.CV 57%

Segmentation-before-Staining Improves Structural Fidelity in Virtual IHC-to-Multiplex IF Translation

在虚拟IHC到多通道IF转换中,先分割再染色提高结构保真度

Junhyeok Lee, Han Jang, Heeseong Eum, Joon Jang, Kyu Sung Choi

机构 * Interdisciplinary Program in Cancer Biology, Seoul National University College of Medicine(癌症生物学跨学科项目,首尔国立大学医学院) Interdisciplinary Program in Bioengineering, Seoul National University(生物工程跨学科项目,首尔国立大学) Department of Biomedical Sciences, Seoul National University(生物医学科学系,首尔国立大学) Department of Radiology, Seoul National University Hospital(放射科,首尔国立大学医院) Department of Radiology, Seoul National University College of Medicine(放射科,首尔国立大学医学院) Healthcare AI Research Institute, Seoul National University Hospital(医疗人工智能研究 institute,首尔国立大学医院)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 本文提出一种无需监督的条件策略,通过预训练的核分割模型生成连续细胞概率图,结合保持局部强度统计的正则化项,提升虚拟染色的核计数保真度和感知质量。

Comments 11 pages, 2 figures, 2 tables. Submitted to MICCAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16151 2026-03-18 cs.CV 57%

EFF-Grasp: Energy-Field Flow Matching for Physics-Aware Dexterous Grasp Generation

EFF-Grasp:基于能量场流匹配的物理感知灵巧抓取生成

Yukun Zhao, Zichen Zhong, Yongshun Gong, Yilong Yin, Haoliang Sun

机构 * Shandong University, Jinan, China(山东大学,济南,中国)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 本文提出EFF-Grasp框架,通过确定性常微分方程过程和物理感知能量引导策略,实现高效稳定的物理感知灵巧抓取生成,优于扩散模型基线。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16122 2026-03-18 cs.CV 57%

Out-of-Distribution Object Detection in Street Scenes via Synthetic Outlier Exposure and Transfer Learning

通过合成异常暴露和迁移学习进行街道场景中的分布外物体检测

Sadia Ilyas, Annika Mütze, Klaus Friedrichs, Thomas Kurbiel, Matthias Rottmann

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 本文提出SynOE-OD框架,利用生成模型和开放词汇物体检测器生成语义丰富的异常数据,提升检测模型对分布外物体的鲁棒性,达到当前最佳的平均精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16099 2026-03-18 cs.CV 57%

OneWorld: Taming Scene Generation with 3D Unified Representation Autoencoder

OneWorld: 通过3D统一表示自编码器驯服场景生成

Sensen Gao, Zhaoqing Wang, Qihang Cao, Dongdong Yu, Changhu Wang, Tongliang Liu, Mingming Gong, Jiawang Bian

机构 * Mohamed bin Zayed University of Artificial Intelligence(莫扎德·本·泽亚德人工智能大学) AISphere Shanghai Jiao Tong University(上海交通大学) University of Melbourne(墨尔本大学) Nanyang Technological University(南洋理工大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 OneWorld通过3D统一表示自编码器直接在3D空间中进行扩散,解决跨视角一致性和几何一致性问题,实验表明其生成的3D场景质量优于现有2D方法。

Comments Code: https://github.com/SensenGao/OneWorld

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.15975 2026-03-18 cs.CV 57%

UMO: Unified In-Context Learning Unlocks Motion Foundation Model Priors

UMO:统一上下文学习解锁运动基础模型先验

Xiaoyan Cong, Zekun Li, Zhiyang Dou, Hongyu Li, Omid Taheri, Chuan Guo, Abhay Mittal, Sizhe An, Taku Komura, Wojciech Matusik, Michael J. Black, Srinath Sridhar

机构 * Brown University(布朗大学) Massachusetts Institute of Technology(麻省理工学院) Meta Reality Lab(Meta现实实验室) Max-Planck Institute for Intelligent Systems(智能系统马克斯·普朗克研究所) University of Hong Kong(香港大学)

专题命中 扩散模型 :inpainting(abstract);分类 cs.CV

AI总结 UMO通过统一框架解锁运动基础模型先验,支持多种跨模态和上下文生成任务,提升文本到运动合成性能。

Comments Project Page: https://oliver-cong02.github.io/UMO.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.13858 2026-03-18 cs.CV 57%

Learning through Creation: A Hash-Free Framework for On-the-Fly Category Discovery

通过创造学习:一种无哈希的实时类别发现框架

Bohan Zhang, Weidong Tang, Zhixiang Chi, Yi Jin, Zhenbo Li, Yang Wang, Yanan Wu

机构 * College of Information and Electrical Engineering(信息与电气工程学院) China Agricultural University(中国农业大学) Department of Electrical and Computer Engineering(电气与计算机工程系) University of Toronto(多伦多大学) School of Computer and Information Technology(计算机与信息科技学院) Beijing Jiaotong University(北京交通大学) Department of Computer Science and Software Engineering(计算机科学与软件工程系) Concordia University(Concordia大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 本文提出LTC框架,通过在线伪未知生成器实现实时类别发现,提升模型对未知区域的识别能力,在七个基准测试中取得显著提升。

Comments Accepted to CVPR 2026 Findings. Code available at https://github.com/brandinzhang/LTC

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.13644 2026-03-18 cs.RO cs.AI cs.CV 57%

World Models for Learning Dexterous Hand-Object Interactions from Human Videos

为从人类视频学习灵巧手-物体交互构建世界模型

Raktim Gautam Goswami, Amir Bar, David Fan, Tsung-Yen Yang, Gaoyue Zhou, Prashanth Krishnamurthy, Michael Rabbat, Farshad Khorrami, Yann LeCun

机构 * FAIR at Meta(Meta 的 FAIR 部门) New York University(纽约大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 本文提出DexWM模型,通过手部关键点提取实现对精细手部动作的建模,提升未来状态预测和零样本迁移能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04761 2026-03-18 cs.CV 57%

Order Matters: 3D Shape Generation from Sequential VR Sketches

顺序至关重要:从连续VR草图生成3D形状

Yizi Chen, Sidi Wu, Tianyi Xiao, Nina Wiedemann, Loic Landrieu

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 本文提出VRSketch2Shape框架,通过顺序感知的草图编码器和扩散生成器,提升从VR草图生成3D形状的几何精度与泛化能力。

Comments Accepted at CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.21659 2026-03-18 math.AP 50%

The Kolmogorov forward equation for a distributed model of regime-switching diffusions

具有切换区域扩散过程的科尔莫戈罗夫前向方程

Alexander S. Bratus, Olga S. Rozanova

专题命中 扩散模型 :diffusion(abstract)

AI总结 本文提出了一种描述连续分布状态密度的积分微分方程,展示了求解柯西问题的构造算法,并讨论了离散隐藏状态模型如何近似为连续分布状态模型。

Comments 16 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16401 2026-03-18 astro-ph.SR physics.plasm-ph physics.space-ph 50%

Quantifying the Effects of Parameters in Widespread SEP Events with EPREM

利用EPREM量化广泛SEP事件中参数的影响

Matthew A. Young, Bala Poduval

专题命中 扩散模型 :diffusion(abstract)

AI总结 研究通过EPREM模型分析不同参数对广泛SEP事件的影响,发现SEP通量随时间与能量变化复杂,受扩散、自由程和激波轮廓等参数影响显著。

Comments 32 pages, 20 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20893 2026-03-18 astro-ph.HE 50%

The role of magnetic fields in shaping $γ$-ray emission from the Fermi bubbles

磁场在塑造费米气泡γ射线发射中的作用

Olivier Tourmente, Donna Rodgers-Lee, Andrew M. Taylor

专题命中 扩散模型 :diffusion(abstract)

AI总结 研究磁场对银河中心喷流中宇宙射线扩散的影响,揭示其在费米气泡γ射线发射形成中的作用。

Journal ref Mon Not R Astron Soc (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16602 2026-03-18 cond-mat.mtrl-sci 50%

Fully anharmonic calculations of the free energy of migration of point defects in UO2 and PuO2

完全非谐计算铀氧化物和钚氧化物中点缺陷迁移自由能

Dillon G. Frost, Johann Bouchet, Mihai-Cosmin Marinica, Clovis Lapointe, Jean-Bernard Maillet, Luca Messina

专题命中 扩散模型 :diffusion(abstract)

AI总结 本文通过PAFI方法研究了非谐效应在UO2和PuO2中点缺陷迁移中的作用,发现非谐贡献显著影响迁移熵和扩散系数,且非谐效应在预测核燃料和其他材料扩散时至关重要。

Comments submitted to Physical Review Materials

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16542 2026-03-18 cs.RO 50%

Conservative Offline Robot Policy Learning via Posterior-Transition Reweighting

通过后过渡重加权实现保守的离线机器人策略学习

Wanpeng Zhang, Hao Luo, Sipeng Zheng, Yicheng Feng, Haiweng Xu, Ziheng Xi, Chaoyi Xu, Haoqi Yuan, Zongqing Lu

机构 * Peking University(北京大学) Tsinghua University(清华大学) BeingBeyond

专题命中 扩散模型 :diffusion(abstract)

AI总结 本文提出PTR方法,通过后验过渡重加权实现保守的离线机器人策略学习,根据样本后作用后果的可归因性分配信用,提升对异质机器人数据的适应性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16290 2026-03-18 math.NA cs.NA 50%

Jin-Xin relaxation as a shock-capturing method for high-order DG/FR schemes

金-辛松弛作为高阶DG/FR方案中的激波捕捉方法

Marco Artiano, Arpit Babbar, Michael Schlottke-Lakemper, Gregor Gassner, Hendrik Ranocha

专题命中 扩散模型 :diffusion(abstract)

AI总结 本文提出利用金-辛松弛方法作为高阶DG/FR方案的激波捕捉方法,通过选择松弛参数ε来控制数值耗散,采用紧凑Runge-Kutta FR方法处理刚性源项,验证了该方法在Burgers方程和可压缩欧拉方程中的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16216 2026-03-18 cs.CE cs.AI cs.ET 50%

Generative AI for Quantum Circuits and Quantum Code: A Technical Review and Taxonomy

生成式AI用于量子电路和量子代码:技术综述与分类

Juhani Merilehto

机构 * University of Vaasa(瓦萨大学) University of Turku(图尔库大学)

专题命中 扩散模型 :diffusion(abstract)

AI总结 综述13个生成系统和5个支持数据集,探讨量子电路和代码生成的分类与训练方法,发现生成系统在语法和语义层面有所覆盖,但缺乏对量子硬件的端到端评估。

Comments 20 pages, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16173 2026-03-18 math.AP 50%

Long time dynamics and anomalous dissipation of energy in viscous forced active scalar equations

黏性强迫主动标量方程中的长时间动力学与异常耗散

Susan Friedlander, Anthony Suen

专题命中 扩散模型 :diffusion(abstract)

AI总结 研究黏性强迫主动标量方程中的长时间动力学与异常耗散,通过分数阶拉普拉斯框架分析扩散参数对能量耗散的影响,证明长时平均解无异常耗散并存在唯一全局吸引子。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16149 2026-03-18 cond-mat.soft 50%

Mechanical anisotropy of 3D-printed digital materials at large strains

3D打印数字材料在大应变下的机械各向异性

Seunghwan Lee, Gisoo Lee, Seounghee Yun, Sumin Lee, Jeonyoon Lee, Hansohl Cho

专题命中 扩散模型 :diffusion(abstract)

AI总结 研究通过实验和建模揭示3D打印数字材料在大应变下的机械各向异性机制,探讨微结构异质性对材料性能的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.16121 2026-03-18 cond-mat.stat-mech cond-mat.dis-nn 50%

Conditional Ergodicity and Universal Fluctuations in Weak Ergodicity Breaking

条件自洽性与弱自洽性破坏中的普遍波动

Dan Shafir, Stanislav Burov

专题命中 扩散模型 :diffusion(abstract)

AI总结 研究揭示在弱自洽性破坏系统中,通过内部时钟条件自洽可使时间平均可观测量自洽,经均值缩放后,输运系数服从Mittag-Leffler分布。

Comments 10 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.15890 2026-03-18 physics.flu-dyn 50%

Mixing with viscoelastic waves at low Reynolds numbers

在低雷诺数下利用粘弹性波实现混合

Enrico Turato, Christelle N. Prinz, Jason P. Beech, Jonas. O Tegenfeldt

专题命中 扩散模型 :diffusion(abstract)

AI总结 本文研究了在低雷诺数条件下,通过粘弹性湍流提高微流控通道中分子和聚合物的混合效率,展示了粘弹性波动对反应速率和大分子混合的增强作用,并探讨了优化策略。

Comments 7 pages, 5 figures, submitted to Lab on a Chip

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.15217 2026-03-18 q-bio.PE math.DS q-bio.QM 50%

A multiscale discrete-to-continuum framework for structured population models

一种多尺度离散到连续框架用于结构化种群模型

Eleonora Agostinelli, Keith L. Chambers, Helen M. Byrne, Mohit P. Dalwadi

专题命中 扩散模型 :diffusion(abstract)

AI总结 本文提出一种多尺度离散到连续框架,用于系统推导结构化种群模型的连续近似,通过多重尺度方法和匹配渐近展开,识别结构空间中的合适区域并推导偏微分方程,验证了该方法在动脉粥样硬化早期脂质结构模型中的应用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.14868 2026-03-18 math.OC cs.SY eess.SY 50%

Free Final Time Adaptive Mesh Covariance Steering via Sequential Convex Programming

自由终时间自适应网格协方差引导 via 顺序凸规划

Joshua Pilipovsky

专题命中 扩散模型 :diffusion(abstract)

AI总结 本文提出了一种顺序凸规划框架,用于非线性随机微分方程的自由终时间协方差引导,通过时间归一化和时间扩张变量实现自适应离散化网格,优化控制策略和时间网格。

Comments Full-length version of paper submitted to L-CSS

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02517 2026-03-18 astro-ph.HE astro-ph.GA 50%

Evidence for a Delayed UV Counterpart to X-ray Quasi-periodic Eruptions in Ansky

Ansky中X射线准周期喷发的延迟紫外 counterpart 的证据

Hengxiao Guo, Zhen Yan, Ya-Ping Li, Joheen Chakraborty, Paula Sánchez-Sáez, Lorena Hernández-García, Wenda Zhang, Jingbo Sun, Shuang-liang Li, Hongping Deng, Wenwen Zuo, Hiromichi Tagawa, Xin Pan, Minghao Zhang, Patricia Arévalo, Paulina Lira, Chichuan Jin, Minfeng Gu

专题命中 扩散模型 :diffusion(abstract)

AI总结 研究发现Ansky中首次检测到与X射线准周期喷发信号时间耦合的紫外响应,紫外发射显示五次周期内相干调制,延迟0.96天,交叉相关系数达0.6,探讨了可能的扩散时间或光行进时间机制。

Comments 12 pages, 5 figures, 1 table; accepted for publication in ApJL

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00698 2026-03-18 cs.LG stat.ML 50%

Flow Matching for Tabular Data Synthesis

表格数据合成的流匹配

Bahrul Ilmi Nasution, Floor Eijkelboom, Mark Elliot, Richard Allmendinger, Christian A. Naesseth

机构 * Department of Social Statistics(社会统计系) The University of Manchester(曼彻斯特大学) Alliance Manchester Business School(曼彻斯特商业联盟学院) Amsterdam Machine Learning Lab(阿姆斯特丹机器学习实验室) University of Amsterdam(阿姆斯特丹大学)

专题命中 扩散模型 :diffusion(abstract)

AI总结 本文探讨了流匹配在表格数据合成中的应用,比较了流匹配与扩散模型的性能,发现流匹配在计算效率和隐私保护方面更具优势。

Comments Published at TMLR

详情

展开后加载摘要…

URL PDF HTML 收藏