arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

2026-03-09 至 2026-03-09 共收录 78 信号源:cs.CV, cs.GR, cs.MM

1. 扩散模型 52 篇

2603.06113 2026-03-09 cs.LG physics.chem-ph 78%

Latent Diffusion-Based 3D Molecular Recovery from Vibrational Spectra

基于潜在扩散模型的振动光谱中3D分子重构

Wenjin Wu, Aleš Leonardis, Linjiang Chen, Jianbo Jiao

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出IR-GeoDiff模型,通过整合光谱信息恢复红外光谱对应的三维分子几何结构,展示了其在分子重构中的有效性。

Comments 27 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05567 2026-03-09 cs.LG 78%

FuseDiff: Symmetry-Preserving Joint Diffusion for Dual-Target Structure-Based Drug Design

FuseDiff: 保留对称性的双目标结构导向药物设计联合扩散

Jianliang Wu, Anjie Qiao, Zhen Wang, Zhewei Wei, Sheng Chen

机构 * Sun Yat-sen University(中山大学) Renmin University of China(中国人民大学) Tsinghua University(清华大学)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 FuseDiff通过端到端扩散模型联合生成双目标结合姿态,保留对称性并实现高精度药物设计。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06261 2026-03-09 cs.RO 78%

Safe Model Predictive Diffusion with Shielding

安全模型预测扩散与防护

Taekyung Kim, Keyvan Majd, Hideki Okamoto, Bardh Hoxha, Dimitra Panagou, Georgios Fainekos

机构 * University of Michigan(密歇根大学) Toyota Motor North America, Research & Development(丰田北美公司,研发部门)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 Safe MPD是一种无需训练的扩散规划器,通过结合基于模型的扩散框架和安全防护,生成安全且动力学可行的轨迹,显著提升规划效率和安全性。

Comments 2026 IEEE International Conference on Robotics and Automation (ICRA). Project page: https://www.taekyung.me/safe-mpd

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05664 2026-03-09 cs.LG 78%

KLASS: KL-Guided Fast Inference in Masked Diffusion Models

KLASS: 在掩码扩散模型中基于KL的快速推理

Seo Hyun Kim, Sunwoo Hong, Hojung Jung, Youngrok Park, Se-Young Yun

机构 * KAIST AI(韩国科学技术院人工智能学院)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 KLASS通过利用令牌级KL散度实现快速稳定生成,显著提升扩散模型推理速度并保持高质量输出。

Comments NeurIPS 2025 Spotlight. Code: https://github.com/shkim0116/KLASS

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.23405 2026-03-09 cs.LG 78%

Planner Aware Path Learning in Diffusion Language Models Training

扩散语言模型训练中的规划感知路径学习

Fred Zhangzhi Peng, Zachary Bezemek, Jarrid Rector-Brooks, Shuibai Zhang, Anru R. Zhang, Michael Bronstein, Alexander Tong, Avishek Joey Bose

机构 * Duke University(杜克大学) Mila Université de Montréal(蒙特利尔大学) California Institute of Technology(加州理工学院) University of Wisconsin–Madison(威斯康星大学麦迪逊分校) University of Oxford(牛津大学) AITHYRA Imperial College London(伦敦帝国学院)

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 本文提出PAPL,一种通过规划感知路径学习提升扩散语言模型训练与推理一致性的方法,实现跨领域性能提升。

Comments Camera ready version for ICLR2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.16861 2026-03-09 q-bio.BM cs.LG physics.bio-ph 78%

BInD: Bond and Interaction-generating Diffusion Model for Multi-objective Structure-based Drug Design

BInD:基于多目标结构的药物设计的结合与相互作用生成扩散模型

Joongwon Lee, Wonho Zhung, Jisu Seo, Woo Youn Kim

专题命中 扩散模型 :diffusion(title,abstract)

AI总结 BInD通过结合知识引导的扩散模型,实现多目标结构导向药物设计,平衡分子相互作用、性质和几何,提升药物结合特异性。

Comments Published in Advanced Science 12(35), e02702 (2025)

Journal ref Advanced Science 12(35), e02702 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06408 2026-03-09 cs.CV cs.AI cs.GR 62%

Physical Simulator In-the-Loop Video Generation

物理模拟器闭环视频生成

Lin Geng Foo, Mark He Huang, Alexandros Lattas, Stylianos Moschoglou, Thabo Beeler, Christian Theobalt

机构 * Max Planck Institute for Informatics(马克斯·普朗克信息研究所) Singapore University of Technology and Design(新加坡科技设计大学) A*STAR(新加坡科技研究局) Google(谷歌) Saarbrücken Research Center for Visual Computing, Interaction and Artificial Intelligence(斯图加特视觉计算、交互与人工智能研究中心)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV、cs.GR

AI总结 PSIVG通过整合物理模拟器与扩散模型,生成符合物理规律的视频,提升生成视频的真实性和一致性。

Comments Accepted to CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06507 2026-03-09 cs.CV 57%

Self-Supervised Flow Matching for Scalable Multi-Modal Synthesis

自监督流匹配用于可扩展的多模态合成

Hila Chefer, Patrick Esser, Dominik Lorenz, Dustin Podell, Vikash Raja, Vinh Tong, Antonio Torralba, Robin Rombach

机构 * Black Forest Labs(黑森林实验室) MIT(麻省理工学院)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 Self-Flow通过自监督流匹配方法,在无需外部监督的情况下提升多模态生成的性能和扩展性。

Comments project webpage: https://bfl.ai/research/self-flow

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06357 2026-03-09 cs.CV 57%

LATO: 3D Mesh Flow Matching with Structured TOpology Preserving LAtents

LATO:基于结构拓扑保持的3D网格流匹配

Tianhao Zhao, Youjia Zhang, Hang Long, Jinshen Zhang, Wenbing Li, Yang Yang, Gongbo Zhang, Jozef Hladký, Matthias Nießner, Wei Yang

机构 * Huazhong University of Science and Technology(华中科技大学) Technical University of Munich(慕尼黑技术大学) Independent Researcher(独立研究者) Peking University(北京大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 LATO通过拓扑保持的潜在表示实现高效3D网格生成,结合流匹配技术生成复杂几何且拓扑结构良好的网格。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06311 2026-03-09 cs.CV 57%

Latent Transfer Attack: Adversarial Examples via Generative Latent Spaces

潜在转移攻击:通过生成性潜在空间进行对抗示例

Eitan Shaar, Ariel Shaulov, Yalcin Tur, Gal Chechik, Ravid Shwartz-Ziv

机构 * Independent Researcher(独立研究者) Tel-Aviv University(特拉维夫大学) Stanford University(斯坦福大学) Bar Ilan University(巴伊兰大学) NVIDIA Research(NVIDIA研究) New York University(纽约大学)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 LTA通过生成性潜在空间优化对抗扰动,提升跨架构转移效果,生成低频、连贯的扰动,增强鲁棒性评估与生成先验的结合。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06147 2026-03-09 cs.CV 57%

Longitudinal NSCLC Treatment Progression via Multimodal Generative Models

多模态生成模型用于非小细胞肺癌治疗进展的纵向预测

Massimiliano Mantegna, Elena Mulero Ayllón, Alice Natalina Caragliano, Francesco Di Feola, Claudia Tacconi, Michele Fiore, Edy Ippolito, Carlo Greco, Sara Ramella, Philippe C. Cattin, Paolo Soda, Matteo Tortora, Valerio Guarrasi

机构 * Unit of Artificial Intelligence and Computer Systems, Department of Engineering, Università Campus Bio-Medico di Roma, Italy(人工智能与计算机系统单位,工程系,罗马大学生物医学学院) Multi-Specialist Clinical Institute for Orthopaedic Trauma Care (COT), Messina, Italy(骨科创伤护理多学科临床研究所(COT),意大利Messina) Department of Diagnostics and Intervention, Radiation Physics, Biomedical Engineering, Umeå University, Sweden(诊断与介入系,放射物理,生物医学工程,乌梅大学,瑞典) Operative Research Unit of Radiation Oncology, Fondazione Policlinico Universitario Campus Bio-Medico, Rome, Italy(放射肿瘤手术研究单位,大学生物医学学院基金会,罗马,意大利) Research Unit of Radiation Oncology, Department of Medicine and Surgery, Università Campus Bio-Medico di Roma, Italy(放射肿瘤研究单位,医学与外科系,罗马大学生物医学学院,意大利) Department of Biomedical Engineering, University of Basel, Allschwil, Switzerland(生物医学工程系,巴塞尔大学,瑞士Allschwil) Department of Naval, Electrical, Electronics and Telecommunications Engineering, University of Genoa, Italy(海军、电气、电子与电信工程系,热那亚大学,意大利)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 本研究提出虚拟治疗框架,利用多模态生成模型预测NSCLC治疗进展,验证扩散模型在生成稳定肿瘤演变轨迹方面的优势。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05936 2026-03-09 cs.CV 57%

OD-RASE: Ontology-Driven Risk Assessment and Safety Enhancement for Autonomous Driving

基于本体的风险评估与安全增强:面向自动驾驶

Kota Shimomura, Masaki Nambata, Atsuya Ishikawa, Ryota Mimura, Takayuki Kawabuchi, Takayoshi Yamashita, Koki Inoue

机构 * Chubu University(楚鸟大学) Elith Inc.(Elith公司) Honda R&D Co., Ltd.(本田研发公司)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 OD-RASE通过本体驱动的数据过滤和视觉语言模型生成,提升自动驾驶系统对事故诱因道路结构的预测和改进能力,增强交通环境的安全性。

Comments Accepted ICCV2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17353 2026-03-09 eess.IV cs.CV physics.optics 57%

Learning Latent Transmission and Glare Maps for Lens Veiling Glare Removal

学习潜在的传输和眩光地图以去除镜头遮蔽眩光

Xiaolong Qian, Qi Jiang, Lei Sun, Zongxi Yu, Kailun Yang, Peixuan Wu, Jiacheng Zhou, Yao Gao, Yaoguang Ma, Ming-Hsuan Yang, Kaiwei Wang

机构 * Zhejiang University(浙江大学) Hunan University(湖南大学) University of California, Merced(加州大学默塞德分校) Google DeepMind(谷歌DeepMind)

专题命中 扩散模型 :diffusion(abstract);分类 cs.CV

AI总结 本文提出VeilGen和DeVeiler,通过学习潜在传输和眩光地图来有效去除遮蔽眩光,提升光学系统成像质量。

Comments Accepted to CVPR 2026. All code and datasets will be publicly released at https://github.com/XiaolongQian/DeVeiler

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.06496 2026-03-09 physics.med-ph 50%

Rotation-invariant graph message passing enables acquisition protocol generalisation in learning-based brain microstructure estimation

旋转不变图消息传递实现基于学习的脑微结构估计中的获取协议泛化

Leevi Kerkelä, Hui Zhang

专题命中 扩散模型 :diffusion(abstract)

AI总结 本文提出一种旋转不变图消息传递网络,通过模拟数据训练实现微结构估计的协议泛化,无需重新训练即可适应未知协议。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12746 2026-03-09 cond-mat.mtrl-sci 50%

Oxygen-vacancy-induced Raman softening in the catalyst Fe$_2$(MoO$_4$)$_3$

氧空位诱导的催化剂Fe$_2$(MoO$_4$)$_3$的拉曼变软

Young-Joon Song, Roser Valentí

专题命中 扩散模型 :diffusion(abstract)

AI总结 本研究通过DFT计算揭示Fe$_2$(MoO$_4$)$_3$中氧空位诱导的拉曼变软机制,表明氧振动主导了拉曼强度变化。

Comments 7 pages, 2 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05983 2026-03-09 math.AP math.CA math.CV 50%

The Planar Coleman--Gurtin model with Beltrami conductivity

平面Coleman-Gurtin模型与Beltrami电导率

Francesco Di Plinio

专题命中 扩散模型 :diffusion(abstract)

AI总结 本文研究了平面Coleman-Gurtin模型中Beltrami电导率对正则化和吸引子构造的影响。

Comments 30 pages; submitted

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05804 2026-03-09 cs.RO 50%

CDF-Glove: A Cable-Driven Force Feedback Glove for Dexterous Teleoperation

CDF-Glove: 一种用于灵巧遥控操作的电缆驱动力反馈手套

Huayue Liang, Ruochong Li, Yaodong Yang, Long Zeng, Yuanpei Chen, Xueqian Wang

机构 * Center for Artificial Intelligence and Robotics, Shenzhen International Graduate School, Tsinghua University(人工智能与机器人中心,深圳国际研究生院,清华大学) PKU-Psibot Joint Lab(北京大学-PsiBot联合实验室) Department of Advanced Manufacturing, Shenzhen International Graduate School, Tsinghua University(先进制造系,深圳国际研究生院,清华大学)

专题命中 扩散模型 :diffusion(abstract)

AI总结 CDF-Glove是一种低成本的电缆驱动力反馈手套,通过实时力反馈提升灵巧遥控操作的演示质量和任务成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05796 2026-03-09 physics.med-ph 50%

CBCT-Based Synthetic CT Generation Using Conditional Flow Matching Model

基于CBCT的合成CT生成使用条件流匹配模型

Junbo Peng, Huiqiao Xie, Tonghe Wang, Xiangyang Tang, Xiaofeng Yang

专题命中 扩散模型 :diffusion(abstract)

AI总结 本研究提出一种条件流匹配模型,用于从CBCT生成高质量合成CT,提高HU精度并减少伪影,从而提升放射治疗中的器官分割和剂量计算可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05745 2026-03-09 nlin.PS physics.bio-ph 50%

Laws of mutual spiral wave interaction in excitable media

可兴奋介质中相互螺旋波相互作用的定律

Tim De Coster, Arstanbek Okenov, Debora Hoogendijk, Arman Nobacht, Mathilde Rivaud, Antoine de Vries, Daniël Pijnappels, Vivi Rottschäfer, Hans Dierckx

专题命中 扩散模型 :diffusion(abstract)

AI总结 研究提出可兴奋介质中螺旋波相互作用的定律,通过边界积分确定螺旋波漂移速度与力的关系,应用于心室颤动分析。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05670 2026-03-09 cs.RO 50%

TransMASK: Masked State Representation through Learned Transformation

TransMASK: 通过学习变换实现状态表示的掩码

Sagar Parekh, Preston Culbertson, Dylan P. Losey

机构 * Mechanical Engineering Department, Virginia Tech(弗吉尼亚理工大学机械工程系) Computer Science Department, Cornell University(康奈尔大学计算机科学系)

专题命中 扩散模型 :diffusion(abstract)

AI总结 TransMASK通过学习变换实现状态表示的掩码,提升机器人在不同环境中的泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05536 2026-03-09 physics.bio-ph 50%

Programmable ultrasonic fields enhance intracellular delivery in cell clusters

可编程超声场增强细胞簇内的细胞内递送

Subhas Nandy, Monica Manohar, Ashis K Sen

专题命中 扩散模型 :diffusion(abstract)

AI总结 PAST通过可编程超声场实现细胞内生物分子递送,具有高通量和非侵入性特点,适用于药物筛选和细胞膜研究。

Comments 35 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05390 2026-03-09 cond-mat.stat-mech 50%

Extreme Values of Infinite-Measure Processes

无穷测度过程的极值

Talia Baravi, Eli Barkai

专题命中 扩散模型 :diffusion(abstract)

AI总结 本文研究了无穷测度过程的极值统计特性,揭示了其与返回指数和无限不变测度的关系,并通过多个实例展示了极值测量对无限密度结构的推断作用。

Comments 13 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03929 2026-03-09 stat.ML cs.LG 50%

Self-Speculative Masked Diffusions

自推测掩码扩散

Andrew Campbell, Valentin De Bortoli, Jiaxin Shi, Arnaud Doucet

机构 * Google DeepMind(谷歌DeepMind)

专题命中 扩散模型 :diffusion(abstract)

AI总结 自推测掩码扩散通过非因子化预测减少计算负担,实现文本和蛋白质序列生成的高效样本生成。

Comments 32 pages, 7 figures, 4 tables

Journal ref ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10638 2026-03-09 astro-ph.HE 50%

Radiation GRMHD Models of Accretion onto Stellar-Mass Black Holes: II. Super-Eddington Accretion

辐射GRMHD模型:恒星级黑洞吸积II. 超埃德顿吸积

Lizhong Zhang, James M. Stone, Christopher J. White, Shane W. Davis, Yan-Fei Jiang, Patrick D. Mullen

专题命中 扩散模型 :diffusion(abstract)

AI总结 本研究通过GRMHD模型分析超埃德顿吸积流,揭示辐射压力支撑厚盘、辐射驱动喷流及喷流对观测特征的影响,适用于多种天文系统。

Comments 38 pages, 25 figures, 3 tables, submitted to ApJ

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.21030 2026-03-09 hep-ex hep-ph 50%

System size and event shape dependence of particle-identified balance functions in proton-proton collisions at $\sqrt{s} = 13$ TeV using PYTHIA 8 and EPOS models

在13 TeV质子-质子碰撞中,使用PYTHIA 8和EPOS模型研究系统大小和事件形状对粒子识别平衡函数的依赖性

Subash Chandra Behera, Arvind Khuntia

专题命中 扩散模型 :diffusion(abstract)

AI总结 本研究通过PYTHIA 8和EPOS模型,探讨了13 TeV质子-质子碰撞中粒子识别平衡函数对系统大小和事件形状的依赖性,揭示了集体动力学和强子化过程的特征。

Comments 14 pages, 8 figures

Journal ref Phys. Rev. C 113, 035201 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.17721 2026-03-09 cs.LG cs.AI cs.MA 50%

Aligning Compound AI Systems via System-level DPO

通过系统级DPO对复合AI系统进行对齐

Xiangwen Wang, Yibo Jacky Zhang, Zhoujie Ding, Katherine Tsai, Haolun Wu, Sanmi Koyejo

机构 * Stanford University(斯坦福大学) University of Illinois Urbana Champaign(伊利诺伊大学厄巴纳-香槟分校) Mila Quebec AI Institute(魁北克AI研究院)

专题命中 扩散模型 :diffusion(abstract)

AI总结 本文提出SysDPO框架,通过系统级DPO实现复合AI系统的联合对齐,解决了组件间非可微分交互和系统偏好转换的问题。

Comments NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13406 2026-03-09 cs.RO cs.AI cs.SY eess.SY 50%

Generative Predictive Control: Flow Matching Policies for Dynamic and Difficult-to-Demonstrate Tasks

生成预测控制:用于动态和难以示范任务的流匹配策略

Vince Kurtz, Joel W. Burdick

专题命中 扩散模型 :diffusion(abstract)

AI总结 本文提出生成预测控制方法,用于解决动态且难以示范任务中的快速动态问题,通过流匹配策略实现高效反馈和高频率控制。

Comments ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.04293 2026-03-09 cond-mat.stat-mech cond-mat.dis-nn cond-mat.soft 50%

Time-dependent dynamics in the confined lattice Lorentz gas

受限晶格洛伦兹气体中的时间依赖动力学

A. Squarcini, A. Tinti, P. Illien, O. Bénichou, T. Franosch

专题命中 扩散模型 :diffusion(abstract)

AI总结 研究受限晶格洛伦兹气体中时间依赖动力学,分析驱动力与约束对扩散系数的影响,揭示约束对非解析行为的改变及超扩散现象的持续性。

Comments 26 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.10430 2026-03-09 astro-ph.EP 50%

Dust-gas dynamics driven by the streaming instability with various pressure gradients

由不同压力梯度驱动的尘-气体动力学与流体不稳定性

Stanley A. Baronett, Chao-Chin Yang, Zhaohuan Zhu

专题命中 扩散模型 :diffusion(abstract)

AI总结 研究探讨了不同压力梯度对尘-气体流体不稳定性的影响,发现梯度变化影响尘埃颗粒分布和涡旋结构,为原行星盘模型提供新见解。

Comments Accepted by MNRAS. 22 pages, 15 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 可控生成 8 篇

2603.05769 2026-03-09 cs.CV 89%

Layer-wise Instance Binding for Regional and Occlusion Control in Text-to-Image Diffusion Transformers

逐层实例绑定用于文本到图像扩散变换器中的区域和遮挡控制

Ruidong Chen, Yancheng Bai, Xuanpu Zhang, Jianhao Zeng, Lanjun Wang, Dan Song, Lei Sun, Xiangxiang Chu, Anan Liu

机构 * Tianjin University(天津大学)

专题命中 可控生成 :text-to-image(title,abstract);diffusion(title,abstract);image generation(abstract);分类 cs.CV

AI总结 LayerBind通过逐层实例绑定实现文本到图像扩散变换器中的区域和遮挡控制,无需训练即可提升生成质量和遮挡管理能力。

Comments Accepted by CVPR26

详情

展开后加载摘要…

URL PDF HTML 收藏