arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

图像生成

图像生成、文生图、图像编辑、扩散模型和可控生成。

共收录 1266 信号源:cs.CV, cs.GR, cs.MM

1. 个性化与一致性 1266 篇

2507.20721 2025-07-29 cs.CV 57%

AIComposer: Any Style and Content Image Composition via Feature Integration

Haowen Li, Zhenfeng Fan, Zhang Wen, Zhengzhou Zhu, Yunjin Li

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

Journal ref ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04599 2025-07-25 cs.CV 57%

QR-LoRA: Efficient and Disentangled Fine-tuning via QR Decomposition for Customized Generation

Jiahui Yang, Yongjia Ma, Donglin Di, Hao Li, Wei Chen, Yan Xie, Jianxun Cui, Xun Yang, Wangmeng Zuo

机构 * Harbin Institute of Technology(哈尔滨工业大学) Li Auto(利汽车) University of Science and Technology of China(中国科学技术大学)

专题命中 个性化与一致性 :text-to-image(abstract);分类 cs.CV

Comments ICCV 2025, 30 pages, 26 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07601 2025-07-23 cs.CV cs.LG 57%

Balanced Image Stylization with Style Matching Score

Yuxin Jiang, Liming Jiang, Shuai Yang, Jia-Wei Liu, Ivor Tsang, Mike Zheng Shou

机构 * Show Lab, National University of Singapore(新加坡国立大学Show实验室) Agency for Science, Technology and Research (A*STAR)(科技研究局) Nanyang Technological University(南洋理工大学) Wangxuan Institute of Computer Technology, Peking University(北京大学王轩计算机技术研究所)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

Comments ICCV 2025. Code: https://github.com/showlab/SMS Project page: https://yuxinn-j.github.io/projects/SMS.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02192 2025-07-22 cs.CV cs.AI 57%

DualReal: Adaptive Joint Training for Lossless Identity-Motion Fusion in Video Customization

Wenchuan Wang, Mengqi Huang, Yijing Tu, Zhendong Mao

机构 * University of Science and Technology of China(中国科学技术大学)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

Comments Accepted by ICCV2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.17609 2025-07-22 physics.med-ph cs.AI cs.CV eess.IV 57%

SynthRAD2025 Grand Challenge dataset: generating synthetic CTs for radiotherapy

Adrian Thummerer, Erik van der Bijl, Arthur Jr Galapon, Florian Kamp, Mark Savenije, Christina Muijs, Shafak Aluwini, Roel J. H. M. Steenbakkers, Stephanie Beuel, Martijn P. W. Intven, Johannes A. Langendijk, Stefan Both, Stefanie Corradini, Viktor Rogowski, Maarten Terpstra, Niklas Wahl, Christopher Kurz, Guillaume Landry, Matteo Maspero

专题命中 个性化与一致性 :image synthesis(abstract);分类 cs.CV

Comments 22 pages, 8 tables, 4 figures; Under submission to Medical Physics, as dataset paper for the SynhtRAD2025 Grand Challenge https://synthrad2025.grand-challenge.org/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.01319 2025-07-16 cs.GR cs.LG 57%

Model See Model Do: Speech-Driven Facial Animation with Style Control

Yifang Pan, Karan Singh, Luiz Gustavo Hafemann

机构 * University of Toronto(多伦多大学)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.GR

Comments 10 pages, 7 figures, SIGGRAPH Conference Papers '25

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.08044 2025-07-14 cs.CV cs.AI 57%

ConsNoTrainLoRA: Data-driven Weight Initialization of Low-rank Adapters using Constraints

Debasmit Das, Hyoungwoo Park, Munawar Hayat, Seokeon Choi, Sungrack Yun, Fatih Porikli

机构 * Qualcomm AI Research(高通人工智能研究)

专题命中 个性化与一致性 :image generation(abstract);分类 cs.CV

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05819 2025-07-09 cs.CV 57%

2D Instance Editing in 3D Space

Yuhuan Xie, Aoxuan Pan, Ming-Xian Lin, Wei Huang, Yi-Hua Huang, Xiaojuan Qi

机构 * The University of Hong Kong(香港大学)

专题命中 个性化与一致性 :image editing(abstract);分类 cs.CV

Comments 8 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02419 2025-07-08 cs.CV 57%

AvatarMakeup: Realistic Makeup Transfer for 3D Animatable Head Avatars

Yiming Zhong, Xiaolin Zhang, Ligang Liu, Yao Zhao, Yunchao Wei

机构 * Institute of Information Science and Visual Intelligence + X International Joint Laboratory, Beijing Jiaotong University(信息科学与视觉智能+X联合实验室,北京交通大学) College of Electrical Engineering and Automation, Shandong University of Science and Technology(电气工程与自动化学院,山东科技大学) School of Mathematical Sciences, University of Science and Technology of China(数学科学学院,中国科学技术大学)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.17254 2025-07-08 cs.CV cs.AI cs.LG 57%

Enhancing Long Video Generation Consistency without Tuning

Xingyao Li, Fengzhuo Zhang, Jiachun Pan, Yunlong Hou, Vincent Y. F. Tan, Zhuoran Yang

机构 * National University of Singapore(新加坡国立大学) Yale University(耶鲁大学)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

Comments ICML 2025 Workshop on Building Physically Plausible World Models (Best Paper), 32 pages, 17 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01792 2025-07-03 cs.CV 57%

FreeLoRA: Enabling Training-Free LoRA Fusion for Autoregressive Multi-Subject Personalization

Peng Zheng, Ye Wang, Rui Ma, Zuxuan Wu

机构 * School of Artificial Intelligence, Jilin University(吉林大学人工智能学院) Shanghai Innovation Institute(上海创新研究院) School of Computer Science, Fudan University(复旦大学计算机学院)

专题命中 个性化与一致性 :image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01055 2025-07-03 eess.IV cs.AI cs.CV 57%

Prompt Mechanisms in Medical Imaging: A Comprehensive Survey

Hao Yang, Xinlong Liang, Zhang Li, Yue Sun, Zheyu Hu, Xinghe Xie, Behdad Dashtbozorg, Jincheng Huang, Shiwei Zhu, Luyi Han, Jiong Zhang, Shanshan Wang, Ritse Mann, Qifeng Yu, Tao Tan

机构 * Faculty of Applied Sciences, Macao Polytechnic University(应用科学学院,澳门理工学院) Netherlands Cancer Institute, Department of Radiology(荷兰癌症研究所,放射科) Radboud University Medical Centre, Department of Radiology and Nuclear Medicine(拉德堡德大学医学中心,放射科和核医学科) College of Aerospace Science and Engineering, National University of Defense Technology(航空航天科学与工程学院,国防科技大学) Medical Department of Breast Cancer, Hunan Cancer Hospital(乳腺癌医学部,湖南癌症医院) the Affiliated Cancer Hospital of Xiangya School of Medicine, Central South University(湘雅医学院附属肿瘤医院,中南大学) Faculty of Biomedical Engineering, Eindhoven University of Technology(生物医学工程学院,埃因霍温理工大学) Laboratory of Advanced Theranostic Materials and Technology, University of Chinese Academy of Sciences(先进诊疗材料与技术实验室,中国科学院大学)

专题命中 个性化与一致性 :image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.09570 2025-07-03 cs.LG cs.AI cs.CV 57%

Improving Consistency Models with Generator-Augmented Flows

Thibaut Issenhuth, Sangchul Lee, Ludovic Dos Santos, Jean-Yves Franceschi, Chansoo Kim, Alain Rakotomamonjy

机构 * Criteo AI Lab, Paris, France(Criteo AI实验室,法国巴黎) Reasoning (AI/R) Laboratory, Korea Institute of Science(推理(AI/R)实验室,韩国科学技术院) Robot Department, University of Science(机器人部门,科学大学)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.20601 2025-06-26 cs.CV 57%

Video Perception Models for 3D Scene Synthesis

Rui Huang, Guangyao Zhai, Zuria Bauer, Marc Pollefeys, Federico Tombari, Leonidas Guibas, Gao Huang, Francis Engelmann

专题命中 个性化与一致性 :image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.06533 2025-06-25 cs.CV cs.CR 57%

DivTrackee versus DynTracker: Promoting Diversity in Anti-Facial Recognition against Dynamic FR Strategy

Wenshu Fan, Minxing Zhang, Hongwei Li, Wenbo Jiang, Hanxiao Chen, Xiangyu Yue, Michael Backes, Xiao Zhang

机构 * University of Electronic Science and Technology of China(电子科技大学) CISPA Helmholtz Center for Information Security(信息安全赫尔姆霍茨中心) The Chinese University of Hong Kong(香港中文大学)

专题命中 个性化与一致性 :image generation(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.09174 2025-06-24 cs.CV cs.AI 57%

DART: An Automated End-to-End Object Detection Pipeline with Data Diversification, Open-Vocabulary Bounding Box Annotation, Pseudo-Label Review, and Model Training

Chen Xin, Andreas Hartel, Enkelejda Kasneci

专题命中 个性化与一致性 :image generation(abstract);分类 cs.CV

Comments Corrected minor typos; no changes to results or conclusions

Journal ref Expert Systems with Applications 258 (2024): 125124

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07497 2025-06-23 cs.CV 57%

Genesis: Multimodal Driving Scene Generation with Spatio-Temporal and Cross-Modal Consistency

Xiangyu Guo, Zhanqian Wu, Kaixin Xiong, Ziyang Xu, Lijun Zhou, Gangwei Xu, Shaoqing Xu, Haiyang Sun, Bing Wang, Guang Chen, Hangjun Ye, Wenyu Liu, Xinggang Wang

机构 * Huazhong University of Science and Technology(华中科技大学) Xiaomi EV(小米电动车)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.15838 2025-06-23 cs.CV 57%

EchoShot: Multi-Shot Portrait Video Generation

Jiahao Wang, Hualian Sheng, Sijia Cai, Weizhan Zhang, Caixia Yan, Yachuang Feng, Bing Deng, Jieping Ye

机构 * Xi’an Jiaotong University(西安交通大学) Alibaba Cloud(阿里巴巴云)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03126 2025-06-04 cs.CV 57%

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Lu Qiu, Yizhuo Li, Yuying Ge, Yixiao Ge, Ying Shan, Xihui Liu

机构 * The University of Hong Kong(香港大学) ARC Lab, Tencent PCG(腾讯PCG ARC实验室)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

Comments Project released at: https://qiulu66.github.io/animeshooter/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02433 2025-06-04 cs.CV 57%

OmniTalker: One-shot Real-time Text-Driven Talking Audio-Video Generation With Multimodal Style Mimicking

Zhongjian Wang, Peng Zhang, Jinwei Qi, Guangyuan Wang, Chaonan Ji, Sheng Xu, Bang Zhang, Liefeng Bo

机构 * Tongyi Lab, Alibaba Group(通义实验室,阿里巴巴集团)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

Comments Project Page https://humanaigc.github.io/omnitalker

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21837 2025-05-29 cs.CV cs.LG 57%

UniMoGen: Universal Motion Generation

Aliasghar Khani, Arianna Rampini, Evan Atherton, Bruno Roy

机构 * Autodesk Research Canada(Autodesk加拿大研究实验室)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18445 2025-05-27 cs.CV 57%

OmniConsistency: Learning Style-Agnostic Consistency from Paired Stylization Data

Yiren Song, Cheng Liu, Mike Zheng Shou

机构 * Show Lab, National University of Singapore(新加坡国立大学展示实验室)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18078 2025-05-26 cs.CV 57%

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation

Junhao Chen, Mingjin Chen, Jianjin Xu, Xiang Li, Junting Dong, Mingze Sun, Puhua Jiang, Hongxiang Li, Yuhang Yang, Hao Zhao, Xiaoxiao Long, Ruqi Huang

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

Comments Our video demos and code are available at https://DanceTog.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07750 2025-05-23 cs.CV 57%

Motion by Queries: Identity-Motion Trade-offs in Text-to-Video Generation

Yuval Atzmon, Rinon Gal, Yoad Tewel, Yoni Kasten, Gal Chechik

机构 * NVIDIA

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

Comments (1) Project page: https://research.nvidia.com/labs/par/MotionByQueries/ (2) The methods and results in section 5, "Consistent multi-shot video generation", are based on the arXiv version 1 (v1) of this work. Starting version 2 (v2), we extend and further analyze those findings to efficient motion transfer (3) in v3 we added: results with WAN 2.1, baselines and more quality metrics

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14028 2025-05-21 cs.CV 57%

OmniStyle: Filtering High Quality Style Transfer Data at Scale

Ye Wang, Ruiqi Liu, Jiang Lin, Fei Liu, Zili Yi, Yilin Wang, Rui Ma

机构 * Jilin University(吉林大学) Nanjing University(南京大学) ByteDance(字节跳动) Adobe Engineering Research Center of Knowledge-Driven Human-Machine Intelligence, MOE, China Project page(知识驱动人机智能工程研究中心,中国)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18225 2025-05-20 cs.LG cs.CL cs.CV 57%

DeLoRA: Decoupling Angles and Strength in Low-rank Adaptation

Massimo Bini, Leander Girrbach, Zeynep Akata

机构 * University of Tübingen(图宾根大学) Tübingen AI Center(图宾根人工智能中心) Helmholtz Munich(海德堡-穆恩医疗中心) Technical University of Munich(慕尼黑技术大学) Munich Center for Machine Learning(慕尼黑机器学习中心) MDSI(慕尼黑数据科学研究所)

专题命中 个性化与一致性 :image generation(abstract);分类 cs.CV

Comments ICLR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05022 2025-05-20 cs.CV 57%

SOAP: Style-Omniscient Animatable Portraits

Tingting Liao, Yujian Zheng, Adilbek Karmanov, Liwen Hu, Leyang Jin, Yuliang Xiu, Hao Li

机构 * Mohamed bin Zayed University of Artificial Intelligence(莫扎德·本·扎耶德人工智能大学) Pinscreen USA(Pinscreen美国公司) Westlake University(西湖大学)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

Journal ref Siggraph 2025, page: https://tingtingliao.github.io/soap/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.09998 2025-05-16 cs.CV 57%

From Air to Wear: Personalized 3D Digital Fashion with AR/VR Immersive 3D Sketching

Ying Zang, Yuanqi Hu, Xinyu Chen, Yuxia Xu, Suhui Wang, Chunan Yu, Lanyun Zhu, Deyi Ji, Xin Xu, Tianrun Chen

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

Comments 8 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06985 2025-05-13 cs.CV 57%

BridgeIV: Bridging Customized Image and Video Generation through Test-Time Autoregressive Identity Propagation

Panwen Hu, Jiehui Huang, Qiang Sun, Xiaodan Liang

机构 * Mohamed bin Zayed University of Artificial Intelligence(莫扎伊德·本·泽亚德人工智能大学) Sun Yat-sen University(孙中山大学) University of Toronto(多伦多大学)

专题命中 个性化与一致性 :text-to-image(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06117 2025-05-12 cs.CV 57%

Photovoltaic Defect Image Generator with Boundary Alignment Smoothing Constraint for Domain Shift Mitigation

Dongying Li, Binyi Su, Hua Zhang, Yong Li, Haiyong Chen

机构 * School of Artificial Intelligence and Data Science, Hebei University of Technology(人工智能与数据科学学院,河北工业大学) China Xiongan Group Digital City Technology Company Ltd.(雄安集团数字城市技术有限公司) State Key Laboratory of Information Security, Institute of Information Engineering, Chinese Academy of Sciences(信息安全国家重点实验室,信息工程研究所,中国科学院) school of Instrumentation and Optoelectronic Engineering, Beihang University(仪器与光电工程学院,北航)

专题命中 个性化与一致性 :diffusion(abstract);分类 cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏