O-DisCo-Edit: Object Distortion Control for Unified Realistic Video Editing
机构 * Tsinghua University(清华大学) ; Huawei Inc.(华为公司) ; Pengcheng National Laboratory(Pengcheng国家实验室)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
视觉与机器人
图像生成、文生图、图像编辑、扩散模型和可控生成。
机构 * Tsinghua University(清华大学) ; Huawei Inc.(华为公司) ; Pengcheng National Laboratory(Pengcheng国家实验室)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
机构 * KT Corporation, South Korea(韩国KT公司) ; University of Illinois Urbana-Champaign, USA(伊利诺伊大学厄巴纳-香槟分校) ; Pohang University of Science and Technology (POSTECH), South Korea(浦项科技大学(POSTECH))
专题命中 可控生成 :image generation(abstract);分类 cs.CV
机构 * Kling Team, Kuaishou Technology(快手科技 Kling 团队) ; Zhejiang University(浙江大学) ; Tsinghua University(清华大学)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
Comments Technical Report. Project Page: https://chenmingthu.github.io/milm/
机构 * Stony Brook University(石溪大学)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
机构 * Alibaba Group(阿里巴巴集团) ; Beijing University of Posts and Telecommunications(北京邮电大学)
专题命中 可控生成 :image generation(abstract);分类 cs.CV
Comments Accepted by CIKM 2025
机构 * School of Translational Medicine, Faculty of Medicine, Nursing and Health Sciences, Monash University(转化医学学院,医学、护理与健康科学学院,墨尔本大学) ; Melbourne Sexual Health Centre, Alfred Health(墨尔本性健康中心,阿尔弗雷德健康中心) ; AIM for Health Lab, Monash University(健康促进实验室,墨尔本大学) ; Faculty of IT, Monash University(信息技术学院,墨尔本大学) ; Faculty of Engineering, Monash University(工程学院,墨尔本大学) ; Faculty of Infectious and Tropical Diseases, London School of Hygiene and Tropical Medicine(传染病与热带医学学院,伦敦热带医学学院)
专题命中 可控生成 :image synthesis(abstract);分类 cs.CV
Comments 11 pages, 4 figures
机构 * School of Sciences, Xi’an University of Technology, Xi’an 710054, China(西安理工大学科学学院) ; Department of Applied Mathematics(应用数学系) ; Theoretical Physics, University of Cambridge, UK(理论物理,剑桥大学,英国) ; Department of Applied Mathematics, The Hong Kong Polytechnic University, Hung Hom, Hong Kong(应用数学系,香港理工大学,九龙)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
机构 * Xi’an Key Laboratory of Big Data and Intelligent Vision, Xidian University, Xi’an 710071, China(西安大数据与智能视觉重点实验室,西安电子科技大学,西安710071,中国) ; Key Laboratory of Collaborative Intelligence Systems, Ministry of Education, Xidian University, Xi’an 710071, China(协同智能系统重点实验室,教育部,西安电子科技大学,西安710071,中国) ; School of Computer Science and Technology, Xidian University, Xi’an 710071, China(计算机科学与技术学院,西安电子科技大学,西安710071,中国)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
机构 * SNU(首尔国立大学)
专题命中 可控生成 :image synthesis(abstract);分类 cs.CV
Comments 6 pages, 4 figures, NeurIPS Creative AI Track 2025
机构 * Department of Electrical and Electronic Engineering, The University of Hong Kong(电子与电气工程系,香港大学)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
机构 * School of Computer Science and Engineering, University of Electronic Science and Technology of China(计算机科学与工程学院,电子科技大学)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
机构 * Jiangxi Normal University, China(江西师范大学) ; Swansea University, United Kingdom(斯旺西大学)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
机构 * Xi'an Jiaotong-Liverpool University(西安交通大学利物浦大学) ; University of Liverpool(利物浦大学) ; Microsoft(微软公司)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
Comments Accepted at Neurocomputing 2025
机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) ; University of Surrey(萨里大学)
专题命中 可控生成 :diffusion(abstract);分类 cs.MM
Comments The first comprehensive survey on controllable TTS. Accepted to the EMNLP 2025 main conference
机构 * MAIS, Institute of Automation, Chinese Academy of Sciecnes(MAIS,自动化研究所,中国科学院) ; Ant Group(蚂蚁集团)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
机构 * School of Electronic Information and Electrical Engineering, Shanghai Jiao Tong University(电子信息与电气工程学院,上海交通大学) ; Ningbo Institute of Digital Twin, Eastern Institute of Technology(宁波数字孪生研究所,东部技术研究所) ; PhiGent Robotics, Beijing, China(PhiGent Robotics,北京,中国)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
Comments IEEE Transactions on Pattern Analysis and Machine Intelligence
机构 * Institute of Artificial Intelligence, Xiamen University(厦门大学人工智能研究所) ; Guangdong Laboratory of Artificial Intelligence(广东省人工智能与数字经济实验室) ; Peking University(北京大学) ; School of Future Technology, South China University of Technology(华南理工大学未来技术学院) ; Hangzhou Dianzi University(杭州电子科技大学) ; Central Laboratory of Lishui Hospital of Wenzhou Medical University, The First Affiliated Hospital of Lishui University, Lishui People's Hospital(丽水市人民医院中央实验室、丽水大学第一附属医院、丽水人民医院)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
机构 * OpenAI
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
Comments 15 pages, 9 figures
专题命中 可控生成 :diffusion(abstract);分类 cs.GR
Comments accepted to IEEE Transactions on Visualization and Computer Graphics
专题命中 可控生成 :inpainting(abstract);分类 cs.CV
Comments Project page: https://whhu7.github.io/IGFuse
机构 * Harbin Institute of Technology(哈尔滨工业大学) ; Tianyijiaotong Technology Ltd.(天翼交通科技有限公司) ; Independent Researcher(独立研究者)
专题命中 可控生成 :inpainting(abstract);分类 cs.CV
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
Comments arXiv admin note: text overlap with arXiv:2505.04410
机构 * College of Computer Science and Engineering, Shandong University of Science and Technology(计算机科学与工程学院,山东科技大学)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
Comments 11 pages, 6 figures
机构 * DAMO Academy, Alibaba Group(阿里达摩院) ; Hupan Lab(华盘实验室) ; INSAIT ; Zhejiang University(浙江大学)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
Comments Project page: https://jingyunliang.github.io/RealisMotion
机构 * The Shenzhen Campus of Sun Yat-sen University, Sun Yat-sen University, Shenzhen, China(中山大学深圳校区,中山大学,深圳,中国) ; Peng Cheng Laboratory, Shenzhen, China(鹏城实验室,深圳,中国) ; SSE, The Chinese University of Hong Kong, Shenzhen, China(香港中文大学深圳校区,香港中文大学,深圳,中国) ; Beijing University of Posts(北京邮电大学)
专题命中 可控生成 :image generation(abstract);分类 cs.CV
Comments IEEE INFOCOM PerAI6G 2024(accepted)
机构 * University of Leeds(利兹大学) ; Carnegie Mellon University(卡内基梅隆大学)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
专题命中 可控生成 :text-to-image(abstract);分类 cs.CV
机构 * ByteDance(字节跳动)
专题命中 可控生成 :image generation(abstract);分类 cs.CV
Comments CVPR 2025; Code and models: https://github.com/ByteVisionLab/TokenFlow
机构 * BEO AI Address(BEO AI)
专题命中 可控生成 :inpainting(abstract);分类 cs.CV