GenVid2Robot: From Video Generation to Robot Manipulation via Rigid-Geometric Consistency
GenVid2Robot:通过刚性几何一致性从视频生成到机器人操作
Haohui Huang, Xi Yuan, Panpan Liao, Tao Teng, Chenguang Yang, Jing Guo, Yi Guo
机构
*
School of Automation, Guangdong University of Technology(广东工业大学自动化学院)
;
University of Liverpool(利物浦大学)
;
Department of Computing, The Hong Kong Polytechnic University(香港理工大学计算学系)
;
State Key Laboratory of Submarine Geoscience, School of Automation and Intelligent Sensing, Shanghai Jiao Tong University(上海交通大学海洋地球科学国家重点实验室,自动化与智能感知学院)
Watching Synthetic Videos: Aligning Cross-modal Representations with Visual Synthesis for Zero-shot Video Captioning
观看合成视频:针对零样本视频字幕生成的视觉合成跨模态表征对齐
Liangyu Fu, Junbo Wang, Yuke Li, Ya Jing, Xuecheng Wu, Zhiyong Wang
机构
*
School of Software, Northwestern Polytechnical University(西北工业大学软件学院)
;
School of Information Science and Technology, Beijing University of Technology(北京工业大学信息科学与技术学院)
;
School of Computer Science and Technology, Xi’an Jiaotong University(西安交通大学计算机科学与技术学院)
;
School of Computer Science, The University of Sydney(悉尼大学计算机科学学院)
VideoRAE: Taming Video Foundation Models for Generative Modeling via Representation Autoencoders
VideoRAE:通过表示自动编码器驯服用于生成建模的视频基础模型
Zhihao Xie, Junfeng Wu, Xinting Hu, Junchao Huang, Li Jiang
机构
*
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Huazhong University of Science and Technology(华中科技大学)
;
Shenzhen Loop Area Institute(深圳河套学院)
;
University of Science and Technology of China(中国科学技术大学)
机构
*
The University of Sydney(悉尼大学)
;
University of Melbourne(墨尔本大学)
;
City University of Hong Kong(香港城市大学)
;
Wuhan University(武汉大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
CommentsThis work provides a basis for the ECCV 2026 LifeGenIP Challenge on Unlearnable Videos against Diffusion-based Customization. Challenge page: https://lifegenip.cc/competition. Evaluation code: ECCV26_LifeGenIP_starting_kit" target="_blank" rel="noopener">https://github.com/tmllab/ECCV26_LifeGenIP_starting_kit. Project page: https://saythe17.github.io/TC-UAP/
Customizing Video Portraits via Identity-ActionDecoupling
通过身份-动作解耦定制视频肖像
Junxiong Lin, Haoran Wang, Xinji Mai, Zeng Tao, Xuan Tong, Ivy Pan, Wenqiang Zhang
机构
*
College of Intelligent Robotics and Advanced Manufacturing, Fudan University(复旦大学智能机器人与先进制造学院)
;
College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院)
;
The University of Hong Kong(香港大学)
Christian Internò, Alexander Pondaven, Habon Issa, Fabio Pizzati, Francesco Pinto, Markus Olhofer, Ivan Laptev, Philip Torr, Eero P. Simoncelli, Barbara Hammer, David Klindt
机构
*
Bielefeld University(比勒费尔德大学)
;
University of Oxford(牛津大学)
;
Cold Spring Harbor Laboratory(冷泉港实验室)
;
MBZUAI(穆罕默德·本·扎耶德人工智能大学)
;
Independent(独立作者)
;
Honda Research Institute EU(本田欧洲研究院)
;
New York University(纽约大学)
;
Flatiron Institute, Simons Foundation(西蒙斯基金会熨斗研究所)
Revealing Artifacts via Noise Amplification: A Novel Perspective for AI-Generated Video Detection
通过噪声放大揭示伪影:AI生成视频检测的新视角
Renxi Cheng, Jie Gui, Hongsong Wang
机构
*
School of Cyber Science and Engineering, Southeast University(东南大学网络空间安全学院)
;
Purple Mountain Laboratories(紫金山实验室)
;
Engineering Research Center of Blockchain Application, Supervision And Management (Southeast University), Ministry of Education(教育部区块链应用监管工程研究中心(东南大学))
;
School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院)
;
Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University), Ministry of Education(教育部新一代人工智能技术及其跨学科应用重点实验室(东南大学))
机构
*
Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身人工智能研究所)
;
Shanghai Collaborative Innovation Center of Intelligent Visual Computing(上海智能视觉计算协同创新中心)
;
Bytedance Seed(字节跳动Seed)
;
The University of Hong Kong(香港大学)