机构
*
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))
;
Great Bay University(大湾区大学)
;
Dongguan Key Laboratory for Intelligence and Information Technology(东莞市智能信息技术重点实验室)
Go-with-the-Track: Video Compositing and Motion Control with Point Tracking
Go-with-the-Track: 基于点追踪的视频合成与运动控制
Koichi Namekata, Yash Kant, Zhizheng Liu, Ryan D Burgert, Yuancheng Xu, Kuan Heng Lin, Emmett Steven, Julien Philip, Li Ma, Andrea Vedaldi, Paul Debevec, Ning Yu
机构
*
Netflix USA(Netflix美国)
;
Eyeline Labs USA(Eyeline Labs美国)
;
University of Oxford(牛津大学)
;
Eyeline Labs Los Angeles USA(Eyeline Labs洛杉矶美国)
;
University of California, Los Angeles(加州大学洛杉矶分校)
;
Stony Brook University USA(石溪大学美国)
;
Columbia University USA(哥伦比亚大学美国)
;
Eyeline Labs United Kingdom(Eyeline Labs英国)
;
Netflix Los Angeles USA(Netflix洛杉矶美国)
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
Alibaba Group(阿里巴巴集团)
;
Tianji KernalMind Co., Ltd.(天机芯智有限公司)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Southeast University(东南大学)
;
Renmin University of China(中国人民大学)
;
The University of Tokyo(东京大学)
FlowWAM: Optical Flow as a Unified Action Representation for World Action Models
FlowWAM:光流作为世界动作模型的统一动作表示
Yixiang Chen, Peiyan Li, Yuan Xu, Qisen Ma, Jiabing Yang, Kai Wang, Jianhua Yang, Dong An, He Guan, Gaoteng Liu, Jianlou Si, Jun Huang, Jing Liu, Nianfeng Liu, Yan Huang, Liang Wang
机构
*
New Laboratory of Pattern Recognition (NLPR), Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所模式识别国家重点实验室)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
FiveAges(无)
;
MBZUAI(无)
;
Alibaba Group(阿里巴巴集团)
The GEST-Engine: From Event Graphs to Synthetic Video. A Full Technical Report
GEST引擎:从事件图到合成视频。完整技术报告
Nicolae Cudlenco, Mihai Masala, Marius Leordeanu
机构
*
Institute of Mathematics of the Romanian Academy(罗马尼亚科学院数学研究所)
;
National University of Science and Technology Politehnica Bucharest(布加勒斯特理工大学)
;
Büchi Labortechnik AG(布赫实验室技术公司)
TraMP-LLaMA: Generative Interpretability with Decoupled Instruction Tuning for Facial Expression Quality Assessment
TraMP-LLaMA: 基于解耦指令微调的生成式可解释性用于面部表情质量评估
Shuchao Duan, Alan Whone, Hossein Rahmani, Jun Liu, Majid Mirmehdi
机构
*
School of Computer Science University of Bristol(布里斯托大学计算机科学学院)
;
Translational Health Sciences University of Bristol(布里斯托大学转化健康科学学院)
;
School of Computing and Communications Lancaster University(兰卡斯特大学计算与通信学院)
机构
*
Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身人工智能研究所)
;
Robbyant, Ant Group(蚂蚁集团 Robbyant)
;
Hongkong University of Science and Technology(香港科技大学)
机构
*
School of Automation, Beijing Institute of Technology(北京理工大学自动化学院)
;
School of Mechanical Engineering, Beijing Institute of Technology(北京理工大学机械工程学院)
CommentsThis work is accepted by CVPR'26, Embodied AI Workshop. This paper represent a part of early result of our official world-action model zero-shot sim-to-real transfer work, which will be released soon