GeoRelight: Learning Joint Geometrical Relighting and Reconstruction with Flexible Multi-Modal Diffusion Transformers
GeoRelight: 基于灵活多模态扩散变换器的联合几何重照明与重建
Yuxuan Xue, Ruofan Liang, Egor Zakharov, Timur Bagautdinov, Chen Cao, Giljoo Nam, Shunsuke Saito, Gerard Pons-Moll, Javier Romero
机构
*
Codec Avatars Lab, Meta(Meta编码器动画实验室)
;
University of Tübingen(图宾根大学)
;
Max Planck Institute for Informatics, Saarland Informatics Campus(马克斯·普朗克信息学院,萨尔兰信息校园)
OmniUMI: Towards Physically Grounded Robot Learning via Human-Aligned Multimodal Interaction
OmniUMI: 通过人对齐的多模态交互实现物理基础的机器人学习
Shaqi Luo, Yuanyuan Li, Youhao Hu, Chenhao Yu, Chaoran Xu, Jiachen Zhang, Guocai Yao, Tiejun Huang, Ran He, Zhongyuan Wang
机构
*
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
;
MAIS & NLPR, Institute of Automation, Chinese Academy of Sciences(自动化研究所,中国科学院)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Beijing Institute of Technology(北京理工大学)
;
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Peking University(北京大学)
MATT-Diff: Multimodal Active Target Tracking by Diffusion Policy
MATT-Diff:基于扩散策略的多模态主动目标跟踪
Saida Liu, Nikolay Atanasov, Shumon Koga
机构
*
Department of Computer Science and Systems Engineering, Kobe University(神户大学计算机科学与系统工程系)
;
Department of Electrical and Computer Engineering, University of California San Diego(加州大学圣地亚哥分校电子与计算机工程系)
机构
*
School of Computer Science and Engineering, Sun Yat-sen University(中山大学计算机科学与工程学院)
;
Guangdong Key Laboratory of Big Data Analysis and Processing(广东省大数据分析与处理重点实验室)
机构
*
Guanghua School of Management, Peking University(北京大学光华管理学院)
;
Department of Industrial Engineering and Decision Analytics, The Hong Kong University of Science and Technology(香港科技大学工业工程与决策分析系)