EgoPoseFormer v2: Accurate Egocentric Human Motion Estimation for AR/VR
EgoPoseFormer v2:面向AR/VR的精准自体中心人体运动估计
Zhenyu Li, Sai Kumar Dwivedi, Filip Maric, Carlos Chacon, Nadine Bertsch, Filippo Arcadu, Tomas Hodan, Michael Ramamonjisoa, Peter Wonka, Amy Zhao, Robin Kips, Cem Keskin, Anastasia Tkach, Chenhongyi Yang
机构
*
Meta
;
KAUST
;
Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所)
Structural Action Transformer for 3D Dexterous Manipulation
结构动作变换器用于3D灵巧操作
Xiaohan Lei, Min Wang, Bohong Weng, Wengang Zhou, Houqiang Li
机构
*
MoE Key Laboratory of Brain-inspired Intelligent Perception and Cognition, University of Science and Technology of China(脑启发智能感知与认知国家重点实验室,中国科学技术大学)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(人工智能研究院,合肥综合性国家科学中心)
InEdit-Bench: Benchmarking Intermediate Logical Pathways for Intelligent Image Editing Models
InEdit-Bench:智能图像编辑模型中间逻辑路径的基准测试
Zhiqiang Sheng, Xumeng Han, Zhiwei Zhang, Zenghui Xiong, Yifan Ding, Aoxiang Ping, Xiang Li, Tong Guo, Yao Mao
机构
*
State Key Laboratory of Optical Field Manipulation Science and Technology, Institute of Optics and Electronics, Chinese Academy of Sciences(光学场操控科学与技术国家重点实验室,光学电子研究所,中国科学院)
;
National Laboratory on Adaptive Optics(自适应光学国家实验室)
;
University of Chinese Academy of Sciences(中国科学院大学)
机构
*
The National Engineering Laboratory for Video Technology, School of Computer Science, Peking University(国家视频技术工程实验室,计算机科学学院,北京大学)
;
Institute for AI Industry Research (AIR), Tsinghua University(人工智能产业研究院(AIR),清华大学)
;
Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院,清华大学)
Training High-Level Schedulers with Execution-Feedback Reinforcement Learning for Long-Horizon GUI Automation
通过执行反馈强化学习训练高阶调度器以实现长周期GUI自动化
Zehao Deng, Tianjie Ju, Zheng Wu, Zhuosheng Zhang, Gongshen Liu
机构
*
School of Computer Science and Technology, Soochow University(苏州大学计算机科学与技术学院)
;
School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学学院)
VideoChat-M1: Collaborative Policy Planning for Video Understanding via Multi-Agent Reinforcement Learning
VideoChat-M1: 通过多智能体强化学习实现视频理解的协作策略规划
Boyu Chen, Zikang Wang, Zhengrong Yue, Kainan Yan, Chenyun Yu, Yi Huang, Zijun Liu, Yafei Wen, Xiaoxin Chen, Yang Liu, Peng Li, Yali Wang
机构
*
Shenzhen Key Lab of Computer Vision and Pattern Recognition(深圳计算机视觉与模式识别重点实验室)
;
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究院)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
VIVO AI Lab(VIVO人工智能实验室)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Shenzhen Campus of Sun Yat-sen University(孙逸仙大学深圳校区)
;
Shanghai Jiao Tong University(上海交通大学)
;
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)
;
Dept. of Comp. Sci. & Tech., Institute for AI, Tsinghua University(清华大学计算机科学与技术系,人工智能研究院)