机构
*
Shanghai Jiao Tong University(上海交通大学)
;
University of Cambridge(剑桥大学)
;
University of Technology Nuremberg(纽伦堡技术大学)
;
Nanyang Technological University(南洋理工大学)
;
Department of Automation, Key Laboratory of System Control and Information Processing of Ministry of Education, State Key Laboratory of Avionics Integration and Aviation System-of-Systems Synthesis, Shanghai Key Laboratory of Navigation and Location Based Services, Shanghai Jiao Tong University(自动化系,教育部系统控制与信息处理重点实验室,航空系统集成与航空系统-of-系统综合国家重点实验室,上海导航与定位基于服务重点实验室,上海交通大学)
VideoAtlas: Navigating Long-Form Video in Logarithmic Compute
VideoAtlas: 在对数计算中导航长视频
Mohamed Eltahir, Ali Habibullah, Yazan Alshoibi, Lama Ayash, Tanveer Hussain, Naeemullah Khan
机构
*
King Abdullah University of Science and Technology (KAUST)(卡斯特大学)
;
Department of Computer Science, King Khalid University (KKU)(国王 Khalid 大学计算机科学系)
;
Department of Computer Science, Edge Hill University(Edge Hill 大学计算机科学系)
Aion: Towards Hierarchical 4D Scene Graphs with Temporal Flow Dynamics
Aion:面向具有时间流动态的分层4D场景图
Iacopo Catalano, Eduardo Montijano, Javier Civera, Julio A. Placed, Jorge Pena-Queralta
机构
*
University of Turku(图尔库大学)
;
Centre for Artificial Intelligence, Zürich University of Applied Sciences(应用科学大学人工智能中心)
;
Instituto Tecnológico de Aragón (ITA) and the University of Zaragoza(阿兰理工大学和萨拉戈萨大学)
;
University of Zaragoza(萨拉戈萨大学)
Unsupervised Decomposition and Recombination with Discriminator-Driven Diffusion Models
无监督分解与重组:基于判别器驱动的扩散模型
Archer Wang, Emile Anand, Yilun Du, Marin Soljačić
机构
*
Research Laboratory of Electronics, MIT, Cambridge, MA 02139, USA(MIT电子研究实验室)
;
Department of Physics, Massachusetts Institute of Technology, MIT, Cambridge, MA, USA(麻省理工学院物理系)
;
NSF AI Institute for Artificial Intelligence(国家科学基金会人工智能与基本相互作用研究所)
;
Kempner Institute, Harvard University, Cambridge, MA, USA(哈佛大学 Kempner 院)
;
Google DeepMind, Mountain View, CA, USA(Google DeepMind)
;
School of Computer Science, Georgia Institute of Technology, Atlanta, USA(佐治亚理工学院计算机科学系)
MagicWorld: Towards Long-Horizon Stability for Interactive Video World Exploration
MagicWorld: 向交互视频世界探索的长时稳定迈进
Guangyuan Li, Bo Li, Jinwei Chen, Xiaobin Hu, Lei Zhao, Peng-Tao Jiang
机构
*
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
vivo BlueImage Lab, vivo Mobile Communication Co., Ltd.(vivo蓝影实验室,vivo移动通信有限公司)
;
National University of Singapore(新加坡国立大学)
Multi-modal 3D Pose and Shape Estimation with Computed Tomography
基于计算机断层扫描的多模态3D姿态和形状估计
Mingxiao Tu, Hoijoon Jung, Alireza Moghadam, Jineel Raythatha, Lachlan Allan, Jeremy Hsu, Andre Kyme, Jinman Kim
机构
*
School of Computer Science, The University of Sydney, Sydney, NSW 2006, Australia(悉尼大学计算机科学学院)
;
School of Biomedical Engineering, The University of Sydney, Sydney, NSW 2006, Australia(悉尼大学生物医学工程学院)
;
Trauma Service, Westmead Hospital, Australia(西墨迪医院创伤服务)
From Digital Twins to World Models:Opportunities, Challenges, and Applications for Mobile Edge General Intelligence
从数字孪生到世界模型:移动边缘通用智能的机会、挑战与应用
Jie Zheng, Dusit Niyato, Changyuan Zhao, Jiawen Kang, Jiacheng Wang
机构
*
State-Province Joint Engineering and Research Center of Advanced Networking and Intelligent Information Services, College of Computer, Northwest University(高级网络与智能信息服务省州联合工程研究中心,计算机学院,西北大学)
;
College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学)
;
Automation of School, Guangdong University of Technology(自动化学院,广东工业大学)
MosaicMem: Hybrid Spatial Memory for Controllable Video World Models
MosaicMem: 用于可控视频世界模型的混合空间记忆
Wei Yu, Runjia Qian, Yumeng Li, Liquan Wang, Songheng Yin, Sri Siddarth Chakaravarthy P, Dennis Anthony, Yang Ye, Yidi Li, Weiwei Wan, Animesh Garg
机构
*
University of Toronto(多伦多大学)
;
Vector Institute(向量研究所)
;
The University of Osaka(大阪大学)
;
Georgia Institute of Technology(佐治亚理工学院)
;
Mujin Inc.(Mujin公司)
;
University of Texas at Austin(德克萨斯大学奥斯汀分校)
;
Taiyuan University of Technology(太原本科技大学)
OccTENS: 3D Occupancy World Model via Temporal Next-Scale Prediction
OccTENS: 通过时间下一尺度预测实现的3D占用世界模型
Bu Jin, Songen Gu, Xiaotao Hu, Yupeng Zheng, Xiaoyang Guo, Qian Zhang, Xiaoxiao Long, Wei Yin
机构
*
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Horizon Robotics(地平线机器人)
;
Nanjing University(南京大学)
机构
*
Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学北京校区人工智能学院)
;
Beijing Key Laboratory of Research on Large Models and Intelligent Governance(北京大型模型与智能治理研究重点实验室)
;
Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(教育部下一代智能搜索与推荐工程研究中心)
;
Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)
;
Institute for AI Industry Research (AIR), Tsinghua University(清华大学人工智能产业研究院)
Rewarding DINO: Predicting Dense Rewards with Vision Foundation Models
奖励DINO:基于视觉基础模型预测密集奖励
Pierre Krack, Tobias Jülg, Wolfram Burgard, Florian Walter
机构
*
Department of Computer Science & Artificial Intelligence, University of Technology Nuremberg(技术大学纽伦堡计算机科学与人工智能系)
;
TUM School of Computation, Information and Technology, Technical University of Munich(慕尼黑技术大学计算、信息与技术学院)
ViSA: Visited-State Augmentation for Generalized Goal-Space Contrastive Reinforcement Learning
ViSA:基于已访问状态的通用目标空间对比强化学习增强
Issa Nakamura, Tomoya Yamanokuchi, Yuki Kadokawa, Jia Qu, Shun Otsub, Ken Miyamoto, Shotaro Miwa, Takamitsu Matsubara
机构
*
Graduate School of Information Science, Nara Institute of Science and Technology (NAIST)(信息科学研究生院,奈良科学与技术研究所)
;
Advanced Technology R&D Center, Mitsubishi Electric Corporation(三菱电机先进技术研发中心)