GUIOdyssey: A Comprehensive Dataset for Cross-App GUI Navigation on Mobile Devices
Quanfeng Lu, Wenqi Shao, Zitao Liu, Lingxiao Du, Fanqing Meng, Boxuan Li, Botong Chen, Siyuan Huang, Kaipeng Zhang, Ping Luo
机构
*
Shanghai AI Laboratory(上海人工智能实验室)
;
The University of Hong Kong(香港大学)
;
Nanjing University(南京大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))
DreamScene: 3D Gaussian-based End-to-end Text-to-3D Scene Generation
Haoran Li, Yuli Tian, Kun Lan, Yong Liao, Lin Wang, Pan Hui, Peng Yuan Zhou
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Nanyang Technological University(南洋理工大学)
;
Hong Kong University of Science and Technology(香港科学理工大学)
;
University of Helsinki(赫尔辛基大学)
;
Aarhus University(奥胡斯大学)
专题命中
视觉空间推理
:planning(abstract)
CommentsExtended version of ECCV 2024 paper "DreamScene"
Homotopy-aware Multi-agent Navigation via Distributed Model Predictive Control
Haoze Dong, Meng Guo, Chengyi He, Zhongkui Li
机构
*
School of Advanced Manufacturing and Robotics, Peking University(先进制造与机器人学院,北京大学)
;
School of Computer Science and Engineering, Beihang University(计算机科学与工程学院,北航)
FreeQ-Graph: Free-form Querying with Semantic Consistent Scene Graph for 3D Scene Understanding
Chenlu Zhan, Yufei Zhang, Gaoang Wang, Hongwei Wang
机构
*
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
College of Biomedical Engineering and Instrument Science, Zhejiang University(浙江大学生物医学工程与仪器科学学院)
;
Zhejiang University-University of Illinois Urbana-Champaign Institute, Zhejiang University(浙江大学-伊利诺伊大学厄巴纳-香槟分校联合学院)
Emerging Properties in Unified Multimodal Pretraining
Chaorui Deng, Deyao Zhu, Kunchang Li, Chenhui Gou, Feng Li, Zeyu Wang, Shu Zhong, Weihao Yu, Xiaonan Nie, Ziang Song, Guang Shi, Haoqi Fan
机构
*
ByteDance Seed(字节跳动种子)
;
Shenzhen Institutes of Advanced Technology(深圳先进技术研究院)
;
Monash University(墨尔本大学)
;
Hong Kong University of Science and Technology(香港科学与技术大学)
;
UC Santa Cruz(加州大学圣克ruz分校)
机构
*
Soochow University(苏州大学)
;
Microsoft(微软)
;
Fudan University(复旦大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
University of Electronic Science and Technology of China(电子科技大学)
;
Sun Yat-sen University(中山大学)
;
Huazhong University of Science and Technology(华中科技大学)
;
The Chinese University of Hong Kong(香港中文大学)
ArtGS:3D Gaussian Splatting for Interactive Visual-Physical Modeling and Manipulation of Articulated Objects
Qiaojun Yu, Xibin Yuan, Yu jiang, Junting Chen, Dongzhe Zheng, Ce Hao, Yang You, Yixing Chen, Yao Mu, Liu Liu, Cewu Lu
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
National University of Singapore(国立新加坡大学)
;
Princeton University(普林斯顿大学)
;
Stanford University(斯坦福大学)
;
Hefei University of Technology(合肥工业大学)
机构
*
Southern University of Science and Technology, Shenzhen, China(南方科技大学,深圳,中国)
;
College of William and Mary, Williamsburg, VA 23185, USA(威廉玛丽学院,威廉斯堡,VA 23185,美国)
;
Peng Cheng Laboratory, Shenzhen, China(鹏城实验室,深圳,中国)
SurgTPGS: Semantic 3D Surgical Scene Understanding with Text Promptable Gaussian Splatting
Yiming Huang, Long Bai, Beilei Cui, Kun Yuan, Guankun Wang, Mobarak I. Hoque, Nicolas Padoy, Nassir Navab, Hongliang Ren
机构
*
The Chinese University of Hong Kong, Hong Kong SAR, China(香港中文大学)
;
Shenzhen Research Institute, CUHK, Shenzhen, China(深圳研究学院)
;
Technical University of Munich, Munich, Germany(慕尼黑技术大学)
;
University of Strasbourg & IHU Strasbourg, Strasbourg, France(斯特拉斯堡大学及斯特拉斯堡IHU)
;
University College London, London, United Kingdom(伦敦大学学院)
Vision Technologies with Applications in Traffic Surveillance Systems: A Holistic Survey
Wei Zhou, Li Yang, Lei Zhao, Runyu Zhang, Yifan Cui, Hongpu Huang, Kun Qie, Chen Wang
机构
*
School of Automation, Nanjing University of Science and Technology(南京理工大学自动化学院)
;
School of Transportation, Southeast University(东南大学交通学院)
;
Beijing Laboratory of General Aviation Technology, Beijing University of Civil Engineering and Architecture(北京建筑大学通用航空技术实验室)
From 2D to 3D Cognition: A Brief Survey of General World Models
Ningwei Xie, Zizi Tian, Lei Yang, Xiao-Ping Zhang, Meng Guo, Jie Li
机构
*
China Mobile Research Institute(中国移动研究院)
;
Shenzhen Ubiquitous Data Enabling Key Lab, Shenzhen International Graduate School, Tsinghua University(深圳无处不在数据赋能重点实验室、深圳国际研究生院、清华大学)
ControlMTR: Control-Guided Motion Transformer with Scene-Compliant Intention Points for Feasible Motion Prediction
Jiawei Sun, Chengran Yuan, Shuo Sun, Shanze Wang, Yuhang Han, Shuailei Ma, Zefan Huang, Anthony Wong, Keng Peng Tee, Marcelo H. Ang
机构
*
Department of Mechanical Engineering, National University of Singapore(新加坡国立大学机械工程系)
;
College of Information Science and Engineering, Northeastern University(东北大学信息科学与工程学院)
;
Moovita Pte Ltd(Moovita公司)
专题命中
视觉空间推理
:planning(abstract)
Journal ref2024 IEEE 27th International Conference on Intelligent Transportation Systems (ITSC)
DyNaVLM: Zero-Shot Vision-Language Navigation System with Dynamic Viewpoints and Self-Refining Graph Memory
Zihe Ji, Huangxuan Lin, Yue Gao
机构
*
SJTU Paris Elite Institute of Technology, Shanghai Jiao Tong University(上海交通大学巴黎精英技术研究所)
;
Department of Automation, Shanghai Jiao Tong University(上海交通大学自动化系)
;
MoE Key Lab of Artificial Intelligence and AI Institute, Shanghai Jiao Tong University(上海交通大学人工智能与AI研究所)