机构
*
SKLCCSE, School of Computer Science and Engineering, Beihang University(软件学院,北京航空航天大学)
;
Beihang University(北京航空航天大学)
;
School of Computer Science, Peking University(北京大学计算机学院)
;
the State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,中国科学院自动化研究所)
;
School of Information Science and Technology, University of Science and Technology of China(信息科学与技术学院,中国科学技术大学)
;
BNRist, Tsinghua University(北京理工大学,清华大学)
;
Zhongguancun Academy(中关村学院)
机构
*
South China University of Technology(南方科技大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Beijing National Research Center for Information Science and Technology, Tsinghua University(北京信息科学研究中心,清华大学)
;
Department of Automation, BNRist, Tsinghua University(清华大学自动化系,北京信息科学研究中心)
机构
*
Beijing Institute of Technology(北京理工大学)
;
Shanghai University(上海大学)
;
OpenNLP Lab(OpenNLP实验室)
;
Beijing University of Technology(北京工业大学)
;
The University of Adelaide(阿德莱德大学)
;
Tsinghua University(清华大学)
;
Inkeverse Group Limited(Inkeverse集团)
DepthPilot: From Controllability to Interpretability in Colonoscopy Video Generation
DepthPilot:从可控性到可解释性在结肠镜视频生成中
Junhu Fu, Ke Chen, Weidong Guo, Shuyu Liang, Jie Xu, Chen Ma, Kehao Wang, Shengli Lin, Zeju Li, Yuanyuan Wang, Yi Guo, Shuo Li
机构
*
College of Biomedical Engineering, Fudan University, Shanghai 200433, China(复旦大学生物医学工程学院)
;
Key Laboratory of Medical Imaging Computing and Computer Assisted Intervention of Shanghai, Shanghai 200032, China(上海医学影像计算与计算机辅助干预重点实验室)
;
Endoscopy Research Institute, Zhongshan Hospital, Fudan University, Shanghai 200032, China(复旦大学中山医院内窥镜研究所)
;
Shanghai Collaborative Innovation Center of Endoscopy, Shanghai 200032, China(上海内窥镜协同创新中心)
;
Department of Biomedical Engineering, Case Western Reserve University, Cleveland, OH 44106, USA(凯斯西储大学生物医学工程系)
;
Department of Computer and Data Science, Case Western Reserve University, Cleveland, OH 44106, USA(凯斯西储大学计算机与数据科学系)
机构
*
The University of Hong Kong(香港大学)
;
Ant Group(蚂蚁集团)
;
Tongyi(通义)
;
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
;
Huazhong University of Science and Technology(华中科技大学)
;
University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合性国家科学中心人工智能研究院)
;
Independent Researcher(独立研究员)
;
University of Macau(澳门大学)
;
The Chinese University of Hong Kong(香港中文大学)
Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generator
Uni-ViGU:通过基于扩散的视频生成器实现视频生成与理解的统一
Luozheng Qin, Jia Gong, Qian Qiao, Tianjiao Li, Li Xu, Haoyu Pan, Chao Qu, Zhiyu Tan, Hao Li
机构
*
Shanghai Academy of AI for Science(上海人工智能实验室)
;
Fudan University(复旦大学)
;
Independent Researcher(独立研究者)
;
Singapore University of Technology and Design(新加坡科技设计大学)
机构
*
Institute of Trustworthy Embodied AI, Fudan University(复旦大学可信具身智能研究所)
;
Shanghai Innovation Institute(上海创新研究院)
;
Shanghai Key Laboratory of Multimodal Embodied AI(上海市多模态具身人工智能重点实验室)
;
College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院)
;
The Chinese University of Hong Kong(香港中文大学)
;
Central South University(中南大学)
;
Fudan University(复旦大学)