Comments35 pages, 14 figures. This work has been accepted by Machine Intelligence Research. The arxiv version is kept updating by adding more novel methods, datasets and insights. The official video interpretation of this paper can be referred at https://youtu.be/2lfNaTkcTHI
Comments8 pages, 3 figures, Association for the Advancement of Artificial Intelligence (AAAI2020). arXiv admin note: substantial text overlap with arXiv:1907.01709
Comments7 pages, 3 figures; This work was presented at the 3rd Workshop on YouTube-8M Large-Scale Video Understanding, at the International Conference on Computer Vision (ICCV 2019) in Seoul, Korea
Commentsaccepted to CVPR'17 Workshop on YouTube-8M Large-Scale Video Understanding (oral presentation); code is at https://github.com/taufikxu/youtube on branches kunxu and zhs
On the Robustness of Temporal Vision-Language Models for Surgical Endoscopy Videos
手术内窥镜视频的时间视觉语言模型的鲁棒性研究
Darakshan Rashid, Raza Imam, Ufaq Khan, Muhammad Bilal, Shazad Ashraf, Dwarikanath Mahapatra, Mohammad Yaqub, Muhammad Haris Khan, Imran Razzak, Brejesh Lall, Lena Maier-Hein, Yutong Xie
机构
*
Indian Institute of Technology Delhi(印度理工学院德里分校)
;
Mohamed bin Zayed University of Artificial Intelligence (MBZUAI)(穆罕默德·本·扎耶德人工智能大学)
;
Birmingham City University(伯明翰城市大学)
;
University Hospitals Birmingham(伯明翰大学医院)
;
Khalifa University(哈利法大学)
;
German Cancer Research Center (DKFZ)(德国癌症研究中心)