Lu Qi, Yi-Wen Chen, Tao Zhang, Xiangtai Li, Xu Yang, Bo Du, Ming-Hsuan Yang
机构
*
Wuhan University(武汉大学)
;
Insta360 Research(Insta360研究院)
;
Department of EECS, University of California, Merced(加州大学默塞德分校电子工程与计算机科学系)
;
Nanyang Technological University(南洋理工大学)
;
Institute of Automation of the Chinese Academy of Sciences(中国科学院自动化研究所)
Knowledge-Preserved Model Tuning in Null-Space for Robust Spatio-Temporal Video Grounding
零空间中知识保留的模型调优用于鲁棒的时空视频定位
Haoxuan Chen, Xianqin Liu, Jian-Fang Hu
机构
*
School of Computer Science and Engineering, Sun Yat-sen University, China(中山大学计算机科学与工程学院)
;
National Information Center of GACC (Guangdong), GuangZhou, China(广东省GACC国家信息中心)
;
Guangdong Province Key Laboratory of Information Security Technology, China(广东省信息安全技术重点实验室)
;
Key Laboratory of Machine Intelligence and Advanced Computing, Ministry of Education, China(教育部机器智能与高级计算重点实验室)
Annotations Are Not All You Need: A Cross-modal Knowledge Transfer Network for Unsupervised Temporal Sentence Grounding
注释并非全部所需:面向无监督时间语句定位的跨模态知识迁移网络
Xiang Fang, Daizong Liu, Wanlong Fang, Pan Zhou, Yu Cheng, Keke Tang, Kai Zou
机构
*
Hubei Key Laboratory of Distributed System Security(湖北分布式系统安全重点实验室)
;
Hubei Engineering Research Center on Big Data Security(湖北大数据安全工程研究中心)
;
School of Cyber Science and Engineering(网络安全学院)
;
Huazhong University of Science and Technology(华中科技大学)
;
Peking University(北京大学)
;
Henan University(河南大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
Guangzhou University(广州大学)
;
Protagolabs Inc.(Protagolabs公司)
The Need for an External Observer Formalizing the Sufficiency Gap: A Mathematical Extension of Mixture Identifiability and Contextual Grounding in Sequence Models
GMOS: Grounding Moving Object Segmentation in 3D Space and Time
GMOS: 在3D空间和时间中定位运动物体分割
Junyu Xie, Tengda Han, Weidi Xie, Andrew Zisserman
机构
*
Visual Geometry Group, Department of Engineering Science, University of Oxford, UK(牛津大学工程科学系视觉几何组)
;
SAI, Shanghai Jiao Tong University, China(上海交通大学SAI)
机构
*
Shenzhen Campus of Sun Yat-sen University(中山大学深圳校区)
;
Guangdong Provincial Key Laboratory of Intelligent Information Processing and Shenzhen Key Laboratory of Media Security(广东省智能信息处理重点实验室和深圳媒体安全重点实验室)
;
Shenzhen University of Advanced Technology and Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences(深圳先进技术大学和深圳先进技术研究所,中国科学院)
;
Alibaba Group(阿里巴巴集团)
;
Shenzhen MSU-BIT University(深圳MSU-BIT大学)
Grounding Text Embeddings in Stakeholder Associations
将文本嵌入与利益相关者关联对齐
Jonathan Rystrøm, Sofie Burgos-Thorsen, Zihao Fu, Johan Irving Søltoft, Kenneth C. Enevoldsen, Chris Russell
机构
*
University of Oxford(牛津大学)
;
Institute for Wicked Problems(复杂问题研究所)
;
The Chinese University of Hong Kong(香港中文大学)
;
Danish Technical University(丹麦技术大学)
;
Aarhus University(奥胡斯大学)
HydraPrompt: An Adaptive and Asymmetric Framework of Vision-Language Models for Synthetic Image Detection
HydraPrompt: 面向合成图像检测的视觉语言模型自适应非对称框架
Senyuan Shi, Hao Tan, Zichang Tan, Shuhan Feng, Ajian Liu, Sergio Escalera, Jun Wan
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
School of Advanced Interdisciplinary Sciences (SAIS), University of Chinese Academy of Sciences(中国科学院大学先进交叉学科学院)
;
Shenzhen Institute of Advanced Technology (SIAT), Chinese Academy of Sciences(中国科学院深圳先进技术研究所)
;
MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所MAIS)
;
University of Barcelona(巴塞罗那大学)
Hierarchical Local-Global Transformer for Temporal Sentence Grounding
层次化局部-全局Transformer用于时间语句定位
Xiang Fang, Daizong Liu, Pan Zhou, Zichuan Xu, Ruixuan Li
机构
*
Hubei Engineering Research Center on Big Data Security, School of Cyber Science and Engineering, Huazhong University of Science and Technology(大数据安全湖北工程研究中心,华中科技大学网络安全科学与工程学院)
;
Wangxuan Institute of Computer fTechnology, Peking University(王宣计算机技术研究院,北京大学)
;
School of software, Dalian University of Technology(软件学院,大连理工大学)
;
School of Computer Science, and Technology, Huazhong University of Science, and Technology(计算机科学与技术学院,华中科技大学)
机构
*
Zhejiang Key Laboratory of Space Information Sensing and Transmission(浙江空间信息感知与传输重点实验室)
;
Hangzhou Dianzi University(杭州电子科技大学)
;
Zhejiang University(浙江大学)
;
Tsinghua University(清华大学)
;
Children's Hospital, Zhejiang University School of Medicine(浙江大学医学院附属儿童医院)
机构
*
School of AI for Science, Peking University(科学人工智能学院,北京大学)
;
School of Electronic and Computer Engineering, Peking University(电子与计算机工程学院,北京大学)
;
School of Computer Science, Peking University(计算机科学学院,北京大学)
SVAG-Bench: A Large-Scale Benchmark for Multi-Instance Spatio-temporal Video Action Grounding
SVAG-Bench:多实例时空视频动作定位的大规模基准
Tanveer Hannan, Shuaicong Wu, Mark Weber, Suprosanna Shit, Jindong Gu, Rajat Koner, Aljoša Ošep, Laura Leal-Taixé, Thomas Seidl
机构
*
LMU Munich(慕尼黑大学)
;
MCML
;
Technical University of Munich(慕尼黑技术大学)
;
University of Zurich(苏黎世大学)
;
University of Oxford(牛津大学)
;
Amazon(亚马逊)
;
NVIDIA(英伟达)
机构
*
Institute of Information Science and Technologies of the National Research Council (ISTI-CNR)(意大利国家研究理事会信息科学与技术研究所)
;
University of Pisa - Department of Information Engineering(比萨大学信息工程系)