机构
*
Massachusetts Institute of Technology(麻省理工学院)
;
Amazon Web Services(亚马逊网络服务)
专题命中
视频多模态
:multi-modal(abstract);分类 cs.CV、cs.AI
Comments8 pages, 5 figures, accepted to the 11th IEEE International Workshop on Computer Vision in Sports (CVSports) at CVPR 2025; supplementary appendix included
机构
*
University of Maryland, College Park(马里兰大学学院市分校)
;
Meta Reality Labs(Meta现实实验室)
;
Worcester Polytechnic Institute(沃斯特理工学院)
;
University of Toronto(多伦多大学)
;
FAIR, Meta AI(Meta AI)
MoNetV2: Enhanced Motion Network for Freehand 3D Ultrasound Reconstruction
Mingyuan Luo, Xin Yang, Zhongnuo Yan, Yan Cao, Yuanji Zhang, Xindi Hu, Jin Wang, Haoxuan Ding, Wei Han, Litao Sun, Dong Ni
机构
*
National-Regional Key Technology Engineering Laboratory for Medical Ultrasound, School of Biomedical Engineering, Shenzhen University Medical School, Shenzhen University, Shenzhen, Guangdong, China(国家级医学超声关键技术研发实验室、生物医学工程学院、深圳大学医学院、深圳大学、深圳、广东、中国)
;
Medical UltraSound Image Computing (MUSIC) Lab, Shenzhen University, Shenzhen, Guangdong, China(医学超声图像计算(MUSIC)实验室、深圳大学、深圳、广东、中国)
;
Shenzhen RayShape Medical Technology Inc.(深圳RayShape医疗科技有限公司)
;
Cancer Center, Department of Ultrasound Medicine, Zhejiang Provincial People’s Hospital, Affiliated People’s Hospital of Hangzhou Medical College, Hangzhou, Zhejiang, China(肿瘤中心、超声医学科、浙江省人民医院、杭州医学院附属人民医院、杭州、浙江、中国)
;
Department of Health Management Center, Qilu Hospital, Cheeloo College of Medicine, Shandong University, Jinan, Shandong, China(健康管理中心、齐鲁医院、山东大学齐鲁医学院、济南、山东、中国)
Context-aware TFL: A Universal Context-aware Contrastive Learning Framework for Temporal Forgery Localization
Qilin Yin, Wei Lu, Xiangyang Luo, Xiaochun Cao
机构
*
School of Computer Science and Engineering, MoE Key Laboratory of Information Technology, Guangdong Province Key Laboratory of Information Security Technology, Sun Yat-sen University(计算机科学与工程学院、信息科技关键实验室、广东信息安全技术重点实验室、中山大学)
;
State Key Laboratory of Mathematical Engineering and Advanced Computing(数学工程与先进计算国家重点实验室)
;
School of Cyber Science and Technology, Shenzhen Campus, Sun Yat-sen University(网络安全科学与技术学院、深圳校区、中山大学)
Audio-Sync Video Generation with Multi-Stream Temporal Control
Shuchen Weng, Haojie Zheng, Zheng Chang, Si Li, Boxin Shi, Xinlong Wang
机构
*
Beijing Academy of Artificial Intelligence(北京人工智能研究院)
;
School of Software and Microelectronics, Peking University(北京大学软件与微电子学院)
;
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
;
Nat’l Key Lab of General AI, School of Intelligence Science and Technology, Peking University(国家通用人工智能实验室,北京大学智能科学与技术学院)
;
Nat’l Eng. Research Ctr. of Visual Tech., School of Computer Science, Peking University(国家视觉技术工程研究中心,北京大学计算机学院)
LLMs as World Models: Data-Driven and Human-Centered Pre-Event Simulation for Disaster Impact Assessment
Lingyao Li, Dawei Li, Zhenhui Ou, Xiaoran Xu, Jingxiao Liu, Zihui Ma, Runlong Yu, Min Deng
机构
*
University of South Florida(佛罗里达州立大学)
;
Arizona State University(亚利桑那州立大学)
;
Massachusetts Institute of Technology(麻省理工学院)
;
New York University(纽约大学)
;
University of Alabama(阿拉巴马大学)
;
Texas Tech University(德克萨斯科技大学)
Prisma: An Open Source Toolkit for Mechanistic Interpretability in Vision and Video
Sonia Joseph, Praneet Suresh, Lorenz Hufe, Edward Stevinson, Robert Graham, Yash Vadi, Danilo Bzdok, Sebastian Lapuschkin, Lee Sharkey, Blake Aaron Richards
机构
*
Mila Quebec(蒙特利尔大学)
;
McGill University(麦吉尔大学)
;
Meta
;
Université de Montréal(蒙特利尔大学)
;
Imperial College London(伦敦帝国理工学院)
;
Fraunhofer Heinrich Hertz Institute(弗劳恩霍夫 Heinrich Hertz 研究所)
;
Technological University Dublin(都柏林技术大学)
;
Apollo Research(Apollo 研究所)
专题命中
视频多模态
:multimodal(abstract);分类 cs.CV、cs.AI
Comments4 pages, 3 figures, 9 tables. Oral and Tutorial at the CVPR Mechanistic Interpretability for Vision (MIV) Workshop
A Survey on Event-driven 3D Reconstruction: Development under Different Categories
Chuanzhi Xu, Haoxian Zhou, Haodong Chen, Vera Chung, Qiang Qu
专题命中
视频多模态
:multimodal(abstract);分类 cs.CV、cs.AI
CommentsWe have decided not to submit this article and plan to withdraw it from public display. The content of this article will be presented in a more comprehensive form in another work