机构
*
Hangzhou Institute for Advanced Study(杭州高等研究院)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
Computer Network Information Center(计算机网络信息中心)
;
Chinese Academy of Sciences(中国科学院)
CommentsThis paper has been accepted to the Late-Breaking Results (LBR) track of the 28th International Conference on Multimodal Interaction (ICMI 2026)
ID-VTG: Image-Disambiguated Video Temporal Grounding
ID-VTG:基于图像消歧的视频时间定位
Minghang Zheng, Jingli Wei, Hongyi Yang, Yang Liu
机构
*
Wangxuan Institute of Computer Technology, Peking University(北京大学王选计算机研究所)
;
State Key Laboratory of General Artificial Intelligence, Peking University(北京大学通用人工智能国家重点实验室)
Comments50 pages, 4 figures, 12 tables. Revised version adds a Supplementary Exploratory Analysis comparing Sociodual pathways with blind developmental/task-process segmentation, with granularity and record-format checks and an unseen-case extension. Main construct specification unchanged; minor editorial and terminology refinements
Self-supervised Learning Matters: A Simple Ensemble Solution for Micro-Gesture Recognition
自监督学习至关重要:一种用于微手势识别的简单集成方案
Tingyi Liu, Kun Li, Fei Wang, Junjie Chen, Zhiliang Wu, Jihao Gu, Haixu Liu, Dan Guo
机构
*
Hefei University of Technology(合肥工业大学)
;
United Arab Emirates University(阿拉伯联合酋长国大学)
;
Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合性国家科学中心人工智能研究院)
;
Anhui Evolution Technology Co., Ltd.(安徽进化科技有限公司)
;
Nanyang Technological University(南洋理工大学)
;
University College London(伦敦大学学院)
;
The University of Sydney(悉尼大学)
;
Beijing QBoson Quantum Technology Co., Ltd.(北京量子芯科技有限公司)
Commentsv2: substantial revision. Cross-architecture ordering statistic and MHA/GQA cohort-split claim retracted per recomputation; corpus and companion-paper citations refreshed. See paper's "Changes from Version 1" section for the full list of superseded values
Comments12 pages, 3 figures. Accepted at the 35th ACM International Conference on Information and Knowledge Management (CIKM 2026). Project page: this https URL (https://lizesheng13.github.io/bridge/)
Exploiting Completeness Perception with Diffusion Transformer for Unified 3D MRI Synthesis
利用扩散变换器的完整性感知实现统一的3D MRI合成
Junkai Liu, Nay Aung, Theodoros N. Arvanitis, Joao A. C. Lima, Steffen E. Petersen, Le Zhang
机构
*
School of Engineering, University of Birmingham, UK(伯明翰大学工程学院)
;
William Harvey Research Institute, Queen Mary University London, UK(女王玛丽大学伦敦威廉·哈里维研究所)
;
Barts Heart Centre, St Bartholomew’s Hospital, Barts Health NHS Trust, UK(巴特勒心脏中心,圣巴塞洛缪医院,巴特勒健康 NHS信托)
;
Division of Cardiology, Johns Hopkins University School of Medicine, US(约翰霍普金斯大学医学院心脏病科)
Guided Diffusion by Optimized Loss Functions on Relaxed Parameters for Inverse Material Design
通过放松参数优化损失函数引导扩散用于逆材料设计
Jens U. Kreber, Christian Weißenfels, Joerg Stueckler
机构
*
Intelligent Perception in Technical Systems Group, University of Augsburg, Germany(技术系统智能感知组,乌尔姆大学,德国)
;
Faculty of Mathematics, Natural Science and Engineering, University of Augsburg, Germany(数学、自然科学与工程学院,乌尔姆大学,德国)
;
Centre for Advanced Analytics and Predictive Sciences, University of Augsburg, Germany(高级分析与预测科学中心,乌尔姆大学,德国)