机构
*
McGill University(麦吉尔大学)
;
Hong Kong University of Science and Technology(香港科技大学)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Tsinghua University(清华大学)
Towards Fast and Effective Long Video Understanding of Multimodal Large Language Models via Adaptive Quasi-Gaussian Sampling
面向多模态大语言模型的长视频快速有效理解:自适应准高斯采样
Kun Zhang, Chenxin Fang, Tao Chen, Baiyang Song, Yunhang Shen, Yiyi Zhou, Rongrong Ji
机构
*
Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(厦门大学多媒体可信感知与高效计算教育部重点实验室)
M^2C-EvDet: Multi-Domain Multi-Order Cross-Modal Knowledge Distillation for Event-based Object Detection
M^2C-EvDet:面向事件目标检测的多域多阶跨模态知识蒸馏
Wei Bao, Siqi Li, Shouan Pan, Yi Xie, Yue Gao
机构
*
BNRist, THUIBCS, BLBCI, School of Software, Tsinghua University(清华大学软件学院、北京信息科学与技术国家研究中心、清华-英特尔先进计算与智能技术联合研究中心、北京国家区块链与物联网技术研究中心)
;
Yangtze Delta Region Institute, Tsinghua University(清华大学长三角研究院)
;
School of Economics and Management, Beijing Forestry University(北京林业大学经济管理学院)
CommentsAccepted for publication in the 17th ACM International Conference on Bioinformatics, Computational Biology and Health Informatics (ACM BCB 2026). DOI: https://doi.org/10.1145/3807503.3819363
Decoding Multimodal Cues: Unveiling the Implicit Meaning Behind Hateful Videos
解码多模态线索:揭示仇恨视频背后的隐含意义
Junyu Lu, Deyi Ji, Liqun Liu, Xiaokun Zhang, Youlin Wu, Roy Ka-Wei Lee, Peng Shu, Huan Yu, Jie Jiang, Bo Xu, Liang Yang, Hongfei Lin
机构
*
Dalian University of Technology(大连理工大学)
;
Tencent(腾讯)
;
City University of Hong Kong(香港城市大学)
;
Singapore University of Technology and Design(新加坡科技设计大学)
D3VL: Understanding Driving Scenes from 3D Time Series Data and Video with Language Models
D3VL:利用语言模型从3D时间序列数据和视频中理解驾驶场景
Heesang Han, A. Lynn Abbott, Abhijit Sarkar
机构
*
Bradley Department of Electrical and Computer Engineering, Virginia Tech(弗吉尼亚理工大学布拉德利电气与计算机工程系)
;
Virginia Tech Transportation Institute(弗吉尼亚理工大学交通研究所)
;
Sanghani Center for Artificial Intelligence and Data Analytics(桑哈尼人工智能与数据分析中心)
Comments5 pages excluding references and supplements, 2 figures and 2 tables, Proceedings of the Workshop on Structured Data for Health at the 43rd International Conference on Machine Learning, Seoul, South Korea
Journal refICML2026 workshop on Structured Data for Health (SD4H)