机构
*
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
Innovation Center of Yangtze River Delta, Zhejiang University(浙江大学长三角创新中心)
AV-Unified: A Unified Framework for Audio-visual Scene Understanding
AV-Unified: 一个用于音频-视觉场景理解的统一框架
Guangyao Li, Xin Wang, Wenwu Zhu
机构
*
Department of Computer Science and Technology, Tsinghua University(计算机科学与技术系,清华大学)
;
Beijing National Research Center for Information Science and Technology(北京信息科学与技术国家研究中心)
机构
*
University of Electronic Science and Technology of China(电子科技大学)
;
Defense Innovation Institute, Academy of Military Sciences(国防科技创新研究院)
;
Peking University(北京大学)
;
High-tech Institute(高新技术研究所)
机构
*
State Key Laboratory of Complex \& Critical Software Environment, Beihang University
;
College of Science, Mathematics
;
Technology, Wenzhou-Kean University
;
360 AI Security Lab
;
Faculty of Innovation Engineering, Macau University of Science
;
Hong Kong University of Science
;
University of Chinese Academy of Sciences
Aeroacoustic signatures reveal fast transient dynamics of vapor-jet-driven cavity oscillations in metallic additive manufacturing
气动声学特征揭示金属增材制造中蒸发喷射驱动腔体振荡的快速瞬态动力学
Haolin Liu, S. Kiana Naghibzadeh, Zhongshu Ren, Yanming Zhang, Jiayun Shao, Samuel J. Clark, Kamel Fezzaa, Xuzhe Zeng, Lin Gao, Wentao Yan, Noel Walkington, Kaushik Dayal, Tao Sun, Anthony D. Rollett, Levent Burak Kara
Language Conditioning Improves Accuracy of Aircraft Goal Prediction in Non-Towered Airspace
语言引导提升非塔台空域飞机目标预测的准确性
Sundhar Vinodh Sangeetha, Chih-Yuan Chiu, Sarah H. Q. Li, Shreyas Kousik
机构
*
Georgia Institute of Technology(佐治亚理工学院)
;
School of Aerospace Engineering(航空航天工程学院)
;
School of Electrical and Computer Engineering(电气与计算机工程学院)
;
School of Mechanical Engineering(机械工程学院)
Cut to the Chase: Training-free Multimodal Summarization via Chain-of-Events
直击核心:一种无需训练的多模态摘要方法 via 事件链
Xiaoxing You, Qiang Huang, Lingyu Li, Xiaojun Chang, Jun Yu
机构
*
School of Computer Science, Hangzhou Dianzi University(杭州电子科技大学计算机科学学院)
;
School of Intelligence Science and Engineering, Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)智能科学与工程学院)
;
School of Information Science and Technology, University of Science and Technology of China(中国科学技术大学信息科学与技术学院)
Spatiotemporal Heterogeneity of AI-Driven Traffic Flow Patterns and Land Use Interaction: A GeoAI-Based Analysis of Multimodal Urban Mobility
人工智能驱动交通流模式与土地利用相互作用的时空异质性:基于GeoAI的多模式城市交通分析
Olaf Yunus Laitinen Imanov
机构
*
Department of Applied Mathematics(应用数学系)
;
Computer Science (DTU Compute), Technical University of Denmark, Kongens Lyngby, Denmark(计算机科学(DTU Compute),丹麦技术大学,Kongens Lyngby)
Learning Next Action Predictors from Human-Computer Interaction
从人机交互中学习下一步动作预测器
Omar Shaikh, Valentin Teutschbein, Kanishk Gandhi, Yikun Chi, Nick Haber, Thomas Robinson, Nilam Ram, Byron Reeves, Sherry Yang, Michael S. Bernstein, Diyi Yang
机构
*
Stanford University(斯坦福大学)
;
Hasso Plattner Institute(哈索普拉特纳研究所)
;
New York University(纽约大学)
机构
*
Jockey Club STEM Laboratory of Quantitative Remote Sensing(裘英俊STEM实验室(定量遥感))
;
Department of Geography, the University of Hong Kong(香港大学地理系)
;
School of Information and Electronics, Beijing Institute of Technology(北京理工大学信息电子学院)
;
Beijing Key Laboratory of Fractional Signals and Systems(北京分数信号与系统重点实验室)
;
Department of Electrical and Computer Engineering, University of Delaware(德雷塞尔大学电气与计算机工程系)
SPARC: Concept-Aligned Sparse Autoencoders for Cross-Model and Cross-Modal Interpretability
SPARC:概念对齐的稀疏自编码器用于跨模型和跨模态可解释性
Ali Nasiri-Sarvi, Hassan Rivaz, Mahdi S. Hosseini
机构
*
Department of Computer Science and Software Engineering (CSSE) Concordia University, Canada(计算机科学与软件工程系(CSSE)康科迪亚大学,加拿大)
;
Department of Electrical and Computer Engineering (ECE) Concordia University, Canada(电气与计算机工程系(ECE)康科迪亚大学,加拿大)