CommentsAccepted to ACL 2026 System Demonstrations. 11 pages, 5 figures, 8 tables
Journal refProceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 3: System Demonstrations), pages 829-839, 2026
机构
*
Department of Computing, Imperial College London(伦敦帝国理工学院计算系)
;
Department of Earth Science & Engineering, Imperial College London(伦敦帝国理工学院地球科学与工程系)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
Department of Computer Science, Heriot-Watt University(赫瑞瓦特大学计算机科学系)
;
School of Computer Science, Northumbria University(诺森比亚大学计算机科学学院)
;
College of Computer Science & Software Engineering, Shenzhen University(深圳大学计算机科学与软件学院)
;
School of Artificial Intelligence, Shenzhen University(深圳大学人工智能学院)
;
Guangdong Provincial Key Laboratory of Intelligent Information Processing, Shenzhen University(深圳大学广东省智能信息处理重点实验室)
;
School of Engineering and Design, Hunan Normal University(湖南师范大学工程与设计学院)
;
Department of Computer Science, University of Oxford(牛津大学计算机科学系)
;
Department of Computer Science, University of Exeter(埃克塞特大学计算机科学系)
Audio-Visual Speech Enhancement: Architectural Design and Deployment Strategies
音频-视觉语音增强:架构设计与部署策略
Anis Hamadouche, Haifeng Luo, Mathini Sellathurai, Amir Hussain, Tharm Ratnarajah
机构
*
School of Engineering & Physical Sciences, Heriot-Watt University(赫瑞斯泰学院,赫瑞斯泰大学)
;
College of Engineering Department of Electrical and Computer Engineering, San Diego State University(工程学院电子与计算机工程系,圣地亚哥州立大学)
;
SDAIA-KFUPM Joint Research Centre for Artificial Intelligence, King Fahd University of Petroleum and Minerals(SDAIA-KFUPM人工智能联合研究中心,国王法赫德石油与矿物大学)
Audio-Language Models for Audio-Centric Tasks: A Systematic Survey
用于以音频为中心任务的音频-语言模型:系统综述
Yi Su, Jisheng Bai, Qisheng Xu, Kele Xu, Yong Dou
机构
*
College of Computer Science and Technology, National University of Defense Technology(计算机科学与技术学院,国防科技大学)
;
School of Communications and Information Engineering, Xi’an University of Posts and Telecommunications(通信与信息工程学院,西安邮电大学)