Focus Then Listen: An Empirical Study of Plug-and-Play Audio Enhancer for Noise-Robust Large Audio Language Models
先聚焦后聆听:探索用于噪声鲁棒的大规模音频语言模型的即插即用音频增强器
机构 * University of California, Berkeley(加州大学伯克利分校) ; University of Washington(华盛顿大学)
专题命中 其他推理 :reasoning(abstract)
AI总结 提出即插即用的音频增强器FTL,通过分离语音与非语音并利用模态路由器预测目标模态,生成任务自适应增强信号,无需微调即可提升LALMs在噪声环境下的性能。
Comments Accepted by ICML 2026 Workshop (Machine Learning for Audio)