Probing Spatial Structure in Pretrained Audio Representations
探究预训练音频表示中的空间结构
机构 * Music and Audio Research Laboratory, New York University, USA(音乐与音频研究实验室,纽约大学,美国)
AI总结 通过提出SARL基准,系统评估预训练音频模型对空间信息的编码能力,发现源因素比房间因素更易解码,且不同编码器对空间变化响应存在异质性。
Comments Accepted to Interspeech 2026