Represented but Ignored: A Causal Account of Prosodic Underuse in Audio-Language Models
被表征却被忽视:音频语言模型中韵律使用不足的因果解释
机构 * University of Connecticut(康涅狄格大学) ; The University of Texas at Austin(德克萨斯大学奥斯汀分校)
AI总结 本研究针对音频语言模型(audio-LLM)韵律使用不足的问题,通过阶段特定探测阶梯与隐藏状态干预实验,发现模型能表征韵律却未在输出中表达,瓶颈在于韵律的使用而非感知。