跳到正文
原文
arXiv · Audio and Speech· Ki Woong Moon, Daniel Brenner·· 7 小时前AI 评分12

韵律训练表征在参数匹配下是否优于可训练融合?——基于冻结 HuBERT 的对比研究

Does a prosody-trained representation help beyond trainable fusion? A parameter-matched study with frozen HuBERT

AI 导读

研究使用冻结 HuBERT 骨干与 64 维韵律表征(log F0、voicing、Delta log F0、log energy、光谱倾斜)进行自动语音识别对比。

来源:arXiv · Audio and Speech · arxiv.org