Skip to content
Hot eventLive

自监督语音模型发现音系向量算术

1 reports1 sources3 hr ago updated

Get the story

AI overview

2026年10月7日,arXiv·Audio and Speech 发布一手研究,显示自监督语音模型(S3M)的表示空间中存在与音系特征对应的线性方向,构成音系向量。研究覆盖96种语言,发现 [d] 与 [t] 的差可得到浊音向量,将该向量加到 [p] 上可生成 [b],缩放该向量则得到连续的浊音变化,即 [b] = [d] - [t] + [p]。这表明 S3M 以可解释、可组合的音系向量编码语音,即音系向量算术;代码与交互式演示已公开,论文被 ACL 2026 Findings 接收。

Generated from reports · updated 3 hr ago

Timeline

Follow the coverage from different angles.

Oct 7, 2026
  1. arXiv · Audio and Speech
    自监督语音模型发现音系向量算术:[b] = [d] - [t] + [p]

    研究显示,自监督语音模型(S3M)的表示空间中存在与音系特征对应的线性方向,构成音系向量。研究覆盖 96 种语言,发现 [d] 与 [t] 的差可得到浊音向量,将该向量加到 [p] 上可生成 [b],缩放该向量则得到连续的浊音变化。这表明 S3M 以可解释、可组合的音系向量编码语音,即音系向量算术;代码与交互式演示已公开,论文被 ACL 2026 Findings 接收。

Heat trend

Current heat 9·Comparable peak 10(Oct 7)·Comparable change over 24 hours –

02.557.510Oct7Oct7Oct7Oct7

The trend compares only the same participants observed continuously; its range may be smaller than the current heat count. Move or click on the chart to inspect hourly heat; use the left and right arrow keys to switch.