Evaluation of Phonetic Encoding Algorithms on Transcription Datasets

Evaluation of Phonetic Encoding Algorithms on Transcription Datasets

转录数据集上语音编码算法的评估

Abstract: In this work, a novel evaluation scheme built on a generalized variant of the Rand Index measure, namely, the Hüllermeier-Rifqi Index, is proposed in order to assess how well phonetic encoding algorithms conform to word-based transcriptions in IPA (International Phonetic Alphabet) notation. 摘要: 在这项工作中,我们提出了一种基于兰德指数(Rand Index)广义变体——即 Hüllermeier-Rifqi 指数——的新型评估方案,旨在评估语音编码算法与国际音标(IPA)符号下的基于单词的转录的一致性程度。

For this objective, the discordance score is obtained by calculating the absolute difference between the pairwise similarity values of ground-truth transcriptions and those of corresponding phonetic encodings, which are computed using normalized edit distance as a permutation dependent string metric. 为了实现这一目标,我们通过计算“真实转录(ground-truth)”的成对相似度值与相应语音编码的成对相似度值之间的绝对差来获得不一致性得分(discordance score),其中相似度是使用作为置换依赖字符串度量的归一化编辑距离计算得出的。

The resulting score is subsequently adjusted with respect to that of a random string generator incorporating the same alphabet as the encoder under consideration. 随后,该得分会根据一个使用与所考虑编码器相同字母表的随机字符串生成器的得分进行调整。

A wide range of phonetic encoders were evaluated as such on multi-lingual transcription datasets along with their recall capabilities based on the collision rate. 我们利用这种方法在多语言转录数据集上评估了多种语音编码器,并结合碰撞率评估了它们的召回能力。

The validity of the proposed scheme is further supported by its applicability in measuring the orthographic transparency of a language when the writing system is viewed as an inherent phonetic representation. 该方案的有效性进一步得到了验证,即当书写系统被视为一种内在的语音表示时,该方案可用于衡量语言的正字法透明度。