ADP 前沿学习

← 板块一 · 研究前沿

Speech vocoding for laboratory phonology

Cernak, Benus, Lazaridis · cs.CL,cs.SD · 2016-09-15 · 原文

Using phonological speech vocoding, we propose a platform for exploring relations between phonology and speech processing, and in broader terms, for exploring relations between the abstract and physical structures of a speech signal. Our goal is to make a step towards bridging phonology and speech processing and to contribute to the program of Laboratory Phonology. We show three application examples for laboratory phonology: compositional phonological speech modelling, a comparison of phonological systems and an experimental phonological parametric text-to-speech (TTS) system. The featural representations of the following three phonological systems are considered in this work: (i) Government Phonology (GP), (ii) the Sound Pattern of English (SPE), and (iii) the extended SPE (eSPE). Comparing GP- and eSPE-based vocoded speech, we conclude that the latter achieves slightly better results than the former. However, GP - the most compact phonological speech representation - performs comparably to the systems with a higher number of phonological features. The parametric TTS based on phonological speech representation, and trained from an unlabelled audiobook in an unsupervised manner, ac

🔮 让 ChatGPT 全网深度追问

讲义

讲义·推断 依据「原文」自动生成的结构化摘要(推断),非原文表述;以原文为准。

1. 人话版

Using phonological speech vocoding, we propose a platform for exploring relations between phonology and speech processing, and in broader terms, for exploring relations between the abstract and physical structures of a speech signal.

Our goal is to make a step towards bridging phonology and speech processing and to contribute to the program of Laboratory Phonology.

2. 领域脉络

本文类目:cs.CL、cs.SD,属于其所在研究脉络的最新进展。

3. 机制拆解

We show three application examples for laboratory phonology: compositional phonological speech modelling, a comparison of phonological systems and an experimental phonological parametric text-to-speech (TTS) system.

The featural representations of the following three phonological systems are considered in this work: (i) Government Phonology (GP), (ii) the Sound Pattern of English (SPE), and (iii) the extended SPE (eSPE).

Comparing GP- and eSPE-based vocoded speech, we conclude that the latter achieves slightly better results than the former.

4. 证据与数字

摘要未给出量化结果——留意原文的实验与数据。

5. 反例与边界

However, GP - the most compact phonological speech representation - performs comparably to the systems with a higher number of phonological features.

6. 跨领域连接与意外收获

横跨 2 个类目(cs.CL、cs.SD),关注其在你兴趣板块间的迁移面。

7. 可复用方法

把本文机制与你手头项目对照,找一个两周内能验证的最小实验。

8. 术语表

精读时把不熟的术语记入此处,作为下次回忆的锚点。