Phonetically Balanced Text Corpus Design Using a Similarity Measure for a Stereo Super-Wideband Speech Database

Yoo Rhee OH; Yong Guk KIM; Mina KIM; Hong Kook KIM; Mi Suk LEE; Hyun Joo BAE

doi:10.1587/transinf.E94.D.1459

IEICE TRANSACTIONS on Information

Phonetically Balanced Text Corpus Design Using a Similarity Measure for a Stereo Super-Wideband Speech Database

Yoo Rhee OH, Yong Guk KIM, Mina KIM, Hong Kook KIM, Mi Suk LEE, Hyun Joo BAE

Full Text Views

0

Cite this

Summary :

In this paper, we propose a text corpus design method for a Korean stereo super-wideband speech database. Since a small-sized text corpus for speech coding is generally required for speech coding, the corpus should be designed to comply with the pronunciation behavior of natural conversation in order to ensure efficient speech quality tests. To this end, the proposed design method utilizes a similarity measure between the phoneme distribution occurring from natural conversation and that from the designed text corpus. In order to achieve this goal, we first collect and refine text data from textbooks and websites. Next, a corpus is designed from the refined text data based on the similarity measure to compare phoneme distributions. We then construct a Korean stereo super-wideband speech (K-SW) database using the designed text corpus, where the recording environment is set to meet the conditions defined by ITU-T. Finally, the subjective quality of the K-SW database is evaluated using an ITU-T super-wideband codec in order to demonstrate that the K-SW database is useful for developing and evaluating super-wideband codecs.

Publication: IEICE TRANSACTIONS on Information Vol.E94-D No.7 pp.1459-1466

Publication Date: 2011/07/01

Publicized

Online ISSN: 1745-1361

DOI: 10.1587/transinf.E94.D.1459

Type of Manuscript: PAPER

Category: Speech and Hearing

Cite this

Copy

Yoo Rhee OH, Yong Guk KIM, Mina KIM, Hong Kook KIM, Mi Suk LEE, Hyun Joo BAE, "Phonetically Balanced Text Corpus Design Using a Similarity Measure for a Stereo Super-Wideband Speech Database" in IEICE TRANSACTIONS on Information, vol. E94-D, no. 7, pp. 1459-1466, July 2011, doi: 10.1587/transinf.E94.D.1459.
Abstract: In this paper, we propose a text corpus design method for a Korean stereo super-wideband speech database. Since a small-sized text corpus for speech coding is generally required for speech coding, the corpus should be designed to comply with the pronunciation behavior of natural conversation in order to ensure efficient speech quality tests. To this end, the proposed design method utilizes a similarity measure between the phoneme distribution occurring from natural conversation and that from the designed text corpus. In order to achieve this goal, we first collect and refine text data from textbooks and websites. Next, a corpus is designed from the refined text data based on the similarity measure to compare phoneme distributions. We then construct a Korean stereo super-wideband speech (K-SW) database using the designed text corpus, where the recording environment is set to meet the conditions defined by ITU-T. Finally, the subjective quality of the K-SW database is evaluated using an ITU-T super-wideband codec in order to demonstrate that the K-SW database is useful for developing and evaluating super-wideband codecs.
URL: https://global.ieice.org/en_transactions/information/10.1587/transinf.E94.D.1459/_p

Copy

@ARTICLE{e94-d_7_1459,
author={Yoo Rhee OH, Yong Guk KIM, Mina KIM, Hong Kook KIM, Mi Suk LEE, Hyun Joo BAE, },
journal={IEICE TRANSACTIONS on Information},
title={Phonetically Balanced Text Corpus Design Using a Similarity Measure for a Stereo Super-Wideband Speech Database},
year={2011},
volume={E94-D},
number={7},
pages={1459-1466},
abstract={In this paper, we propose a text corpus design method for a Korean stereo super-wideband speech database. Since a small-sized text corpus for speech coding is generally required for speech coding, the corpus should be designed to comply with the pronunciation behavior of natural conversation in order to ensure efficient speech quality tests. To this end, the proposed design method utilizes a similarity measure between the phoneme distribution occurring from natural conversation and that from the designed text corpus. In order to achieve this goal, we first collect and refine text data from textbooks and websites. Next, a corpus is designed from the refined text data based on the similarity measure to compare phoneme distributions. We then construct a Korean stereo super-wideband speech (K-SW) database using the designed text corpus, where the recording environment is set to meet the conditions defined by ITU-T. Finally, the subjective quality of the K-SW database is evaluated using an ITU-T super-wideband codec in order to demonstrate that the K-SW database is useful for developing and evaluating super-wideband codecs.},
keywords={},
doi={10.1587/transinf.E94.D.1459},
ISSN={1745-1361},
month={July},}

Copy

TY - JOUR
TI - Phonetically Balanced Text Corpus Design Using a Similarity Measure for a Stereo Super-Wideband Speech Database
T2 - IEICE TRANSACTIONS on Information
SP - 1459
EP - 1466
AU - Yoo Rhee OH
AU - Yong Guk KIM
AU - Mina KIM
AU - Hong Kook KIM
AU - Mi Suk LEE
AU - Hyun Joo BAE
PY - 2011
DO - 10.1587/transinf.E94.D.1459
JO - IEICE TRANSACTIONS on Information
SN - 1745-1361
VL - E94-D
IS - 7
JA - IEICE TRANSACTIONS on Information
Y1 - July 2011
AB - In this paper, we propose a text corpus design method for a Korean stereo super-wideband speech database. Since a small-sized text corpus for speech coding is generally required for speech coding, the corpus should be designed to comply with the pronunciation behavior of natural conversation in order to ensure efficient speech quality tests. To this end, the proposed design method utilizes a similarity measure between the phoneme distribution occurring from natural conversation and that from the designed text corpus. In order to achieve this goal, we first collect and refine text data from textbooks and websites. Next, a corpus is designed from the refined text data based on the similarity measure to compare phoneme distributions. We then construct a Korean stereo super-wideband speech (K-SW) database using the designed text corpus, where the recording environment is set to meet the conditions defined by ITU-T. Finally, the subjective quality of the K-SW database is evaluated using an ITU-T super-wideband codec in order to demonstrate that the K-SW database is useful for developing and evaluating super-wideband codecs.
ER -

IEICE TRANSACTIONS on Information

Phonetically Balanced Text Corpus Design Using a Similarity Measure for a Stereo Super-Wideband Speech Database

Summary :

Authors

Keyword

Latest Issue

Contents

Links

Call for Papers

Submit to IEICE Trans.

Transactions NEWS

Popular articles

IEICE TRANSACTIONS on Information

Phonetically Balanced Text Corpus Design Using a Similarity Measure for a Stereo Super-Wideband Speech Database

Summary :

Authors

Keyword

Latest Issue

Contents

Copyrights notice of machine-translated contents

Cite this

Links

Call for Papers

Submit to IEICE Trans.

Transactions NEWS

Popular articles