DSpace at KOASAS: Deep MOS Predictor for Synthetic Speech Using Cluster-Based Modeling

DSpace at KOASAS

College of Engineering(공과대학)School of Electrical Engineering(전기및전자공학부)EE-Conference Papers(학술회의논문)

Deep MOS Predictor for Synthetic Speech Using Cluster-Based Modeling

Cited 0 time in webofscience

Cited 10 time in scopus

Hit : 161
Download : 0

Export

Choi, Yeunju / Jung, Youngmoon / Kim, Hoi-Rin researcher

While deep learning has made impressive progress in speech synthesis and voice conversion, the assessment of the synthesized speech is still carried out by human participants. Several recent papers have proposed deep-learning-based assessment models and shown the potential to automate the speech quality assessment. To improve the previously proposed assessment model, MOSNet, we propose three models using cluster-based modeling methods: using a global quality token (GQT) layer, using an Encoding Layer, and using both of them. We perform experiments using the evaluation results of the Voice Conversion Challenge 2018 to predict the mean opinion score of synthesized speech and similarity score between synthesized speech and reference speech. The results show that the GQT layer helps to predict human assessment better by automatically learning the useful quality tokens for the task and that the Encoding Layer helps to utilize frame-level scores more precisely.

Publisher: ISCA

Issue Date: 2020-10-27

Language: English

Citation: Interspeech 2020, pp.1743 - 1747

DOI: 10.21437/Interspeech.2020-2111

URI: http://hdl.handle.net/10203/277892

Appears in Collection: EE-Conference Papers(학술회의논문)

Files in This Item: There are no files associated with this item.

Display Full Item Record

qr_code

트윗하기

KOASAS

Knowledge Service Development Team, KAIST 291 Daehak-ro, Yuseong-gu, Daejeon 34141, Republic of Korea. T. 82-42-350-4493 Email. koasas@kaist.ac.kr
Copyright © 2016. Korea Advanced Institute of Science and Technology. All Rights Reserved.

KOASAS

KOASAS

Browse

Deep MOS Predictor for Synthetic Speech Using Cluster-Based Modeling

KOASAS

Communities & Collections