跳至主導覽 跳至搜尋 跳過主要內容

Personalized Taiwanese Speech Synthesis using Cascaded ASR and TTS Framework

  • Yuan Fu Liao
  • , Wen Han Hsu
  • , Chen Ming Pan
  • , Wern Jun Wang
  • , Matus Pleva
  • , Daniel Hladek

研究成果: Conference contribution同行評審

3 引文 斯高帕斯(Scopus)

摘要

To bring endangered Taiwanese language back to life, this paper leveraged a large-scale Taiwanese across Taiwan (TAT) corpus to construct cascaded automatic speech recognition (ASR) and text-to-speech (TTS)-based personalized Taiwanese speech synthesizers to help young people to learn how to speak Taiwanese. This paradigm not only alleviates the low resource, nonparallel corpus and cross-lingual training data problems but also dramatically reduces the fine-tuning data size and training time. Experimental results on a Taiwanese-to-Taiwanese and Mandarin-to-Taiwanese voice conversion tasks had shown that it allows us to successfully produce good personalized Taiwanese TTS with only approximately 3 minutes of data in both cases.

原文English
主出版物標題2022 32nd International Conference Radioelektronika, RADIOELEKTRONIKA 2022 - Proceedings
發行者Institute of Electrical and Electronics Engineers Inc.
ISBN(電子)9781728186863
DOIs
出版狀態Published - 2022
事件32nd International Conference Radioelektronika, RADIOELEKTRONIKA 2022 - Kosice, 斯洛伐克
持續時間: 21 4月 202222 4月 2022

出版系列

名字2022 32nd International Conference Radioelektronika, RADIOELEKTRONIKA 2022 - Proceedings

Conference

Conference32nd International Conference Radioelektronika, RADIOELEKTRONIKA 2022
國家/地區斯洛伐克
城市Kosice
期間21/04/2222/04/22

指紋

深入研究「Personalized Taiwanese Speech Synthesis using Cascaded ASR and TTS Framework」主題。共同形成了獨特的指紋。

引用此