Bibliographie complète
Neural Text-to-Speech Synthesis for an Under-Resourced Language in a Diglossic Environment: the Case of Gascon Occitan
Type de ressource
Conference Paper
Auteurs/contributeurs
- Corral, Ander (Author)
- Leturia, Igor (Author)
- Séguier, Aure (Author)
- Barret, Michäel (Author)
- Dazéas, Benaset (Author)
- Boula de Mareüil, Philippe (Author)
- Quint, Nicolas (Author)
- Beermann, Dorothee (Editor)
- Besacier, Laurent (Editor)
- Sakti, Sakriani (Editor)
- Soria, Claudia (Editor)
Title
Neural Text-to-Speech Synthesis for an Under-Resourced Language in a Diglossic Environment: the Case of Gascon Occitan
Abstract
Occitan is a minority language spoken in Southern France, some Alpine Valleys of Italy, and the Val d'Aran in Spain, which only very recently started developing language and speech technologies. This paper describes the first project for designing a Text-to-Speech synthesis system for one of its main regional varieties, namely Gascon. We used a state-of-the-art deep neural network approach, the Tacotron2-WaveGlow system. However, we faced two additional difficulties or challenges: on the one hand, we wanted to test if it was possible to obtain good quality results with fewer recording hours than is usually reported for such systems; on the other hand, we needed to achieve a standard, non-Occitan pronunciation of French proper names, therefore we needed to record French words and test phoneme-based approaches. The evaluation carried out over the various developed systems and approaches shows promising results with near production-ready quality. It has also allowed us to detect the phenomena for which some flaws or fall of quality occur, pointing at the direction of future work to improve the quality of the actual system and for new systems for other language varieties and voices.
Date
2020-05
Proceedings Title
Proceedings of the 1st Joint Workshop on Spoken Language Technologies for Under-resourced languages (SLTU) and Collaboration and Computing for Under-Resourced Languages (CCURL)
Conference Name
SLTU 2020
Place
Marseille, France
Publisher
European Language Resources association
Pages
53–60
Language
English
ISBN
979-10-95546-35-1
Short Title
Neural Text-to-Speech Synthesis for an Under-Resourced Language in a Diglossic Environment
Accessed
13/05/2024 09:01
Library Catalog
ACLWeb
Référence
Corral, A., Leturia, I., Séguier, A., Barret, M., Dazéas, B., Boula de Mareüil, P., & Quint, N. (2020). Neural Text-to-Speech Synthesis for an Under-Resourced Language in a Diglossic Environment: the Case of Gascon Occitan. In D. Beermann, L. Besacier, S. Sakti, & C. Soria (Eds.), Proceedings of the 1st Joint Workshop on Spoken Language Technologies for Under-resourced languages (SLTU) and Collaboration and Computing for Under-Resourced Languages (CCURL) (pp. 53–60). European Language Resources association. https://aclanthology.org/2020.sltu-1.8
Langue
Tâche
Lien vers cette notice