Back to Main Conference 2012
LREC 2012main

Versatile Speech Databases for High Quality Synthesis for Basque

Proceedings of the Eighth International Conference on Language Resources and Evaluation (LREC 2012)

DOI:10.63317/4pphewyngnxz

Abstract

This paper presents three new speech databases for standard Basque. They are designed primarily for corpus-based synthesis but each database has its specific purpose: 1) AhoSyn: high quality speech synthesis (recorded also in Spanish), 2) AhoSpeakers: voice conversion and 3) AhoEmo3: emotional speech synthesis. The whole corpus design and the recording process are described with detail. Once the databases were collected all the data was automatically labelled and annotated. Then, an HMM-based TTS voice was built and subjectively evaluated. The results of the evaluation are pretty satisfactory: 3.70 MOS for Basque and 3.44 for Spanish. Therefore, the evaluation assesses the quality of this new speech resource and the validity of the automated processing presented.

Details

Paper ID
lrec2012-main-014
Pages
pp. 3308-3312
BibKey
sainz-etal-2012-versatile
Editor
N/A
Publisher
European Language Resources Association (ELRA)
ISSN
2522-2686
ISBN
978-2-9517408-7-7
Conference
Eighth International Conference on Language Resources and Evaluation
Location
Istanbul, Turkey
Date
21 May 2012 27 May 2012

Authors

  • IS

    Iñaki Sainz

  • DE

    Daniel Erro

  • EN

    Eva Navas

  • IH

    Inma Hernáez

  • JS

    Jon Sanchez

  • IS

    Ibon Saratxaga

  • IO

    Igor Odriozola

Links