Back to Main Conference 2012
LREC 2012main

Building Text-To-Speech Voices in the Cloud

Proceedings of the Eighth International Conference on Language Resources and Evaluation (LREC 2012)

DOI:10.63317/4joy56e74b7m

Abstract

The AT&T VoiceBuilder provides a new tool to researchers and practitioners who want to have their voices synthesized by a high-quality commercial-grade text-to-speech system without the need to install, configure, or manage speech processing software and equipment.It is implemented as a web service on the AT&T Speech Mashup Portal.The system records and validates users' utterances, processes them to build a synthetic voice and provides a web service API to make the voice available to real-time applications through a scalable cloud-based processing platform. All the procedures are automated to avoid human intervention. We present experimental comparisons of voices built using the system.

Details

Paper ID
lrec2012-main-416
Pages
pp. 3317-3321
BibKey
conkie-etal-2012-building
Editor
N/A
Publisher
European Language Resources Association (ELRA)
ISSN
2522-2686
ISBN
978-2-9517408-7-7
Conference
Eighth International Conference on Language Resources and Evaluation
Location
Istanbul, Turkey
Date
21 May 2012 27 May 2012

Authors

  • AC

    Alistair Conkie

  • TO

    Thomas Okken

  • YK

    Yeon-Jun Kim

  • GD

    Giuseppe Di Fabbrizio

Links