Towards Fully Automatic Annotation of Audio Books for TTS
Olivier Boeffard, Laure Charonnat, Sébastien Le Maguer, Damien Lolive
Abstract
Building speech corpora is a first and crucial step for every text-to-speech synthesis system. Nowadays, the use of statistical models implies the use of huge sized corpora that need to be recorded, transcribed, annotated and segmented to be usable. The variety of corpora necessary for recent applications (content, style, etc.) makes the use of existing digital audio resources very attractive. Among all available resources, audiobooks, considering their quality, are interesting. Considering this framework, we propose a complete acquisition, segmentation and annotation chain for audiobooks that tends to be fully automatic. The proposed process relies on a data structure, Roots, that establishes the relations between the different annotation levels represented as sequences of items. This methodology has been applied successfully on 11 hours of speech extracted from an audiobook. A manual check, on a part of the corpus, shows the efficiency of the process.- Anthology ID:
- L12-1363
- Volume:
- Proceedings of the Eighth International Conference on Language Resources and Evaluation (LREC'12)
- Month:
- May
- Year:
- 2012
- Address:
- Istanbul, Turkey
- Editors:
- Nicoletta Calzolari, Khalid Choukri, Thierry Declerck, Mehmet Uğur Doğan, Bente Maegaard, Joseph Mariani, Asuncion Moreno, Jan Odijk, Stelios Piperidis
- Venue:
- LREC
- SIG:
- Publisher:
- European Language Resources Association (ELRA)
- Note:
- Pages:
- 975–980
- Language:
- URL:
- http://www.lrec-conf.org/proceedings/lrec2012/pdf/632_Paper.pdf
- DOI:
- Cite (ACL):
- Olivier Boeffard, Laure Charonnat, Sébastien Le Maguer, and Damien Lolive. 2012. Towards Fully Automatic Annotation of Audio Books for TTS. In Proceedings of the Eighth International Conference on Language Resources and Evaluation (LREC'12), pages 975–980, Istanbul, Turkey. European Language Resources Association (ELRA).
- Cite (Informal):
- Towards Fully Automatic Annotation of Audio Books for TTS (Boeffard et al., LREC 2012)
- PDF:
- http://www.lrec-conf.org/proceedings/lrec2012/pdf/632_Paper.pdf