Back to results

Universidade Federal do Rio de Janeiro

Comparison between rule-based and data-driven natural language processing algorithms for Brazilian Portuguese speech synthesis

Abstract

dc:description.abstract

Due to the exponential growth in the use of computers, personal digital assistants and smartphones, the development of Text-to-Speech (TTS) systems have become highly demanded during the last years. An important part of these systems is the Text Analysis block, that converts the input text into linguistic specifications that are going to be used to generate the final speech waveform. The Natural Language Processing algorithms presented in this block are crucial to the quality of the speech generated by synthesizers. These algorithms are responsible for important tasks such as Grapheme-to-Phoneme Conversion, Syllabification and Stress Determination. For Brazilian Portuguese (BP), solutions for the algorithms presented in the Text Analysis block have been focused in rule-based approaches. These algorithms perform well for BP but have many disadvantages. On the other hand, there is still no research to evaluate and analyze the performance of data-driven approaches that reach state-of-the-art results for complex languages, such as English. So, in this work, we compare different data-driven approaches and rule-based approaches for NLP algorithms presented in a TTS system. Moreover, we propose, as a novel application, the use of Sequence-to-Sequence models as solution for the Syllabification and Stress Determination problems. As a brief summary of the results obtained, we show that data-driven algorithms can achieve state-of-the-art performance for the NLP algorithms presented in the Text Analysis block of a BP TTS system.

Degree

thesis:*
Grantor dc:publisher
Universidade Federal do Rio de Janeiro
Year dc:date.issued
2017

Author and committee

dc:creator, dc:contributor.*
Author dc:creator
  • Vecchietti, Luiz Felipe Santos
Advisor dc:contributor.advisor
  • Resende Junior, Fernando Gil Vianna

Subjects

dc:subject × 3

Rights

dc:rights
Statement dc:rights
  • Acesso Aberto
Language dc:language
eng

Identifiers

dc:identifier.*
Handle dc:identifier.uri
http://hdl.handle.net/11422/6486
OAI identifier oai:identifier
oai:pantheon.ufrj.br:11422/6486

Chain of custody

source
Harvested from
Brazil UERJ
Base URL
pantheon.ufrj.br/oai/request
Last updated
2026-07-24
Source record
OAI-PMH GetRecord
citation

Vecchietti, Luiz Felipe Santos. Comparison between rule-based and data-driven natural language processing algorithms for Brazilian Portuguese speech synthesis. Universidade Federal do Rio de Janeiro, 2017. http://hdl.handle.net/11422/6486