TTS-Portuguese Corpus: a corpus for speech synthesis in Brazilian Portuguese.

Speech provides a natural way for human–computer interaction. In particular, speech synthesis systems are popular in different applications, such as personal assistants, GPS applications, screen readers and accessibility tools. However, not all languages are on the same level when in terms of resour...

Descripción completa

Detalles Bibliográficos
Publicado en:Language Resources & Evaluation Vol. 56; no. 3; pp. 1043 - 1056
Autores principales: Casanova, Edresson, Junior, Arnaldo Candido, Shulby, Christopher, Oliveira, Frederico Santos de, Teixeira, João Paulo, Ponti, Moacir Antonelli, Aluísio, Sandra
Formato: Artículo
Publicado: Springer Nature Sep2022
Materias:
Acceso en línea:Ver este registro en EBSCOhost
fields @attributes:
  recordID: 1
pdfLink:
plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=158609438&site=ehost-live
header:
  @attributes:
    shortDbName: hlh
    uiTerm: 158609438
    longDbName: Humanities International Complete
    uiTag: AN
  controlInfo:
    bkinfo:
    jinfo:
      jid:
        1574020X
        179V
      jtl: Language Resources & Evaluation
      issn: 1574020X
      maglogo: N
    pubinfo:
      dt: Sep2022
      vid: 56
      iid: 3
      pid: 237
      pub: Springer Nature
    artinfo:
      ui:
        158609438
        10.1007/s10579-021-09570-4
      ppf: 1043
      ppct: 13
      formats:
        fmt:
          – @attributes:
              type: T
          – @attributes:
              type: P
              size: 409KB
      tig:
        atl: TTS-Portuguese Corpus: a corpus for speech synthesis in Brazilian Portuguese.
      aug:
        au:
          Casanova, Edresson
          Junior, Arnaldo Candido
          Shulby, Christopher
          Oliveira, Frederico Santos de
          Teixeira, João Paulo
          Ponti, Moacir Antonelli
          Aluísio, Sandra
        affil:
          Instituto de Ciências Matemáticas e de Computação, University of São Paulo, São Carlos, Brazil
          Federal University of Technology – Paraná (UTFPR), Medianeira, Brazil
          DefinedCrowd Corp., Seattle, USA
          Federal University of Mato Grosso, Cuiabá, Brazil
          Research Center in Digitalization and Intelligent Robotics (CEDRI) - Instituto Politecnico de Braganca, Bragança, Portugal
      su:
        Speech synthesis
        Portuguese language
        Human-computer interaction
        Deep learning
        Personal assistants
      sug:
        subj:
          Speech synthesis
          Portuguese language
          Human-computer interaction
          Deep learning
          Personal assistants
      keyword:
        Corpora
        Portuguese
        TTS
      ab: Speech provides a natural way for human–computer interaction. In particular, speech synthesis systems are popular in different applications, such as personal assistants, GPS applications, screen readers and accessibility tools. However, not all languages are on the same level when in terms of resources and systems for speech synthesis. This work consists of creating publicly available resources for Brazilian Portuguese in the form of a novel dataset along with deep learning models for end-to-end speech synthesis. Such dataset has 10.5 h from a single speaker, from which a Tacotron 2 model with the RTISI-LA vocoder presented the best performance, achieving a 4.03 MOS value. The obtained results are comparable to related works covering English language and the state-of-the-art in European Portuguese.
      pubtype: Academic Journal
      doctype: Article
      src: R
    language: English
    refInfo:
    copyright:
      @attributes:
        flag: Y
      custom: Language Resources & Evaluation is a copyright of Springer, 2022. All Rights Reserved.
      item: Language Resources & Evaluation
      holder: Springer Nature
      dt:
        @attributes:
          year: 2022
    holdings:
      @attributes:
        islocal: N