TTS-Portuguese Corpus: a corpus for speech synthesis in Brazilian Portuguese.
Speech provides a natural way for human–computer interaction. In particular, speech synthesis systems are popular in different applications, such as personal assistants, GPS applications, screen readers and accessibility tools. However, not all languages are on the same level when in terms of resour...
| Publicado en: | Language Resources & Evaluation Vol. 56; no. 3; pp. 1043 - 1056 |
|---|---|
| Autores principales: | , , , , , , |
| Formato: | Artículo |
| Publicado: |
Springer Nature
Sep2022
|
| Materias: | |
| Acceso en línea: | Ver este registro en EBSCOhost |
| fields | @attributes: recordID: 1 pdfLink: plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=158609438&site=ehost-live header: @attributes: shortDbName: hlh uiTerm: 158609438 longDbName: Humanities International Complete uiTag: AN controlInfo: bkinfo: jinfo: jid: 1574020X 179V jtl: Language Resources & Evaluation issn: 1574020X maglogo: N pubinfo: dt: Sep2022 vid: 56 iid: 3 pid: 237 pub: Springer Nature artinfo: ui: 158609438 10.1007/s10579-021-09570-4 ppf: 1043 ppct: 13 formats: fmt: – @attributes: type: T – @attributes: type: P size: 409KB tig: atl: TTS-Portuguese Corpus: a corpus for speech synthesis in Brazilian Portuguese. aug: au: Casanova, Edresson Junior, Arnaldo Candido Shulby, Christopher Oliveira, Frederico Santos de Teixeira, João Paulo Ponti, Moacir Antonelli Aluísio, Sandra affil: Instituto de Ciências Matemáticas e de Computação, University of São Paulo, São Carlos, Brazil Federal University of Technology – Paraná (UTFPR), Medianeira, Brazil DefinedCrowd Corp., Seattle, USA Federal University of Mato Grosso, Cuiabá, Brazil Research Center in Digitalization and Intelligent Robotics (CEDRI) - Instituto Politecnico de Braganca, Bragança, Portugal su: Speech synthesis Portuguese language Human-computer interaction Deep learning Personal assistants sug: subj: Speech synthesis Portuguese language Human-computer interaction Deep learning Personal assistants keyword: Corpora Portuguese TTS ab: Speech provides a natural way for human–computer interaction. In particular, speech synthesis systems are popular in different applications, such as personal assistants, GPS applications, screen readers and accessibility tools. However, not all languages are on the same level when in terms of resources and systems for speech synthesis. This work consists of creating publicly available resources for Brazilian Portuguese in the form of a novel dataset along with deep learning models for end-to-end speech synthesis. Such dataset has 10.5 h from a single speaker, from which a Tacotron 2 model with the RTISI-LA vocoder presented the best performance, achieving a 4.03 MOS value. The obtained results are comparable to related works covering English language and the state-of-the-art in European Portuguese. pubtype: Academic Journal doctype: Article src: R language: English refInfo: copyright: @attributes: flag: Y custom: Language Resources & Evaluation is a copyright of Springer, 2022. All Rights Reserved. item: Language Resources & Evaluation holder: Springer Nature dt: @attributes: year: 2022 holdings: @attributes: islocal: N |
|---|