DeepMine-multi-TTS: a Persian speech corpus for multi-speaker text-to-speech.
Speech synthesis has made significant progress in recent years thanks to deep neural networks (DNNs). However, one of the challenges of DNN-based models is the requirement for large and diverse data, which limits their applicability to many languages and domains. To date, no multi-speaker text-to-sp...
| Publicado en: | Language Resources & Evaluation Vol. 59; no. 3; pp. 2245 - 2265 |
|---|---|
| Autores principales: | , , |
| Formato: | Conference Paper/Materials |
| Publicado: |
Springer Nature
Sep2025
|
| Materias: | |
| Acceso en línea: | Ver este registro en EBSCOhost |