Collecting and evaluating speech recognition corpora for 11 South African languages.
We describe the Lwazi corpus for automatic speech recognition (ASR), a new telephone speech corpus which contains data from the eleven official languages of South Africa. Because of practical constraints, the amount of speech per language is relatively small compared to major corpora in world langua...
| Publicado en: | Language Resources & Evaluation Vol. 45; no. 3; pp. 289 - 310 |
|---|---|
| Autores principales: | , , , |
| Formato: | Artículo |
| Publicado: |
Springer Nature
Aug2011
|
| Materias: | |
| Acceso en línea: | Ver este registro en EBSCOhost |