Involving language professionals in the evaluation of machine translation.

Significant breakthroughs in machine translation (MT) only seem possible if human translators are taken into the loop. While automatic evaluation and scoring mechanisms such as BLEU have enabled the fast development of systems, it is not clear how systems can meet real-world (quality) requirements i...

Descripción completa

Detalles Bibliográficos
Publicado en:Language Resources & Evaluation Vol. 48; no. 4; pp. 541 - 560
Autores principales: Popović, Maja, Avramidis, Eleftherios, Burchardt, Aljoscha, Hunsicker, Sabine, Schmeier, Sven, Tscherwinka, Cindy, Vilar, David, Uszkoreit, Hans
Formato: Artículo
Publicado: Springer Nature Dec2014
Materias:
Acceso en línea:Ver este registro en EBSCOhost
Descripción
Sumario:Significant breakthroughs in machine translation (MT) only seem possible if human translators are taken into the loop. While automatic evaluation and scoring mechanisms such as BLEU have enabled the fast development of systems, it is not clear how systems can meet real-world (quality) requirements in industrial translation scenarios today. The taraXŰ project has paved the way for wide usage of multiple MT outputs through various feedback loops in system development. The project has integrated human translators into the development process thus collecting feedback for possible improvements. This paper describes results from detailed human evaluation. Performance of different types of translation systems has been compared and analysed via ranking, error analysis and post-editing.