Involving language professionals in the evaluation of machine translation.
Significant breakthroughs in machine translation (MT) only seem possible if human translators are taken into the loop. While automatic evaluation and scoring mechanisms such as BLEU have enabled the fast development of systems, it is not clear how systems can meet real-world (quality) requirements i...
| Published in: | Language Resources & Evaluation Vol. 48; no. 4; pp. 541 - 560 |
|---|---|
| Main Authors: | , , , , , , , |
| Format: | Article |
| Published: |
Springer Nature
Dec2014
|
| Subjects: | |
| Online Access: | View this record in EBSCOhost |
| fields | @attributes: recordID: 1 pdfLink: plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=99708560&site=ehost-live header: @attributes: shortDbName: hlh uiTerm: 99708560 longDbName: Humanities International Complete uiTag: AN controlInfo: bkinfo: jinfo: jid: 1574020X 179V jtl: Language Resources & Evaluation issn: 1574020X maglogo: N pubinfo: dt: Dec2014 vid: 48 iid: 4 pid: 237 pub: Springer Nature artinfo: ui: 99708560 10.1007/s10579-014-9286-z ppf: 541 ppct: 19 formats: fmt: @attributes: type: P size: 267KB tig: atl: Involving language professionals in the evaluation of machine translation. aug: au: Popović, Maja Avramidis, Eleftherios Burchardt, Aljoscha Hunsicker, Sabine Schmeier, Sven Tscherwinka, Cindy Vilar, David Uszkoreit, Hans affil: DFKI - Language Technology Lab, Berlin Germany euroscript Deutschland, Berlin Germany su: Computational linguistics Translators Linguists Error analysis in mathematics Translating & interpreting sug: subj: Computational linguistics Translators Linguists Error analysis in mathematics Translating & interpreting keyword: Error analysis Human evaluation Machine translation ab: Significant breakthroughs in machine translation (MT) only seem possible if human translators are taken into the loop. While automatic evaluation and scoring mechanisms such as BLEU have enabled the fast development of systems, it is not clear how systems can meet real-world (quality) requirements in industrial translation scenarios today. The taraXŰ project has paved the way for wide usage of multiple MT outputs through various feedback loops in system development. The project has integrated human translators into the development process thus collecting feedback for possible improvements. This paper describes results from detailed human evaluation. Performance of different types of translation systems has been compared and analysed via ranking, error analysis and post-editing. pubtype: Academic Journal doctype: Article src: R language: English refInfo: copyright: @attributes: flag: Y custom: Language Resources & Evaluation is a copyright of Springer, 2014. All Rights Reserved. item: Language Resources & Evaluation holder: Springer Nature dt: @attributes: year: 2014 holdings: @attributes: islocal: N |
|---|