A large and evolving cognate database.
We present CogNet, a large-scale, automatically-built database of sense-tagged cognates—words of common origin and meaning across languages. CogNet is continuously evolving: its current version contains over 8 million cognate pairs over 338 languages and 35 writing systems, with new releases already...
| Publicado en: | Language Resources & Evaluation Vol. 56; no. 1; pp. 165 - 190 |
|---|---|
| Autores principales: | , , |
| Formato: | Artículo |
| Publicado: |
Springer Nature
Mar2022
|
| Materias: | |
| Acceso en línea: | Ver este registro en EBSCOhost |
| fields | @attributes: recordID: 1 pdfLink: plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=155080015&site=ehost-live header: @attributes: shortDbName: hlh uiTerm: 155080015 longDbName: Humanities International Complete uiTag: AN controlInfo: bkinfo: jinfo: jid: 1574020X 179V jtl: Language Resources & Evaluation issn: 1574020X maglogo: N pubinfo: dt: Mar2022 vid: 56 iid: 1 pid: 237 pub: Springer Nature artinfo: ui: 155080015 10.1007/s10579-021-09544-6 ppf: 165 ppct: 25 formats: fmt: – @attributes: type: T – @attributes: type: P size: 805KB tig: atl: A large and evolving cognate database. aug: au: Batsuren, Khuyagbaatar Bella, Gábor Giunchiglia, Fausto affil: Department of Information and Computer Science, National University of Mongolia, Ikh surguuliin gudamj 1, 14200, Ulaanbaatar, Mongolia Department of Information Engineering and Computer Science, University of Trento, via Sommarive 5, 38123, Trento, Italy College of Computer Science and Technology, Jilin University, Changchun, China su: Etymology Databases Knowledge base Data analysis Quantitative research sug: subj: Etymology Databases Knowledge base Data analysis Quantitative research keyword: Cognate Lexical database Lexical semantics ab: We present CogNet, a large-scale, automatically-built database of sense-tagged cognates—words of common origin and meaning across languages. CogNet is continuously evolving: its current version contains over 8 million cognate pairs over 338 languages and 35 writing systems, with new releases already in preparation. The paper presents the algorithm and input resources used for its computation, an evaluation of the result, as well as a quantitative analysis of cognate data leading to novel insights on language diversity. Furthermore, as an example on the use of large-scale cross-lingual knowledge bases for improving the quality of multilingual applications, we present a case study on the use of CogNet for bilingual lexicon induction in the framework of cross-lingual transfer learning. pubtype: Academic Journal doctype: Article src: R language: English refInfo: copyright: @attributes: flag: Y custom: Language Resources & Evaluation is a copyright of Springer, 2022. All Rights Reserved. item: Language Resources & Evaluation holder: Springer Nature dt: @attributes: year: 2022 holdings: @attributes: islocal: N |
|---|