From legacy encodings to Unicode: the graphical and logical principles in the scripts of South Asia.
Much electronic text in the languages of South Asia has been published on the Internet. However, while Unicode has emerged as the favoured encoding system of corpus and computational linguists, most South Asian language data on the web uses one of a wide range of non-standard legacy encodings. This...
| Publicado en: | Language Resources & Evaluation Vol. 41; no. 1; pp. 1 - 26 |
|---|---|
| Autor principal: | |
| Formato: | Artículo |
| Publicado: |
Springer Nature
Feb2007
|
| Materias: | |
| Acceso en línea: | Ver este registro en EBSCOhost |