Outperforming Multilingual Models: Character-level Processing and Custom Word Embeddings for Old English.
We evaluate a Stanza-based pipeline for Old English that combines character-level processing with language-specific word embeddings derived from the Dictionary of Old English Corpus. On a 25,000-word dataset annotated with Universal Dependencies, character-level models yield consistent gains over to...
| Publicado en: | International Journal of Humanities & Arts Computing: A Journal of Digital Humanities Vol. 20; no. 1; pp. 18 - 36 |
|---|---|
| Autores principales: | , |
| Formato: | Artículo |
| Publicado: |
Edinburgh University Press
Mar2026
|
| Materias: | |
| Acceso en línea: | Ver este registro en EBSCOhost |