How different is different? Systematically identifying distribution shifts and their impacts in NER datasets.

When processing natural language, we are frequently confronted with the problem of distribution shift. For example, using a model trained on a news corpus to subsequently process legal text exhibits reduced performance. While this problem is well-known, to this point, there has not been a systematic...

Descripción completa

Detalles Bibliográficos
Publicado en:Language Resources & Evaluation Vol. 59; no. 2; pp. 1111 - 1151
Autores principales: Li, Xue, Groth, Paul
Formato: Artículo
Publicado: Springer Nature Jun2025
Materias:
Acceso en línea:Ver este registro en EBSCOhost