A web-based Bengali news corpus for named entity recognition.

The rapid development of language resources and tools using machine learning techniques for less computerized languages requires appropriately tagged corpus. A tagged Bengali news corpus has been developed from the web archive of a widely read Bengali newspaper. A web crawler retrieves the web pages...

Descripción completa

Detalles Bibliográficos
Publicado en:Language Resources & Evaluation Vol. 42; no. 2; pp. 173 - 183
Autores principales: Ekbal, Asif, Bandyopadhyay, Sivaji
Formato: Artículo
Publicado: Springer Nature May2008
Materias:
Acceso en línea:Ver este registro en EBSCOhost