Discovery of Language Resources on the Web: Information Extraction from Heterogeneous Documents.

The present article is concerned with the problem of automatic database population via information extraction (IE) from web pages obtained from heterogeneous sources, such as those retrieved by a domain crawler. Specifically, we address the task of filling single multi-field templates from individua...

Descripción completa

Detalles Bibliográficos
Publicado en:Literary & Linguistic Computing Vol. 22; no. 3; pp. 329 - 344
Autores principales: Pekar, Viktor, Evans, Richard
Formato: Artículo
Publicado: Oxford University Press / USA Sep2007
Materias:
Acceso en línea:Ver este registro en EBSCOhost