SemGraph: Extracting Keyphrases Following a Novel Semantic Graph-Based Approach.
Keyphrases represent the main topics a text is about. In this article, we introduce SemGraph, an unsupervised algorithm for extracting keyphrases from a collection of texts based on a semantic relationship graph. The main novelty of this algorithm is its ability to identify semantic relationships be...
| Publicado en: | Journal of the Association for Information Science & Technology Vol. 67; no. 1; pp. 71 - 83 |
|---|---|
| Autores principales: | , , |
| Formato: | equations & formulas research tables/charts Journal Article |
| Publicado: |
Wiley-Blackwell
Jan2016
|
| Acceso en línea: | Ver este registro en EBSCOhost |
| fields | @attributes: recordID: 1 pdfLink: plink: https://search.ebscohost.com/login.aspx?direct=true&db=ccm&AN=112228402&site=ehost-live header: @attributes: shortDbName: ccm uiTerm: 112228402 longDbName: CINAHL Complete uiTag: AN controlInfo: bkinfo: dissinfo: jinfo: jid: 23301635 H6JN jtl: Journal of the Association for Information Science & Technology issn: 23301635 maglogo: N pubinfo: dt: Jan2016 vid: 67 iid: 1 pid: 480 pub: Wiley-Blackwell place: Malden, Massachusetts artinfo: ui: 112228402 112228402 112228402 10.1002/asi.23365 112228402 ppf: 71 ppct: 12 formats: tig: atl: SemGraph: Extracting Keyphrases Following a Novel Semantic Graph-Based Approach. aug: au: Martinez‐Romo, Juan Araujo, Lourdes Duque Fernandez, Andres affil: NLP & IR Group, Dpto. Lenguajes y Sistemas Informáticos, Universidad Nacional de Educación a Distancia (UNED), Juan del Rosal, 16. 28040, Madrid Spain sug: subj: Data Mining Methods Semantics Human Algorithms Funding Source ab: Keyphrases represent the main topics a text is about. In this article, we introduce SemGraph, an unsupervised algorithm for extracting keyphrases from a collection of texts based on a semantic relationship graph. The main novelty of this algorithm is its ability to identify semantic relationships between words whose presence is statistically significant. Our method constructs a co-occurrence graph in which words appearing in the same document are linked, provided their presence in the collection is statistically significant with respect to a null model. Furthermore, the graph obtained is enriched with information from WordNet. We have used the most recent and standardized benchmark to evaluate the system ability to detect the keyphrases that are part of the text. The result is a method that achieves an improvement of 5.3% and 7.28% in F measure over the two labeled sets of keyphrases used in the evaluation of SemEval-2010. pubtype: Academic Journal doctype: equations & formulas research tables/charts Journal Article ougenre: Article language: English refInfo: holdings: @attributes: islocal: N |
|---|