Spanish corpora for sentiment analysis: a survey.

Corpora play an important role when training machine learning systems for sentiment analysis. However, Spanish is underrepresented in these corpora, as most primarily include English texts. This paper describes 20 Spanish-language text corpora—collected to support different tasks related to sentimen...

Descripción completa

Detalles Bibliográficos
Publicado en:Language Resources & Evaluation Vol. 54; no. 2; pp. 303 - 341
Autores principales: Navas-Loro, María, Rodríguez-Doncel, Víctor
Formato: Artículo
Publicado: Springer Nature Jun2020
Materias:
Acceso en línea:Ver este registro en EBSCOhost
fields @attributes:
  recordID: 1
pdfLink:
plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=143152358&site=ehost-live
header:
  @attributes:
    shortDbName: hlh
    uiTerm: 143152358
    longDbName: Humanities International Complete
    uiTag: AN
  controlInfo:
    bkinfo:
    jinfo:
      jid:
        1574020X
        179V
      jtl: Language Resources & Evaluation
      issn: 1574020X
      maglogo: N
    pubinfo:
      dt: Jun2020
      vid: 54
      iid: 2
      pid: 237
      pub: Springer Nature
    artinfo:
      ui:
        143152358
        10.1007/s10579-019-09470-8
      ppf: 303
      ppct: 38
      formats:
        fmt:
          – @attributes:
              type: T
          – @attributes:
              type: P
              size: 539KB
      tig:
        atl: Spanish corpora for sentiment analysis: a survey.
      aug:
        au:
          Navas-Loro, María
          Rodríguez-Doncel, Víctor
        affil: Ontology Engineering Group, Universidad Politécnica de Madrid, Madrid, Spain
      su:
        Sentiment analysis
        Corpora
        Task analysis
        System analysis
        Machine learning
      sug:
        subj:
          Sentiment analysis
          Corpora
          Task analysis
          System analysis
          Machine learning
      keyword:
        Emotion
        Opinion mining
        Polarity
      ab: Corpora play an important role when training machine learning systems for sentiment analysis. However, Spanish is underrepresented in these corpora, as most primarily include English texts. This paper describes 20 Spanish-language text corpora—collected to support different tasks related to sentiment analysis, ranging from polarity to emotion categorization. We present a brand-new framework for the characterization of corpora. This includes a number of features to help analyze resources at both corpus level and document level. This survey—besides depicting the overall landscape of corpora in Spanish—supports sentiment analysis practitioners with the task of selecting the most suitable resources.
      pubtype: Academic Journal
      doctype: Article
      src: R
    language: English
    refInfo:
    copyright:
      @attributes:
        flag: Y
      custom: Language Resources & Evaluation is a copyright of Springer, 2020. All Rights Reserved.
      item: Language Resources & Evaluation
      holder: Springer Nature
      dt:
        @attributes:
          year: 2020
    holdings:
      @attributes:
        islocal: N