Automatic dialect identification system for Kannada language using single and ensemble SVM algorithms.

In this paper, an automatic dialect identification (ADI) system is proposed by extracting spectral and prosodic features for Kannada language. A new dialect dataset is collected from native speakers of Kannada language (A Dravidian language). This dataset includes five distinct dialects of Kannada l...

Full description

Bibliographic Details
Published in:Language Resources & Evaluation Vol. 54; no. 2; pp. 553 - 586
Main Authors: Chittaragi, Nagaratna B., Koolagudi, Shashidhar G.
Format: Article
Published: Springer Nature Jun2020
Subjects:
Online Access:View this record in EBSCOhost
fields @attributes:
  recordID: 1
pdfLink:
plink: https://search.ebscohost.com/login.aspx?direct=true&db=hlh&AN=143152359&site=ehost-live
header:
  @attributes:
    shortDbName: hlh
    uiTerm: 143152359
    longDbName: Humanities International Complete
    uiTag: AN
  controlInfo:
    bkinfo:
    jinfo:
      jid:
        1574020X
        179V
      jtl: Language Resources & Evaluation
      issn: 1574020X
      maglogo: N
    pubinfo:
      dt: Jun2020
      vid: 54
      iid: 2
      pid: 237
      pub: Springer Nature
    artinfo:
      ui:
        143152359
        10.1007/s10579-019-09481-5
      ppf: 553
      ppct: 33
      formats:
        fmt:
          – @attributes:
              type: T
          – @attributes:
              type: P
              size: 933KB
      tig:
        atl: Automatic dialect identification system for Kannada language using single and ensemble SVM algorithms.
      aug:
        au:
          Chittaragi, Nagaratna B.
          Koolagudi, Shashidhar G.
        affil:
          Dept. of Computer Science and Engg., National Institute of Technology Karnataka, Surathkal, India
          Dept. of Information Science and Engg., Siddaganga Institute of Technology, Tumkur, Karnataka, India
      su:
        Automatic identification
        System identification
        Facial expression
        Support vector machines
        Native language
        Parameter identification
        Karnataka (India)
      sug:
        subj:
          Karnataka (India)
          Automatic identification
          System identification
          Facial expression
          Support vector machines
          Native language
          Parameter identification
      keyword:
        Derived features
        Dialect identification
        Ensemble SVM
        IViE dialect dataset
        Kannada dialect dataset
        Single SVM
        Spectral and prosodic features
      ab: In this paper, an automatic dialect identification (ADI) system is proposed by extracting spectral and prosodic features for Kannada language. A new dialect dataset is collected from native speakers of Kannada language (A Dravidian language). This dataset includes five distinct dialects of Kannada language representing five geographical regions of Karnataka state. Investigation of the significance of spectral and prosodic variations on five Kannada dialects is carried out. Mel-frequency cepstral coefficients (MFCCs), spectral flux, and entropy are used as representatives of spectral features. Besides, pitch and energy features are extracted as representatives of prosodic parameters for identification of dialects. These raw feature vectors are further processed to get a new derived feature vectors by using statistical processing. In this paper, a single classifier based multi-class support vector machine (SVM) and multiple classifier based ensemble SVM (ESVM) techniques are employed for classification of dialects. The effectiveness and performance evaluation of the explored features are carried out on newly collected Kannada speech corpus, with five Kannada dialects and internationally known standard Intonation Variation in English (IViE) dataset with nine British English dialects. Experimental results have demonstrated that the derived feature vectors performs better when compared to raw feature vectors. However, ESVM technique has demonstrated better performance over a single SVM. Spectral and prosodic features have resulted individually with the dialect recognition performance of 83.12% and 44.52% respectively. Further, the complementary nature of both spectral and prosodic features is evaluated by combining both feature vectors for dialect recognition. However, an increase in dialect recognition performance of about 86.25% is observed. This indicates the existence of complementary dialect specific evidence with spectral and prosodic features. The experiments conducted on standard IViE corpus have shown a higher recognition rate of 91.38% using ESVM. Proposed ADI systems with derived features have shown better performance over the state-of-the-art i-vector feature based systems on both datasets.
      pubtype: Academic Journal
      doctype: Article
      src: R
    language: English
    refInfo:
    copyright:
      @attributes:
        flag: Y
      custom: Language Resources & Evaluation is a copyright of Springer, 2020. All Rights Reserved.
      item: Language Resources & Evaluation
      holder: Springer Nature
      dt:
        @attributes:
          year: 2020
    holdings:
      @attributes:
        islocal: N