Reproducible Speech Research With the Artificial Intelligence–Ready PERCEPT Corpora.

Background: Publicly available speech corpora facilitate reproducible research by providing open-access data for participants who have consented/assented to data sharing among different research teams. Such corpora can also support clinical education, including perceptual training and training in th...

Descripción completa

Detalles Bibliográficos
Publicado en:Journal of Speech, Language & Hearing Research Vol. 66; no. 6; pp. 1986 - 2010
Autores principales: Benway, Nina R., Preston, Jonathan L., Hitchcock, Elaine, Rose, Yvan, Salekin, Asif, Liang, Wendy, McAllister, Tara
Formato: Artículo
Publicado: American Speech-Language-Hearing Association Jun2023
Materias:
Acceso en línea:Ver este registro en EBSCOhost
fields @attributes:
  recordID: 1
pdfLink:
plink: https://search.ebscohost.com/login.aspx?direct=true&db=ssf&AN=164422020&site=ehost-live
header:
  @attributes:
    shortDbName: ssf
    uiTerm: 164422020
    longDbName: Social Sciences Full Text (H.W. Wilson)
    uiTag: AN
  controlInfo:
    bkinfo:
    jinfo:
      jid:
        10924388
        1SM
      jtl: Journal of Speech, Language & Hearing Research
      issn: 10924388
      maglogo: N
    pubinfo:
      dt: Jun2023
      vid: 66
      iid: 6
      pid: 42
      pub: American Speech-Language-Hearing Association
    artinfo:
      ui:
        164422020
        10.1044/2023_JSLHR-22-00343
      ppf: 1986
      ppct: 24
      formats:
        fmt:
          @attributes:
            type: P
            size: 4.6MB
      tig:
        atl: Reproducible Speech Research With the Artificial Intelligence–Ready PERCEPT Corpora.
      aug:
        au:
          Benway, Nina R.
          Preston, Jonathan L.
          Hitchcock, Elaine
          Rose, Yvan
          Salekin, Asif
          Liang, Wendy
          McAllister, Tara
        affil:
          Department of Communication Sciences & Disorders, Syracuse University, NY.
          Haskins Laboratories, New Haven, CT.
          Department of Communication Sciences and Disorders, Montclair State University, NJ.
          Department of Linguistics, Memorial University, St. John’s, Newfoundland and Labrador, Canada.
          Department of Electrical Engineering and Computer Science, Syracuse University, NY.
          Department of Communicative Sciences and Disorders, New York University, NY.
      su:
        Speech disorders
        Artificial intelligence
        Phonetics
        Research evaluation
        Speech evaluation
        Articulation disorders
        Software architecture
        Descriptive statistics
        Data analysis software
        Evaluation
      sug:
        subj:
          Speech disorders
          Artificial intelligence
          Phonetics
          Computer systems design and related services (except video game design and development)
          Research evaluation
          Speech evaluation
          Articulation disorders
          Software architecture
          Descriptive statistics
          Data analysis software
          Evaluation
      ab: Background: Publicly available speech corpora facilitate reproducible research by providing open-access data for participants who have consented/assented to data sharing among different research teams. Such corpora can also support clinical education, including perceptual training and training in the use of speech analysis tools. Purpose: In this research note, we introduce the PERCEPT (Perceptual Error Rating for the Clinical Evaluation of Phonetic Targets) corpora, PERCEPT-R (Rhotics) and PERCEPT-GFTA (Goldman-Fristoe Test of Articulation), which together contain over 36 hr of speech audio (> 125,000 syllable, word, and phrase utterances) from children, adolescents, and young adults aged 6– 24 years with speech sound disorder (primarily residual speech sound disorders impacting /ɹ/) and age-matched peers. We highlight PhonBank as the repository for the corpora and demonstrate use of the associated speech analysis software, Phon, to query PERCEPT-R. A worked example of research with PERCEPT-R, suitable for clinical education and research training, is included as an appendix. Support for end users and information/descriptive statistics for future releases of the PERCEPT corpora can be found in a dedicated Slack channel. Finally, we discuss the potential for PERCEPT corpora to support the training of artificial intelligence clinical speech technology appropriate for use with children with speech sound disorders, the development of which has historically been constrained by the limited representation of either children or individuals with speech impairments in publicly available training corpora. Conclusions: We demonstrate the use of PERCEPT corpora, PhonBank, and Phon for clinical training and research questions appropriate to child citation speech. Increased use of these tools has the potential to enhance reproducibility in the study of speech development and disorders.
      pubtype: Academic Journal
      doctype: Article
      src: R
    language: English
    refInfo:
    copyright:
      @attributes:
        flag: N
    holdings:
      @attributes:
        islocal: N