Automatically Calculated Context-Sensitive Features of Connected Speech Improve Prediction of Impairment in Alzheimer's Disease.

Purpose: Early detection is critical for effective management of Alzheimer's disease (AD) and other dementias. One promising approach for predicting AD status is to automatically calculate linguistic features from open-ended connected speech. Past work has focused on individual word-level features s...

Descripción completa

Detalles Bibliográficos
Publicado en:Journal of Speech, Language & Hearing Research Vol. 68; no. 11; pp. 5341 - 5363
Autores principales: Flick, Graham, Ostrand, Rachel
Formato: Artículo
Publicado: American Speech-Language-Hearing Association Nov2025
Materias:
Acceso en línea:Ver este registro en EBSCOhost
fields @attributes:
  recordID: 1
pdfLink:
plink: https://search.ebscohost.com/login.aspx?direct=true&db=ssf&AN=189241811&site=ehost-live
header:
  @attributes:
    shortDbName: ssf
    uiTerm: 189241811
    longDbName: Social Sciences Full Text (H.W. Wilson)
    uiTag: AN
  controlInfo:
    bkinfo:
    jinfo:
      jid:
        10924388
        1SM
      jtl: Journal of Speech, Language & Hearing Research
      issn: 10924388
      maglogo: N
    pubinfo:
      dt: Nov2025
      vid: 68
      iid: 11
      pid: 42
      pub: American Speech-Language-Hearing Association
    artinfo:
      ui:
        189241811
        10.1044/2025_JSLHR-24-00297
      ppf: 5341
      ppct: 22
      formats:
        fmt:
          @attributes:
            type: P
            size: 1.5MB
      tig:
        atl: Automatically Calculated Context-Sensitive Features of Connected Speech Improve Prediction of Impairment in Alzheimer's Disease.
      aug:
        au:
          Flick, Graham
          Ostrand, Rachel
        affil:
          Department of Psychology, New York University, NY
          Rotman Research Institute, Baycrest Centre, Toronto, Ontario, Canada
          IBM Research, Yorktown Heights, NY
      su:
        Linguistics
        Discourse analysis
        Speech evaluation
        Statistical models
        Pearson correlation (Statistics)
        Alzheimer's disease
        Research funding
        Secondary analysis
        Receiver operating characteristic curves
        Data analysis
        Multiple regression analysis
        Logistic regression analysis
        Descriptive statistics
        Natural language processing
        Research
        Statistics
      sug:
        subj:
          Linguistics
          Discourse analysis
          Speech evaluation
          Statistical models
          Pearson correlation (Statistics)
          Alzheimer's disease
          Research funding
          Secondary analysis
          Receiver operating characteristic curves
          Data analysis
          Multiple regression analysis
          Logistic regression analysis
          Descriptive statistics
          Natural language processing
          Research
          Statistics
      ab: Purpose: Early detection is critical for effective management of Alzheimer's disease (AD) and other dementias. One promising approach for predicting AD status is to automatically calculate linguistic features from open-ended connected speech. Past work has focused on individual word-level features such as part of speech counts, total word production, and lexical richness, with less emphasis on measuring the relationship between words and the context in which they are produced. Here, we assessed whether linguistic features that take into account where a word was produced in the discourse context improved the ability to predict AD patients' Mini-Mental State Examination (MMSE) scores and classify AD patients from healthy control participants. Method: Seventeen linguistic features were automatically computed from transcriptions of spoken picture descriptions from individuals with probable or possible AD (n = 176 transcripts). This included 12 word-level features (e.g., part of speech counts) and five features capturing contextual word choices (linguistic surprisal, computed from a computational large language model, and properties of words produced following filled pauses). We examined whether (a) the full set jointly predicted MMSE scores, (b) the addition of contextual features improved prediction, and (c) linguistic features could classify AD patients (n = 130) versus healthy participants (n = 93). Results: Linguistic features accurately predicted MMSE scores in individuals with probable or possible AD and successfully identified up to 87% of AD participants versus healthy controls. Statistical models that contained linguistic sur-prisal (a contextual feature) performed better than those that included only word-level and demographic features. Overall, AD patients with lower MMSE scores produced more empty words, fewer nouns and definite articles, and words that were higher frequency yet more surprising given the previous context. Conclusion: These results provide novel evidence that metrics related to con-textualized word choices, particularly the surprisal of an individual's words, capture variance in degree of cognitive decline in AD.
      pubtype: Academic Journal
      doctype: Article
      src: R
    language: English
    refInfo:
    copyright:
      @attributes:
        flag: N
    holdings:
      @attributes:
        islocal: N