Speech model flags cognitive decline and generates explanations
Tested across six dataset and task conditions, the bilingual framework outperformed three baselines while directly generating natural-language explanations for clinicians.
In human speech datasets from individuals with Alzheimer's disease and mild cognitive impairment, researchers evaluated a bilingual speech large language model designed for automated cognitive screening. The preprint describes a framework that processes raw spontaneous speech directly, avoiding transcription errors and capturing subtle acoustic and prosodic features alongside semantics. Using the newly collected PUTH-AD dataset and multiple open-source corpora, the model learned to simultaneously classify cognitive status and generate natural-language explanations. Across six dataset and task conditions, the system achieved higher average accuracy and AUROC than three representative baselines. It also maintained classification accuracy on an unseen cognitive task subset without task-specific fine-tuning. Furthermore, clinician reviews verified that the generated diagnostic explanations were clinically relevant and largely consistent with underlying speech evidence.
Why it matters
Subtle linguistic and acoustic shifts often appear early in mild cognitive impairment and Alzheimer's disease. Scalable, explainable tools could support accessible screening to catch age-related cognitive decline earlier than resource-intensive traditional diagnostics.
Caveats
The study is a preprint that has not yet undergone peer review. Further prospective validation in everyday healthcare settings is needed to confirm the tool's clinical utility.
The paper
Explainable and Generalisable LLM-based Cognitive Decline Detection with Spontaneous Speech
Cui Z, Wu W, Shi C et al.
arXiv · 28 Sep 2026 · Preprint, not yet peer-reviewed

