Mini-mental status examination phenotyping for Alzheimer's disease patients using both structured and narrative electronic health record features

Betina Idnay; Gongbo Zhang; Fangyi Chen; Casey N Ta; Matthew W Schelke; Karen Marder; Chunhua Weng

doi:10.1093/jamia/ocae274

Mini-mental status examination phenotyping for Alzheimer's disease patients using both structured and narrative electronic health record features

J Am Med Inform Assoc. 2025 Jan 1;32(1):119-128. doi: 10.1093/jamia/ocae274.

Authors

Betina Idnay¹, Gongbo Zhang¹, Fangyi Chen¹, Casey N Ta¹, Matthew W Schelke², Karen Marder², Chunhua Weng¹

Affiliations

¹ Department of Biomedical Informatics, Columbia University Irving Medical Center, New York, NY 10032, United States.
² Department of Neurology, Columbia University Irving Medical Center, New York, NY 10032, United States.

PMID: 39520712
PMCID: PMC11648712 (available on 2025-11-09)
DOI: 10.1093/jamia/ocae274

Abstract

Objective: This study aims to automate the prediction of Mini-Mental State Examination (MMSE) scores, a widely adopted standard for cognitive assessment in patients with Alzheimer's disease, using natural language processing (NLP) and machine learning (ML) on structured and unstructured EHR data.

Materials and methods: We extracted demographic data, diagnoses, medications, and unstructured clinical visit notes from the EHRs. We used Latent Dirichlet Allocation (LDA) for topic modeling and Term-Frequency Inverse Document Frequency (TF-IDF) for n-grams. In addition, we extracted meta-features such as age, ethnicity, and race. Model training and evaluation employed eXtreme Gradient Boosting (XGBoost), Stochastic Gradient Descent Regressor (SGDRegressor), and Multi-Layer Perceptron (MLP).

Results: We analyzed 1654 clinical visit notes collected between September 2019 and June 2023 for 1000 Alzheimer's disease patients. The average MMSE score was 20, with patients averaging 76.4 years old, 54.7% female, and 54.7% identifying as White. The best-performing model (ie, lowest root mean squared error (RMSE)) is MLP, which achieved an RMSE of 5.53 on the validation set using n-grams, indicating superior prediction performance over other models and feature sets. The RMSE on the test set was 5.85.

Discussion: This study developed a ML method to predict MMSE scores from unstructured clinical notes, demonstrating the feasibility of utilizing NLP to support cognitive assessment. Future work should focus on refining the model and evaluating its clinical relevance across diverse settings.

Conclusion: We contributed a model for automating MMSE estimation using EHR features, potentially transforming cognitive assessment for Alzheimer's patients and paving the way for more informed clinical decisions and cohort identification.

Keywords: Alzheimer’s disease; electronic health records; machine learning; natural language processing; phenotyping.

MeSH terms

Aged
Aged, 80 and over
Alzheimer Disease* / diagnosis
Electronic Health Records*
Female
Humans
Machine Learning*
Male
Mental Status and Dementia Tests*
Natural Language Processing*
Phenotype

Abstract

MeSH terms

Grants and funding