JMIR Aging

Multimodal Dementia Prediction With Large Language Models: Cross-Attention Over Text, Audio, and Image

Figure 1. Proposed trimodal fusion architecture for dementia prediction.
Open the figure at full size
Figure 1. Proposed trimodal fusion architecture for dementia prediction.Proposed trimodal fusion architecture for dementia prediction.Agbavor et al.
Computational studyBiomarkers

The paper

Drexel University; Hong Kong Polytechnic University

JMIR Aging, 21 Sep 2026

doi.org/10.2196/93279PubMed 42766599