cs.MMOct 7, 2026

MemoCare: An Interactive Multimodal Mobile System for Automated Cognitive Screening

Authors: Duy-Cat Can, Mau Minh Phuc Le, Tuan-Khoa Hoang, Hai-Dang Nguyen, Trung-Hieu Do, Dang Minh Ly, Minh-Duc Nguyen, Nghia TT Hoang, +5 more

Organizations: Lausanne University Hospital, Switzerland · Faculty of Biology and Medicine, University of Lausanne, Switzerland · VNU University of Engineering and Technology, Vietnam · VinUni-Illinois Smart Health Center, VinUniversity, Hanoi, Vietnam · University of Science, VNU-HCM, Vietnam · Hanoi Medical University, Vietnam · National Geriatric Hospital, Vietnam · Department of Neurology, Military Hospital 175, Vietnam · Department of Neurology, School of Medicine, University of Medicine and Pharmacy at Ho Chi Minh City, Vietnam · International University, VNU-HCM, Vietnam · Vietnam National University Ho Chi Minh City, Vietnam

Abstract

MemoCare is an interactive mobile system for automated multimodal cognitive screening. A React Native application combines spoken responses, temporal and spatial orientation, touchscreen actions, and visuoconstruction in complete English and Vietnamese workflows. Speech is transcribed by Google Speech-to-Text and scored locally with deterministic task-specific natural language processing rules; GPS coordinates are resolved by the MemoCare spatial module before answer matching; touch tasks are scored from interaction events; and the drawing task uses a three-model convolutional neural network consensus with separate visual interpretation. Software tests pass 151/151 predefined cases across speech/language, spatial-answer, and touch-interaction scoring, while spatial regression passes 48/48 four-country coordinate-resolution cases. For the drawing module, validation-selected ShuffleNetV2 x1.5 achieved 91.33% mean balanced accuracy and 78.87% exact three-criterion accuracy on a locked 71-image test set. Four clinician co-authors additionally inspected the end-to-end workflow, yielding a pooled median rating of 4/5 across eight criteria, with item-level medians ranging from 3 to 4.5. At MMM, attendees can directly try a shortened multimodal screening workflow and inspect automatic item-level and total scoring.

Figures & tables

Explore similar work

CardsList
  1. Explainable and Generalisable LLM-based Cognitive Decline Detection with Spontaneous Speech

    Sep 28, 2026Ziyun Cui, Wen Wu, Chuan Shi +10Cognitive DiagnosisScreening

  2. M3^3Exam: Benchmarking Multimodal Memory for Realistic User-Agent Interactions

    Jun 5, 2026Zhengjun Huang, Wenxuan Liu, Zhoujin Tian +6Multimodal MemoryModalities

  3. Toward Generalizable Cognitive Impairment Detection with Speech-Based Multimodal Large Language Models

    Jul 23, 2026Yingchao Huang, Xin Wang, Yuhan Su +1Mild Cognitive ImpairmentAlzheimer