paper-with-me

Papers

Towards Synthesizing Normative Data for Cognitive Assessments Using Generative Multimodal Large Language Models

2025-08-25 · Victoria Yan, Honor Chotkowski, Fengran Wang, Xinhui Li, Carl Yang, Jiaying Lu, Runze Yan, Xiao Hu, Alex Fedorov arxiv

Cognitive assessments require normative data as essential benchmarks for evaluating individual performance. Hence, developing new cognitive tests based on novel image stimuli is challenging due to the lack of readily available normative data. Traditional data collection methods are costly, time-consuming, and infrequently updated, limiting their practical utility. Recent advancements in generative multimodal large language models (MLLMs) offer a new approach to generate synthetic normative data from existing cognitive test images. We investigated the feasibility of using MLLMs, specifically GPT-4o and GPT-4o-mini, to synthesize normative textual responses for established image-based cognitive assessments, such as the "Cookie Theft" picture description task. Two distinct prompting strategies-naive prompts with basic instructions and advanced prompts enriched with contextual guidance-were evaluated. Responses were analyzed using embeddings to assess their capacity to distinguish diagnostic groups and demographic variations. Performance metrics included BLEU, ROUGE, BERTScore, and an LLM-as-a-judge evaluation. Advanced prompting strategies produced synthetic responses that more effectively distinguished between diagnostic groups and captured demographic diversity compared to naive prompts. Superior models generated responses exhibiting higher realism and diversity. BERTScore emerged as the most reliable metric for contextual similarity assessment, while BLEU was less effective for evaluating creative outputs. The LLM-as-a-judge approach provided promising preliminary validation results. Our study demonstrates that generative multimodal LLMs, guided by refined prompting methods, can feasibly generate robust synthetic normative data for existing cognitive tests, thereby laying the groundwork for developing novel image-based cognitive assessments without the traditional limitations.

📄 PDF Abstract BibTeX arXiv:2508.17675

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Anatomy-Guided Surface Diffusion Model for Alzheimer's Disease Normative Modeling

2024-03-07 · Jianwei Zhang, Yonggang Shi

Normative modeling has emerged as a pivotal approach for characterizing heterogeneity and individual variance in neurodegenerative diseases, notably Alzheimer's disease(AD). One of the challenges of cortical normative mo…

Anatomy

Neuropsychiatric Deviations From Normative Profiles: An MRI-Derived Marker for Early Alzheimer's Disease Detection

2026-04-01 · Synne Hjertager Osenbroch, Lisa Ramona Rosvold, Yao Lu, Alvaro Fernandez-Quilez arxiv

Neuropsychiatric symptoms (NPS) such as depression and apathy are common in Alzheimer's disease (AD) and often precede cognitive decline. NPS assessments hold promise as early detection markers due to their correlation w…

Alzheimer's Disease Detection

Parsing altered brain connectivity in neurodevelopmental disorders by integrating graph-based normative modeling and deep generative networks

2024-10-14 · Rui Sherry Shen, Yusuf Osmanlıoğlu, Drew Parker, Darien Aunapu 외

Divergent brain connectivity is thought to underlie the behavioral and cognitive symptoms observed in many neurodevelopmental disorders. Quantifying divergence from neurotypical connectivity patterns offers a promising p…

Diffusion MRI

Normative Reasoning in Large Language Models: A Comparative Benchmark from Logical and Modal Perspectives

2025-10-30 · Kentaro Ozeki, Risako Ando, Takanobu Morishita, Hirohiko Abe 외 arxiv

Normative reasoning is a type of reasoning that involves normative or deontic modality, such as obligation and permission. While large language models (LLMs) have demonstrated remarkable performance across various reason…

Learn Like Humans: Use Meta-cognitive Reflection for Efficient Self-Improvement

2026-01-17 · Xinmeng Hou, Peiliang Gong, Bohao Qu, Wuqi Wang 외 arxiv

While Large Language Models (LLMs) enable complex autonomous behavior, current agents remain constrained by static, human-designed prompts that limit adaptability. Existing self-improving frameworks attempt to bridge thi…