A Multimodal Vision Transformer for Interpretable Fusion of Functional and Structural Neuroimaging Data

Author:

Bi Yuda,Abrol Anees,Fu Zening,Calhoun Vince D.ORCID

Abstract

AbstractDeep learning models, despite their potential for increasing our understanding of intricate neuroimaging data, can be hampered by challenges related to interpretability. Multimodal neuroimaging appears to be a promising approach that allows us to extract supplementary information from various imaging modalities. It’s noteworthy that functional brain changes are often more pronounced in schizophrenia, albeit potentially less reproducible, while structural MRI effects are more replicable but usually manifest smaller effects. Instead of conducting isolated analyses for each modality, the joint analysis of these data can bolster the effects and further refine our neurobiological understanding of schizophrenia. This paper introduces a novel deep learning model, the multimodal vision transformer (MultiViT), specifically engineered to enhance the accuracy of classifying schizophrenia by using structural MRI (sMRI) and functional MRI (fMRI) data independently and simultaneously leveraging the combined information from both modalities. This study uses functional network connectivity data derived from a fully automated independent component analysis method as the fMRI features and segmented gray matter volume (GMV) as the sMRI features. These offer sensitive, high-dimensional features for learning from structural and functional MRI data. The resulting MultiViT model is lightweight and robust, outperforming unimodal analyses. Our approach has been applied to data collected from control subjects and patients with schizophrenia, with the MultiViT model achieving an AUC of 0.833, which is significantly higher than the average 0.766 AUC for unimodal baselines and 0.78 AUC for multimodal baselines. Advanced algorithmic approaches for predicting and characterizing these disorders have consistently evolved, though subject and diagnostic heterogeneity pose significant challenges. Given that each modality provides only a partial representation of the brain, we can gather more comprehensive information by harnessing both modalities than by relying on either one independently. Furthermore, we conducted a saliency analysis to gain insights into the co-alterations in structural gray matter and functional network connectivity disrupted in schizophrenia. While it’s clear that the MultiViT model demonstrates differences compared to previous multimodal methods, the specifics of how it compares to methods such as MCCA and JICA are still under investigation, and more research is needed in this area. The findings underscore the potential of interpretable multimodal data fusion models like the MultiViT, highlighting their robustness and potential in the classification and understanding of schizophrenia.

Publisher

Cold Spring Harbor Laboratory

Reference45 articles.

1. “Deep learning encodes robust discriminative neuroimaging representations to outperform standard machine learning;Nature communications,2021

2. “3d-cnn based discrimination of schizophrenia using resting-state fmri;Artificial Intelligence in Medicine,2019

3. SSPNet: An interpretable 3D-CNN for classification of schizophrenia using phase maps of resting-state complex-valued fMRI data

4. “Classification of schizophrenia and normal controls using 3d convolutional neural network and outcome visualization;Schizophrenia Research,2019

5. A. Dosovitskiy , L. Beyer , A. Kolesnikov , D. Weissenborn , X. Zhai , T. Unterthiner , M. Dehghani , M. Minderer , G. Heigold , S. Gelly et al., “An image is worth 16×16 words: Transformers for image recognition at scale,” arXiv preprint arXiv:2010.11929, 2020.

同舟云学术

1.学者识别学者识别

2.学术分析学术分析

3.人才评估人才评估

"同舟云学术"是以全球学者为主线,采集、加工和组织学术论文而形成的新型学术文献查询和分析系统,可以对全球学者进行文献检索和人才价值评估。用户可以通过关注某些学科领域的顶尖人物而持续追踪该领域的学科进展和研究前沿。经过近期的数据扩容,当前同舟云学术共收录了国内外主流学术期刊6万余种,收集的期刊论文及会议论文总量共计约1.5亿篇,并以每天添加12000余篇中外论文的速度递增。我们也可以为用户提供个性化、定制化的学者数据。欢迎来电咨询!咨询电话:010-8811{复制后删除}0370

www.globalauthorid.com

TOP

Copyright © 2019-2024 北京同舟云网络信息技术有限公司
京公网安备11010802033243号  京ICP备18003416号-3