Skip to main content
Briefings in Bioinformatics logoLink to Briefings in Bioinformatics
. 2025 Aug 1;26(4):bbaf366. doi: 10.1093/bib/bbaf366

NeuroFANN: identification of neuropathological subtypes in dementia with plasma proteins by using functionally annotated neural network

Sunghong Park 1,#, Doyoon Kim 2,#, Ji-Hye Choi 3, Chang Hyung Hong 4, Sang Joon Son 5, Hyun Woong Roh 6, Hyunjung Shin 7,8,✉, Hyun Goo Woo 9,10,11,✉
PMCID: PMC12315549  PMID: 40748327

Abstract

Dementia diagnosis relies on identifying neuropathological features, such as beta-amyloid (Aβ) deposition, medial temporal lobe atrophy (MTA), and white matter hyperintensity (WMH). Recently, plasma protein biomarkers have emerged as a cost-effective and less invasive tool for identifying neuropathological features, enhanced by machine learning (ML) for precise diagnosis. However, most ML studies fail to account for protein–protein interactions (PPIs) and synergetic effects between proteins, overlooking their collective contributions to disease mechanisms. Additionally, the lack of consideration for functional properties may result in the redundant and imbalanced representation of proteins and their functions, potentially limiting the effectiveness of dementia diagnosis. In this study, we propose NeuroFANN, a method designed to classify three neuropathological subtypes in dementia—positivity for Aβ, MTA, and WMH—using plasma protein biomarkers. A key feature of NeuroFANN is the combination of the PPI network-based synergetic effects with the functional annotation-based protein biomarker clustering. NeuroFANN extracts synergetic effects by propagating independent effects of proteins across the PPI network, which are then aggregated in functional protein clusters, thereby enabling global PPI awareness and capturing the biological properties of protein biomarkers. From a South Korean cohort, 54 proteins were identified as plasma protein biomarkers for dementia subtypes and grouped into 16 clusters. NeuroFANN outperformed comparison methods in classifying dementia subtypes, with its core components validated as key contributors to superior performance. Additionally, the risk scores predicted by NeuroFANN showed a strong association with longitudinal cognitive decline, demonstrating its potential as a valuable diagnostic tool in clinical settings.

Keywords: dementia, neuropathological subtype, plasma proteomic profile, protein–protein interaction, biologically functional annotation, graph-based machine learning

Introduction

Dementia is a disease of degenerative brain changes characterized by cognitive decline and impairment in daily living functions [1]. Several subtypes of dementia have been identified, including the most common type of Alzheimer's disease (AD) and vascular dementia [2], each associated with distinct neuropathological mechanisms [2]. Key pathological features include amyloid-beta (Aβ) accumulation, tau protein deposition, medial temporal lobe atrophy (MTA), and white matter hyperintensity (WMH) [3–5], all of which are closely associated with clinical symptoms and patient prognosis [6, 7].

Specifically, Aβ pathology, a hallmark of AD, involves the aggregation of Aβ proteins and subsequent plaque formation in the brain [8]. MTA pathology represents structural degeneration resulting primarily from neuronal injury, synaptic loss, and neuroinflammation, particularly within the hippocampus and medial temporal lobes. This makes MTA a crucial biomarker for diagnosing AD and distinguishing it from other dementia subtypes [9]. WMH pathology arises due to small vessel disease and is characterized by endothelial dysfunction, chronic vascular inflammation, and blood–brain barrier disruption [10]. These neuropathological features align well with established frameworks such as ATN (amyloid, tau, neurodegeneration) and vascular pathology [11, 12], playing critical roles in the classification of dementia subtypes and informing clinical management strategies.

Conventional diagnostic approaches for identifying neuropathological features include neuroimaging and cerebrospinal fluid (CSF) biomarkers. Positron emission tomography (PET), for instance, is used to detect Aβ deposition, while magnetic resonance imaging (MRI) assesses the severity of MTA [13]. Additionally, CSF biomarkers, such as Aβ42 and phosphorylated tau proteins, are widely utilized [14, 15]. However, PET and MRI are costly and have limited accessibility, and CSF testing is invasive with the potential for serious complications. Moreover, both methods have limitations in fully capturing the complex molecular pathology of dementia.

Recent advancements in neurology have highlighted plasma protein biomarkers as promising alternatives for diagnosing various neurodegenerative diseases and understanding their underlying mechanisms. Plasma biomarkers offer advantages over CSF biomarkers and neuroimaging by being less invasive and more cost-effective [16, 17]. Integrating these biomarkers into existing medical systems could reduce patient burden and improve diagnostic accessibility. Consequently, plasma biomarkers have gained increasing adoption across various disorders beyond neurodegenerative diseases [18, 19].

Plasma protein biomarkers effectively reflect distinct dementia pathologies at the molecular level, including Aβ, MTA, and WMH. First, proteins involved in amyloid processing and clearance, such as apolipoprotein E, complement, and inflammation-related proteins, significantly correlate with brain amyloid deposition [20–22]. Second, plasma neurofilament light chain, indicative of neuronal injury, and glial fibrillary acidic protein, a marker of neuroinflammation, reliably reflect MTA pathology [23–25]. Third, plasma proteins indicative of endothelial dysfunction, vascular inflammation, and blood–brain barrier integrity have been extensively linked to white matter lesion burden, thereby biologically reflecting cerebrovascular injury associated with WMH [26–28]. These molecular associations enhance our understanding of dementia pathogenesis and improve diagnostic accuracy, particularly in early disease stages.

Furthermore, the integration of machine learning (ML) has significantly advanced plasma biomarker-based diagnostic approaches, providing deeper insights into dementia and other diseases [29, 30]. However, existing ML studies on dementia biomarkers often emphasize only the independent effects of individual proteins, overlooking their synergetic interactions within biological networks [31–33]. This narrow approach contrasts with broader trends observed in other omics-based research fields—such as cancer subtype prediction, multi-omics drug-response modeling, and rare-disease gene discovery—where network-based aggregation of PPIs has become a standard analytical practice [34–36]. Similarly, plasma biomarker-based dementia studies have rarely utilized functional annotations guided feature grouping methods [37, 38], despite their widespread use in transcriptomic and proteomic analyses to elucidate underlying biological processes and ensure balanced functional representation [39, 40]. Neglecting these strategies may lead to redundancy among selected biomarkers and imbalanced representations of biological functions [41], ultimately compromising the effectiveness and interpretability of diagnostic models.

To address these challenges, the present study aims to develop a novel ML model based on a graph neural network (GNN), namely NeuroFANN, for predicting three neuropathological subtypes in dementia—positivity for Aβ, MTA, and WMH—by employing plasma biomarkers. A distinguishing feature of NeuroFANN is its utilization of the PPI network and functional annotations for plasma biomarkers. This strategy enables the extraction of the synergetic effects between biomarkers, capturing the collective contribution of multiple proteins to distinct dementia subtypes. Furthermore, it facilitates a functionally balanced analysis informed by understanding the underlying biological processes.

Methods

Overview of the study

As shown in Fig. 1a, this study begins with the identification of plasma proteins associated with neuropathological subtypes of dementia through differential expression analysis (DEA), followed by the clustering of these protein biomarkers based on their functional annotations. Next, NeuroFANN is used to predict individual risks for dementia subtypes. As illustrated in Fig. 1b, the model projects the independent effects of proteins that passed the DEA criteria onto the PPI network, capturing synergetic effects that reflect the global properties of PPIs. The synergetic effects are then aggregated within predefined clusters, reflecting the functional annotations of plasma protein biomarkers. The key contributions of this study are summarized below.

Figure 1.

Overview diagram of the NeuroFANN pipeline, illustrating the integration of functionally clustered plasma proteins and PPI-based synergetic effects.

Overview of the study. (a) Represents the overview of the study design. The proposed method begins by identifying plasma biomarkers associated with neuropathological subtypes in dementia using differential expression analysis. These biomarkers are then grouped based on their functional annotations. Individual risks for dementia subtypes are subsequently predicted using NeuroFANN. (b) Schematically describes the proposed method. NeuroFANN propagates the independent effects of biomarkers onto the PPI network, extracting synergetic effects that account for global interactions between proteins. These synergetic effects are aggregated within predefined clusters, aligning with the functional annotations of the plasma biomarkers.

  • We propose a novel GNN-based method, NeuroFANN, for identifying neuropathological subtypes in dementia with plasma proteins.

  • NeuroFANN extracts the synergetic effects between proteins through the PPI network, capturing the collective impact of multiple proteins on the dementia subtypes.

  • NeuroFANN achieves improved outcomes by utilizing biologically informed clustering based on the functional annotations of plasma protein biomarkers.

Differential expression analysis

The proposed method initially identifies differentially expressed proteins (DEPs) for neuropathological subtypes of dementia. The expression levels of each protein are compared between the positive and negative groups for each subtype, and the statistical significance of those differences is subsequently analyzed. To carry out this analysis, the “limma” R/Bioconductor package [42] is utilized. Consequently, a union of proteins that are differentially expressed for the three subtypes is employed as plasma protein biomarkers among a total of assayed proteins.

Biologically informed protein clustering

Biologically informed protein clustering comprises two steps: functional grouping and biomarker clustering. This process begins with identifying Gene Ontology (GO) groups consisting of GO terms significantly associated with the plasma biomarkers. Subsequently, a hierarchical clustering analysis is performed based on the incidence matrix for the GO groups and their corresponding plasma biomarkers.

Functional grouping

The proposed method firstly identifies the functional annotations for the plasma biomarkers by using the “‘Functional Annotation Clustering” tool, which is provided by the Database for Annotation, Visualization, and Integrated Discovery (DAVID version 2021 with DAVID Knowledgebase version 2024q2, https://david.ncifcrf.gov). The running options are as follows: similarity term overlap = 3, similarity threshold = 1, initial group membership = 2, final group membership = 2, multiple linkage threshold = 1, and enrichment thresholds = 0.05. The resulting outcomes are represented as the binary incidence matrix for the GO groups and their corresponding plasma protein biomarkers.

Biomarker clustering

The proposed method subsequently conducts a hierarchical clustering analysis on the protein biomarkers. The distances between proteins are calculated by the incidence matrix from the functional grouping using the Jaccard metric. The distances between clusters are measured using Ward’s method, and the cutoff is set as a minimum value that ensures all clusters contain a minimum of two biomarkers.

Functionally annotated neural network

Model implementation

NeuroFANN consists of three main components: network propagation, cluster aggregation, and subtype prediction. First, the independent effects of plasma biomarkers are propagated through a PPI network to capture the synergetic effects among biomarkers. Next, the synergetic effects of biomarkers within the same cluster are aggregated, with each biomarker’s contribution modulated by a trainable importance parameter. Finally, these aggregated cluster-level effects are used to predict risks of neuropathological subtypes.

Network propagation

Let Inline graphic and Inline graphic denote the numbers of plasma biomarkers and study participants, respectively. We define Inline graphic as the matrix of independent effect, and Inline graphic as the matrix representing the PPI network. We aim to derive a propagated signal matrix Inline graphic capturing the synergetic effects by propagating Inline graphic through Inline graphic. Formally, we address the following optimization problem:

graphic file with name DmEquation1.gif (1)

where Inline graphic denotes the normalized graph Laplacian of Inline graphic, defined as Inline graphic, with Inline graphic being a diagonal matrix whose elements are Inline graphic. Inline graphic is a diagonal matrix with diagonal entries Inline graphic (Inline graphic), balancing the smoothness term (Inline graphic) against the steadiness term (Inline graphic). Biologically, the smoothness term ensures minimal differences in propagated signals between adjacent proteins within the PPI network, emphasizing coordinated actions and functional coherence among interacting proteins. Conversely, the steadiness term preserves original individual biomarker signals, preventing excessive dilution of protein-specific effects during network propagation. The closed-form solution to this optimization problem is given by:

graphic file with name DmEquation2.gif
Cluster aggregation

The derived synergetic effects in Inline graphic are then aggregated at the cluster level to produce Inline graphic, where Inline graphic is the number of biomarker clusters. For the Inline graphic-th cluster Inline graphic, its aggregated effect Inline graphic is:

graphic file with name DmEquation2a.gif

Here, Inline graphic denotes the subset of Inline graphic corresponding to biomarkers within cluster Inline graphic, and Inline graphic indicates the trainable importance parameter vector for Inline graphic. The parameters Inline graphic are initialized with uniform weights inversely proportional to their cluster size, specifically Inline graphic. After training, the optimized Inline graphic values represent the relative importance of proteins within Inline graphic. The exponentiated terms Inline graphic normalize to ensure the contribution weights within each cluster sum to 1. Consequently, higher Inline graphic values indicate that the corresponding proteins contribute more significantly to the functional role of Inline graphic in relation to neuropathological subtypes, compared to other proteins in the same cluster.

Subtype prediction

Finally, NeuroFANN predicts the risk for each neuropathological subtype based on the aggregated cluster effects. Denote Inline graphic by the prediction parameters for a subtype Inline graphic. The predicted risk vector Inline graphic is obtained via a logistic function:

graphic file with name DmEquation9.gif

Parameter optimization

NeuroFANN jointly optimizes three parameter sets: the propagation parameter (Inline graphic) for network propagation, the importance parameter (Inline graphic) for cluster aggregation, and the prediction parameter (Inline graphic) for subtype prediction. These parameters are trained to minimize the binary cross-entropy:

graphic file with name DmEquation10.gif

where Inline graphic indicates subtype presence. The overall objective, including regularization, is:

graphic file with name DmEquation11.gif (2)

where Inline graphic is the Inline graphic regularization term defined as Inline graphic, and Inline graphic is the regularization coefficient. The objective function in equation (2) is optimized via gradient descent method [43], updating each parameter set in turn.

Minimization over Inline graphic: For each subtype Inline graphic, we first differentiate the terms in equation (2) as:

graphic file with name DmEquation12.gif

Combining these terms, the gradient with respect to Inline graphic becomes:

graphic file with name DmEquation14.gif

Minimization over Inline graphic: To find the gradient with respect to the importance parameter Inline graphic, we first derive Inline graphic as below:

graphic file with name DmEquation15.gif

By partitioning Inline graphic as Inline graphic, we compute Inline graphic by combining Inline graphic with Inline graphic. Consequently, the gradient with respect to Inline graphic (the parameter vector for cluster Inline graphic) obtained as follows.

graphic file with name DmEquation16.gif

Minimization over Inline graphic: Lastly, for the propagation parameter Inline graphic, we first derive:

graphic file with name DmEquation17.gif

Combining Inline graphic with Inline graphic gives:

graphic file with name DmEquation20.gif

Since Inline graphic is diagonal, the gradient with respect to Inline graphic (its diagonal entries) is:

graphic file with name DmEquation21.gif

Statistical analysis

We used MATLAB (version R2024a), Python (version 3.9.7), and R (version 3.6.1) for all analyses in this study. Continuous variables are presented as mean with standard deviation (SD) or median with interquartile range, with group-wise comparison using a one-way analysis of variance for Gaussian distribution or the Kruskal–Wallis test for non-Gaussian distribution. Categorical variables are presented as count (percentage), with group-wise comparison using the chi-squared test. Statistical significance was indicated by a two-sided P-value.

Materials

Study participants

We recruited participants from the Biobank Innovations for chronic Cerebrovascular disease With ALZheimer’s disease Study at Ajou University Hospital (Suwon, Republic of Korea) [44]. The recruited study participants underwent three-dimensional T1-weighted MRI using a 3.0 Tesla scanner and 18F-flutemetamol PET scan using a Discovery STE/690 PET/CT scanner. An average of 185 MBq of 18F-flutemetamol was intravenously injected as a bolus via the antecubital vein. After a 90-minute uptake period, PET scans were acquired as four consecutive 5-minute frames, totaling 20 minutes. All neuroimaging data were independently reviewed by neuroradiologists.

Subsequently, clinicians determined the neuropathological subtypes of the participants based on MRI and PET scans. Participants were classified as amyloid-positive if their global cortical standardized uptake value ratio exceeded 0.634, a cutoff previously established in elderly Koreans [45]. MTA positivity was assessed using the Scheltens’ scale [46], which grades atrophy severity from 0 to 4 in each hemisphere; participants were classified as MTA-positive if the combined score from both hemispheres was greater than 4. WMH was classified into three severity categories (mild, moderate, and severe) using the Fazekas scale [47]; participants identified with moderate or severe WMH were considered WMH-positive.

In total, among the 475 participants included in the study cohort, 142 (29.9%), 106 (22.3%), and 172 (36.2%) individuals were identified as positive for Aβ, MTA, and WMH, respectively, while 181 (38.1%) participants were negative for all three subtypes. Importantly, these three neuropathological subtypes were not mutually exclusive, reflecting common clinical observations of multiple concurrent pathologies within individuals. Specifically, 104 (21.9%) participants exhibited positivity for two or more subtypes: 82 (17.3%) participants were positive for exactly two subtypes—23 (4.8%) for both Aβ and MTA, 24 (5.1%) for both Aβ and WMH, and 35 (7.4%) for both MTA and WMH. Additionally, 22 (4.6%) participants were classified as positive for all three subtypes.

Finally, we divided the participants into discovery and validation cohorts. The validation cohort comprised participants who underwent cognitive assessments at the two-year follow-up, whereas participants with only baseline assessments were allocated to the discovery cohort. Detailed demographic and clinical characteristics of the study participants are summarized in Table 1.

Table 1.

Demographic and clinical characteristics of participants

Characteristics Total participants (N = 475) Discovery cohort (N = 348) Validation cohort (N = 127) P-valuea
Age, median (IQR), yr 73 (67–77) 72 (67–77) 73 (67–78) .6343
Female, No. (%) 332 (69.9) 241 (69.3) 91 (71.7) .6146
MMSE, median (IQR) 24 (20–27) 24 (20–27) 24 (22–27) .0525
CDR-SB, median (IQR) 2.5 (1.5–4.0) 2.5 (1.5–4.5) 2.0 (1.5–3.0) .0921
GDS, median (IQR) 3 (3–4) 3 (3–4) 3 (3–4) .2841
APOE ε2 carrier, No. (%) 68 (14.3) 48 (13.8) 20 (15.7) .5912
APOE ε4 carrier, No. (%) 129 (27.2) 91 (26.1) 38 (29.9) .4144
Aβ positivity, No. (%) 142 (29.9) 100 (28.7) 42 (33.1) .3621
MTA positivityb, No. (%) 106 (22.3) 81 (23.3) 25 (19.7) .4065
WMH positivityc, No. (%) 172 (36.2) 121 (34.8) 51 (40.2) .2806
Only Aβ positivity, No. (%) 73 (15.4) 53 (15.2) 20 (15.7) .8901
Only MTA positivity, No. (%) 26 (5.5) 20 (5.7) 6 (4.7) .6653
Only WMH positivity, No. (%) 91 (19.2) 67 (19.3) 24 (18.9) .9308
Aβ and MTA positivity, No. (%) 23 (4.8) 18 (5.2) 5 (3.9) .5797
Aβ and WMH positivity, No. (%) 24 (5.1) 11 (3.2) 13 (10.2) .0018
MTA and WMH positivity, No. (%) 35 (7.4) 25 (7.2) 10 (7.9) .7994
Aβ, MTA, and WMH positivity, No. (%) 22 (4.6) 18 (5.2) 4 (3.1) .3543
Aβ, MTA, and WMH negativity, No. (%) 181 (38.1) 136 (39.1) 45 (35.4) .4698

IQR, interquartile range; MMSE, Mini-Mental Status Examination; CDR-SB, Clinical Dementia Rating Sum of Boxes; GDS, Global Deterioration Scale; APOE, apolipoprotein E; Aβ, amyloid beta; MTA, medial temporal lobe atrophy score; WMH, white matter hyperintensity.

a P-values were calculated for each characteristic by means of a comparison between the discovery and validation cohorts.

bMTA scale was divided into left and right and subdivided into 0–4 according to severity, and in this study, MTA-positive was set for cases where the sum of left and right sides was five or more.

cWMH scale was divided into three types (mild, moderate, and severe), and WMH-positive was set for moderate and severe.

Proteomic assays

Blood samples from study participants were collected concurrently with neuroimaging and cognitive function assessments. The samples were centrifuged at room temperature at 3000 rpm for 10 minutes to separate plasma and serum supernatants. To ensure high-purity samples, the plasma and serum supernatants underwent an additional centrifugation step under identical conditions, after which the purified supernatants were immediately collected and stored in an ultra-low temperature freezer at −80°C. The plasma samples were profiled by the Olink Target 96 Neurology and Olink Target 48 Cytokine panels. The former comprises 92 established assays associated with neurobiological diseases, while the latter includes 45 selected assays highly relevant to inflammatory processes, more details can be found at: https://olink.com, and the assayed proteins are listed in Supplementary Table S1. The raw data for protein expression underwent quality control (QC) through both internal and external controls in the panels. We excluded proteins with missing frequency over 10% and imputed missing values by using the k-nearest neighbor method. Subsequently, the protein expression values were standardized by using Z-score normalization and then transformed by the logistic function for scaling.

Results

Identified plasma biomarkers and protein clusters

We conducted DEA for 127 QC-passed proteins out of 137 assayed proteins, identifying 10, 23, and 46 proteins as significant plasma biomarkers for Aβ, MTA, and WMH, respectively (Fig. 2a; Supplementary Table S2). In total, 54 proteins were identified as plasma biomarkers for three neuropathological subtypes, comprising 43 (79.6%) upregulated and 11 (20.4%) downregulated proteins (Fig. 2b). Of those, BCAN and CNTN5 were significantly downregulated proteins in all three subtypes, where those significances associated with dementia have been shown in clinical studies [48, 49], with an average log2Fold-Change (FC) of −0.182, which represents a 16.9% reduction compared to an average log2FC of −0.156 for the remaining downregulated proteins. NTRK3 was the most significantly upregulated protein for Aβ, demonstrating that NTRK3 signaling is associated with AD by promoting synaptic plasticity and interacting with Aβ-related pathways [50], while GFRA1 indicated the highest degree of upregulation in both MTA and WMH, where GFRA1 has known to bind with the glial cell line-derived neurotrophic factor (GDNF) [51], and the GDNF-GFRA1 complex has revealed to affect hippocampal-related neurological disorders and cerebral white matter damages [52, 53].

Figure 2.

Summary of 54 plasma protein biomarkers identified for dementia subtypes through differential expression analysis and their grouping into 16 functional clusters.

Identified plasma biomarkers and protein clusters for the neuropathological subtypes. (a) Of the 127 proteins that passed quality control process, differential expression analysis identified 10, 23, and 46 proteins as significant plasma biomarkers for aβ, MTA, and WMH, respectively. (b) In total, 54 proteins were identified as plasma biomarkers for three neuropathological subtypes of dementia, comprising 43 upregulated and 11 downregulated proteins. (c) Subsequent clustering analysis grouped the identified biomarkers into 16 clusters.

Next, we identified 86 GO groups for 239 GO terms that are significantly associated with the plasma biomarkers based on the functional annotation analysis, with each group containing an average of 7.6 (SD, 5.9) biomarkers. By performing clustering analysis, the biomarkers could be grouped into 16 clusters (Fig. 2c; Supplementary Table S3). The “Cytokines” cluster had the highest number of biomarkers (n = 9), followed by the ‘Positive regulation of metabolic process’ cluster (n = 5), demonstrating the importance of the inflammatory response of cytokines that are related to dementia subtypes based on the association between neurodegeneration and neuroinflammation [54], including the significance of metabolic functions that are associated with dementia [55].

Predicted outcomes for neuropathological subtypes

The predicted risks by NeuroFANN demonstrated significant differences between the diagnostic groups for each subtype, with P-values of 1.96 × 10−12, 7.63 × 10−7, and 2.95 × 10−13 for Aβ, MTA, and WMH, respectively (Fig. 3a). For the true positive group for each subtype, the proportion of participants in the interval of the highest value in the predicted risk distribution (0.9 or higher) was 35.39% (95% CI, 34.44%–36.33%) on average, with the highest proportion in MTA at 39.76% (95% CI, 38.44%–41.08%), followed by Aβ at 36.81% (95% CI, 35.98%–37.64%) and WMH at 29.59% (95% CI, 28.89%–30.28%). The log10odds ratios (log10ORs) between the highest and lowest predicted risk groups were 2.95 (95% CI, 2.93%–2.97%), 3.13 (95% CI, 3.11%–3.16%), and 3.10 (95% CI, 3.08%–3.11%) for ABT, MTA, and WMH, respectively, representing an increasing pattern of log10ORs in all risk groups across the three subtypes (Fig. 3b). Our model achieved an average AUROC of 0.832 (95% CI, 0.830%–0.833%), where the highest AUROC was observed for WMH at 0.847 (95% CI, 0.843%–0.850%), followed by Aβ at 0.846 (95% CI, 0.843%–0.849%) and MTA at 0.803 (95% CI, 0.800%–0.805%; Fig. 3c).

Figure 3.

Comparative plots of predicted risk distributions, log-odds ratios, and AUROC values across diagnostic groups for each dementia subtype.

Comparative analyses of predicted outcomes for neuropathological subtypes by NeuroFANN. For each neuropathological dementia subtype, (a) compares the risk-level differences between the diagnostic groups, (b) presents odds ratios according to the risk groups in comparison to the lowest risk group, and (c) evaluates the AUROC performances.

As shown in Fig. 4, we further investigated the longitudinal changes in cognitive function of at-risk patients for each subtype using the MMSE, CDR-SB, and GDS scores assessed over a two-year follow-up. Of the 127 participants in the validation cohort, 58 (45.7%), 42 (33.1%), and 63 (49.6%) were identified as at-risk patients for Aβ, MTA, and WMH, respectively, whose predicted risks for each subtype exceeded the positivity threshold in more than half of the whole predictions. The longitudinal changes for all three test scores in the at-risk patients for Aβ and WMH were statistically significant, and the MTA-at-risk patients indicated significant baseline-to-follow-up differences for CDR-SB and GDS, except MMSE, while the longitudinal changes for all scores in the non-at-risk patients for all subtypes were not statistically significant (Fig. 4a–c). Our investigation also revealed a statistically significant association between the cognitive decline and the predicted risks for all three subtypes. The incidence of cognitive decline was determined by the simultaneous deterioration of MMSE, CDR-SB, and GDS scores from the baseline assessment to the follow-up evaluation (Fig. 4d). The hazard ratios were as follows: Aβ at 2.802 (95% CI, 1.278%–6.143%), MTA at 1.872 (95% CI, 0.880%–3.986%), and WMH at 1.958 (95% CI, 0.937%–4.087%), respectively.

Figure 4.

Two-year longitudinal trajectories of MMSE, CDR-SB, and GDS in at-risk patients, and survival analysis of cognitive decline based on predicted subtype risk.

Associations between predicted outcomes and longitudinal changes in cognitive assessments. For patients at risk of each subtype, two-year trajectories were examined for the MMSE (a), CDR-SB (b), and GDS (c). The relationship between cognitive decline and predicted risks was evaluated using survival analysis (d), where cognitive decline was defined as a concurrent deterioration in MMSE, CDR-SB, and GDS scores from baseline to follow-up.

Model evaluation with performance comparison

Experimental settings

Our model was compared with seven GNN-based models: graph convolutional network (GCN) [56], simple graph convolution (SGC) [57], exponential graph convolution (EGC) [58], linear graph convolution (LGC) [58], MixHop [59], universal graph convolutional network (UGCN) [60], and mixed-order graph convolutional network (MOGCN) [61]. The discovery cohort was employed for model training, and the validation cohort was fixed as the test set. The model performance was evaluated by measuring six metrics: area under the receiver operating characteristic curve (AUROC), sensitivity, specificity, positive predictive value (PPV), and negative predictive value (NPV). Each model derived 100 repeated predictions, with 20 iterations of five-fold cross-validation.

Performance comparison

The performance of NeuroFANN was compared with seven other GNN-based methods (Supplementary Table S4). The average AUROC of the comparison methods was 0.776 (95% CI, 0.775%–0.777%), indicating that our model demonstrated 7.28% (95% CI, 7.06%–7.49%) higher AUROC on average (Fig. 5a). For each subtype, the comparison methods showed average AUROCs as follows: 0.781 (95% CI, 0.779%–0.784%) for Aβ, 0.763 (95% CI, 0.762%–0.764%) for MTA, and 0.784 (95% CI, 0.782%–0.785%) for WMH; thereby demonstrating that our model exhibited improved AUROC for all subtypes: 8.51% (95% CI, 8.11%–8.91%) for Aβ, 5.22% (95% CI, 5.04%–5.40%) for MTA, and 8.17% (95% CI, 7.86%–8.48%) for WMH (Fig. 5b). We further compared the performance of models according to the number of hops each model reflects in the PPI network: 2 hops for GCN and SGC, 3 hops for EGC, 4 hops for LGC and MixHop, 5 hops for UGCN, 6 hops for MOGCN, and all hops for NeuroFANN (defined as 7 hops for comparability). This comparison demonstrated significant performance differences according to the hop count, with a strong correlation between performance and the number of hops (Pearson correlation coefficient = 0.931, P–value <0.001; Jonckheere trend test, P–value = 2.20 × 10−16; Fig. 5c). These findings highlight the importance of utilizing the PPI in a global range; specifically, NeuroFANN’s network propagation effectively captured interactions among distant proteins, while the comparison models only incorporated local interactions between neighboring proteins.

Figure 5.

Performance comparison of NeuroFANN versus seven GNN models, showing AUROC differences by subtype and impact of PPI network hop counts.

Model evaluation with performance comparison and ablation study. The AUROCs of the ML models were evaluated in terms of their overall performance for the three neuropathological subtypes (a), as well as their individual performance for each subtype (b). The performance of the models was further compared by the number of hops each model reflects in the PPI network, with NeuroFANN defined as having 7 hops for comparability (c).

Empirical analysis of the proposed method

Ablation study on key components of NeuroFANN

First, we assessed the contributions of NeuroFANN’s key components—network propagation and cluster aggregation—to predictive performance by developing four ablated models (AMs 1–4): AM 1 retained network propagation but employed equal aggregation, where protein clusters were aggregated using uniform weights instead of trained cluster-specific weights; AM 2 excluded the cluster aggregation; AM 3 excluded the network propagation; and AM 4 excluded both network propagation and cluster aggregation. Among these, AM 1 showed the highest AUROC of 0.817 (±0.010), followed by AM 2 at 0.806 (±0.007), and AM 3 at 0.769 (±0.010). AM 4 exhibited the lowest AUROC performance of 0.688 (±0.002). Compared to the average AUROC of these four ablated models, the complete NeuroFANN model achieved an average improvement of 7.65% (Fig. 6a; Supplementary Table S5). These results clearly indicate that the network propagation and cluster aggregation components substantially enhance NeuroFANN’s predictive performance, underscoring their critical roles in effectively integrating PPI information.

Figure 6.

Ablation study of NeuroFANN’s key components, comparison of protein selection strategies, and sensitivity analysis of clustering parameters.

Empirical analysis of the proposed method. First, the contribution of NeuroFANN’s key components—Network propagation and cluster aggregation—To predictive performance was evaluated through four ablated models (a). Next, the impact of different initial protein selection methods on NeuroFANN’s predictive performance was investigated (b). Last, the sensitivity of protein clustering parameters was analyzed (c).

Comparative evaluation of initial protein selection methods

Next, we evaluated the impact of different initial protein selection strategies on the predictive performance of NeuroFANN. In addition to the univariate DEA-based P–value filtering used in this study, multivariate approaches including backward and forward selection methods, as well as an inclusive ‘entire selection’ involving all 127 proteins, were examined. As shown in Fig. 6b, the backward selection identified 83 proteins grouped into 31 clusters, whereas forward selection identified 19 proteins grouped into 15 clusters. In the case of entire selection, all 127 proteins were grouped into 42 clusters. Subsequently, NeuroFANN achieved AUROC values of 0.819, 0.749, and 0.807 for backward, forward, and entire selections, respectively, representing decreases of 1.6%, 11.1%, and 3.1% compared to the baseline DEA-based selection performance of 0.832. These results suggest that a parsimonious feature set derived from DEA-based selection optimally balances informative signal and noise, highlighting the importance of careful biomarker pre-selection in enhancing NeuroFANN’s predictive performance.

Sensitivity analysis of protein clustering parameters

Finally, we conducted a sensitivity analysis of protein clustering parameters: (i) inter-cluster distance measures, (ii) the minimum number of shared proteins between GO terms, and (iii) the minimum number of GO terms within protein clusters (Fig. 6c). First, in addition to the baseline Jaccard distance measure, we evaluated Cosine, Euclidean, and Manhattan distances. Compared to the 16 clusters derived by the Jaccard method (AUROC = 0.832), Cosine, Euclidean, and Manhattan distances yielded fewer clusters—13, 8, and 7, respectively—and correspondingly reduced AUROC values (0.810, 0.791, and 0.768, respectively). Second, when increasing the minimum number of shared proteins between GO terms from the original setting of 3 to 4, 5, and 6, the number of clusters changed marginally (16, 17, 16, and 13 clusters, respectively), accompanied by slight decreases in AUROC performance (0.832, 0.830, 0.828, and 0.824, respectively). Third, raising the minimum number of GO terms per cluster from the original setting of 2 to 3, 4, and 5 markedly increased the number of clusters (16, 20, 35, and 36, respectively) while progressively decreasing AUROC performance (0.832, 0.823, 0.815, and 0.812, respectively). Collectively, these findings suggest that NeuroFANN achieves optimal predictive performance under the original, conservatively tuned clustering parameters, highlighting the necessity of biologically balanced parameter settings to preserve informative signals while minimizing performance loss due to either excessive or insufficient aggregation.

Interpretation of features and parameters

We evaluated predictive importance of biomarker clusters using SHapley Additive exPlanations (SHAP) [62] (Fig. 7a). The “neuron development associated signal transduction” cluster exhibited the highest importance for Aβ positivity, with a SHAP value of 0.776 (95% CI, 0.732%–0.819%). For MTA and WMH positivity, the most significant clusters were “extracellular matrix” (0.838; 95% CI, 0.796%–0.879%) and “positive regulation of metabolic process” (0.821; 95% CI, 0.786%–0.856%), respectively. These three clusters ranked among the top three overall feature importance, followed closely by clusters such as “signal transduction,” “molecular transducer activity,” “cytokine associated signal transduction,” “neuron development associated cell motility,” and “programmed cell death.” Together, these eight clusters exhibited higher feature importance than the mean value of 1.321 observed across all clusters, identifying them as key clusters for predicting neuropathological subtypes. Subsequently, we analyzed the protein-level importance within the 16 clusters derived from the cluster aggregation step of NeuroFANN (Fig. 7b). In the “positive regulation of metabolic process” cluster, GFRA1 demonstrated the highest protein-level importance (0.366; 95% CI, 0.324%–0.407%), followed by NTRK3 (0.220; 95% CI, 0.192%–0.248%). Additionally, proteins such as BCAN, CNTN, DKK4, NCAN, CD300LF, UNC5C, and ROBO2 showed the greatest importance within their respective clusters. These findings suggest that SHAP-based importance highlight the functional significance of each protein cluster for neuropathological subtypes by capturing their marginal contributions to the model output, while the model-driven importance, determined by the relative weights assigned to individual proteins within clusters, identify proteins that serve as functional representatives within biologically coherent modules. This complementary interpretation underscores the value of our integrated modeling approach, effectively capturing both individual and collective effects of biomarkers.

Figure 7.

SHAP-based feature importance of biomarker clusters and model-derived protein-level importance scores across clusters.

Predictive importance for biomarker clusters and proteins. The lists of SHAP-based feature importance for biomarker clusters (a) and model-driven protein importance for plasma biomarkers (b) are compared.

Lastly, we analyzed the prediction parameters of biomarker clusters (Fig. 8). The clusters with consistently positive prediction parameters across all dementia subtypes were “positive regulation of metabolic process,” “macromolecule-associated signal transduction,” and “signal transduction.” Among these clusters, the “positive regulation of metabolic process” cluster exhibited the highest values across all subtypes, indicating the strongest association with increased risk. GFRA1, identified as the key protein in this cluster, had its significance validated through DEA. Moreover, MMP12 and DKK4, the most critical proteins in the other two clusters, have been confirmed as plasma biomarkers significantly associated with dementia in various studies [16, 63, 64]. These findings align closely with existing clinical research, underscoring the robustness of our results.

Figure 8.

Estimated prediction parameters for biomarker clusters across dementia subtypes, highlighting clusters with consistent positive contributions to risk.

Estimated prediction parameters of biomarker clusters. The prediction parameters of biomarker clusters for each dementia subtype estimated by NeuroFANN were compared.

Discussion

This study aimed to predict the neuropathological subtypes of dementia with plasma biomarkers by developing a novel ML model. We designed NeuroFANN to be further biologically informed by reflecting the interactive and functional properties of proteins. Our model utilized the PPI-based synergetic effects of proteins, encompassing the collective contribution of multiple proteins to the neuropathological subtypes, in addition to the functional annotation-based clustering of biomarkers, considering the balanced and diversified biological functions of proteins. To the best of our knowledge, this is the first attempt to simultaneously implement those two components in a single ML model, while they have been applied to ML separately in previous studies [31, 65, 66].

NeuroFANN outperformed seven GNN-based models in prediction for all three subtypes, even when it only applied the network propagation, with a significantly enhanced performance compared to that of training the independent effects of biomarkers, suggesting the validity of utilizing the PPI-based synergetic effects. The performance comparison results also implicate the necessity of utilizing the PPI in a global range. The network propagation of our model enabled reflecting interactions between distant proteins in the PPI network, while the comparison models only reflected the interactions between neighboring proteins within a local range: 2-hop for GCN and SGC, 3-hop for EGC, 4-hop for MixHop and UGCN, 5-hop for LGC, and 6-hop for MOGCN, indicating the performance of comparison models tended to be proportional to the number of hops.

This study enabled understanding the contribution of functionally clustered biomarkers to the neuropathological subtypes, covering metabolic function [55, 67, 68], neuron development [69], extracellular region [70, 71], immune response [72], and cellular communication [73]. We identified the “positive regulation of metabolic process” as the most important biomarker cluster for predicting neuropathological subtypes. The most informative proteins in this cluster were GFRA1 and NTRK3, consistent with the DEP analysis results indicating that these were the remarkably upregulated proteins for the neuropathological subtypes, including clinical implications revealed that GFRA1 and NTRK3 are significantly associated with neurodegenerative changes [74–77], binding with the nerve growth factor or the glial cell line-derived neurotrophic factor that are related to metabolic dysfunction [78–80]. Our findings also covered well-known dementia-associated proteins, such as BCAN, CNTN5, and NCAN [81, 82], along with identifying recently promising biomarkers, including IL18, MMP12, TNFSF12, and UNC5C [63, 83–85], where these biomarkers presented the highest protein importance in their clusters.

Limitations

This study has limitations as follows. First, the plasma biomarker identification was limited to assayed proteins included in the neurology and cytokine panels. Although both panels yielded valuable biomarkers for neuropathological subtypes, this study should be extended to utilize the additional panels to gain a more comprehensive understanding of dementia’s biological nature. Second, this study did not consider tauopathy to be a neuropathological subtype due to the absence of tau PET scans for the participants. Since tauopathy is an important factor in dementia, our model should be complemented to include it. Third, the findings in this study were limited to the single cohort for the Korean and to the internal validation. Therefore, it is challenging to ensure that our model will perform consistently well across different cohorts with various ethnicities, requiring a further demonstration of the model using the external validation datasets.

Conclusion

In this diagnostic study on neuropathological subtypes of dementia, we developed a predictive ML model that reflects the in-depth biological properties of plasma biomarkers. By identifying PPI-based synergetic effects, clustering biomarkers based on functional annotations, and utilizing them in the ML model, we gained a deeper understanding of the molecular nature of plasma biomarkers, leading to more accurate prediction. Our findings suggest that the developed model can contribute to the clinical field as a diagnostic aid for more precise screening of patients with dementia.

Key Points

  • We propose NeuroFANN, a novel machine learning model designed to identify neuropathological subtypes in dementia—positivity for amyloid beta deposition, medial temporal lobe atrophy, and white matter hyperintensity—using plasma protein biomarkers.

  • NeuroFANN provides biologically informed risk predictions for dementia subtypes by incorporating the protein–protein interaction (PPI) network and functional annotation of protein biomarkers.

  • By integrating the PPI network with independent effects of proteins, NeuroFANN captures synergetic effects through globality-based feature aggregation, effectively representing the complex properties of PPIs.

  • NeuroFANN employs functional annotation-based protein clustering, enabling a biologically meaningful and balanced analysis of dementia-related processes.

  • Experimental validation demonstrates NeuroFANN’s superior classification performance, showcasing its ability to identify biologically relevant proteins and generate interpretable results.

Supplementary Material

NeuroFANN_Supplement_bbaf366

Contributor Information

Sunghong Park, Department of Physiology, Ajou University School of Medicine, Worldcup-ro 164, Yeongtong-gu, Suwon, 16499, Republic of Korea.

Doyoon Kim, Department of Physiology, Ajou University School of Medicine, Worldcup-ro 164, Yeongtong-gu, Suwon, 16499, Republic of Korea.

Ji-Hye Choi, Department of Physiology, Ajou University School of Medicine, Worldcup-ro 164, Yeongtong-gu, Suwon, 16499, Republic of Korea.

Chang Hyung Hong, Department of Psychiatry, Ajou University School of Medicine, Worldcup-ro 164, Yeongtong-gu, Suwon, 16499, Republic of Korea.

Sang Joon Son, Department of Psychiatry, Ajou University School of Medicine, Worldcup-ro 164, Yeongtong-gu, Suwon, 16499, Republic of Korea.

Hyun Woong Roh, Department of Psychiatry, Ajou University School of Medicine, Worldcup-ro 164, Yeongtong-gu, Suwon, 16499, Republic of Korea.

Hyunjung Shin, Department of Industrial Engineering, Ajou University, Worldcup-ro 206, Yeongtong-gu, Suwon, 16499, Republic of Korea; Department of Artificial Intelligence, Ajou University, Worldcup-ro 206, Yeongtong-gu, Suwon, 16499, Republic of Korea.

Hyun Goo Woo, Department of Physiology, Ajou University School of Medicine, Worldcup-ro 164, Yeongtong-gu, Suwon, 16499, Republic of Korea; Department of Biomedical Science, Graduate School of Ajou University, Worldcup-ro 164, Yeongtong-gu, Suwon, 16499, Republic of Korea; Ajou Translational Omics Center, Research Institute for Innovative Medicine, Ajou University Medical Center, Worldcup-ro 164, Yeongtong-gu, Suwon, 16499, Republic of Korea.

Author contributions

Sunghong Park (Conceptualization, Formal analysis, Funding acquisition, Investigation, Methodology, Software, Validation, Visualization, Writing—original draft, Writing—review & editing), Doyoon Kim (Conceptualization, Formal analysis, Methodology, Investigation, Validation, Writing – original draft, Writing—review & editing), Ji-Hye Choi (Investigation, Investigation, Validation, Writing—original draft, Writing—review & editing), Chang Hyung Hong (Data Curation, Funding acquisition), Sang Joon Son (Data Curation), Hyun Woong Roh (Data Curation), Hyunjung Shin (Conceptualization, Funding acquisition, Methodology, Supervision, Writing—original draft, Writing—review & editing), Hyun Goo Woo (Conceptualization, Funding acquisition, Supervision, Writing—original draft, Writing—review & editing).

Conflict of interest

None declared.

Funding

This study was conducted using biospecimens and data from the Biobank Innovations for Chronic Cerebrovascular Disease with ALZheimer’s Disease Study (BICWALZS) consortium funded by the National Institute of Health (NIH), Republic of Korea (2024-ER0505-01). This study was supported by the Basic Science Research Program through the National Research Foundation of Korea (NRF) funded by the Ministry of Education (MOE), Republic of Korea (2022R1A6A3A01086784) and Ajou University Research Fund. This work was also supported by the NRF grants funded by the Ministry of Science and ICT (MSIT), Republic of Korea (2019R1A5A2026045, RS-2022-001653, and 2021R1A2C2003474), the Institute of Information & Communications Technology Planning & Evaluation grant funded by the MSIT (RS-2023-00255968), grants of Korea Health Technology R&D Project through Korea Health Industry Development Institute (KHIDI) funded by the Ministry of Health and Welfare, Republic of Korea (RS-2021-KH113821 and RS-2024-00407544), and a grant of ‘Korea Government Grant Program for Education and Research in Medical AI’ through the KHIDI funded by the MOE and MOHW.

Data availability

The data used in this study are available upon request from the BICWALZS consortium biobank (http://www.bicwalzs.com) as well as the National Biobank of Korea (https://biobank.nih.go.kr) which is a central biobank supporting for the Korea Biobank Project. The source codes of this study are available at https://github.com/sunghongpark-ai/NeuroFANN.

References

  • 1. Elahi  FM, Miller  BL. A clinicopathological approach to the diagnosis of dementia. Nat Rev Neurol  2017;13:457–76. 10.1038/nrneurol.2017.96 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 2. Hebert  LE, Weuve  J, Scherr  PA. et al.  Alzheimer disease in the United States (2010–2050) estimated using the 2010 census. Neurology  2013;80:1778–83. 10.1212/WNL.0b013e31828726f5 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 3. Hampel  H, Cummings  J, Blennow  K. et al.  Developing the ATX (N) classification for use across the Alzheimer disease continuum. Nat Rev Neurol  2021;17:580–9. 10.1038/s41582-021-00520-w [DOI] [PubMed] [Google Scholar]
  • 4. Van Der Flier  WM, Scheltens  P. The ATN framework—moving preclinical Alzheimer disease to clinical relevance. JAMA Neurol  2022;79:968–70. 10.1001/jamaneurol.2022.2967 [DOI] [PubMed] [Google Scholar]
  • 5. Chun  MY, Jang  H, Kim  SJ. et al.  Emerging role of vascular burden in AT (N) classification in individuals with Alzheimer’s and concomitant cerebrovascular burdens. J Neurol Neurosurg Psychiatry  2024;95:44–51. 10.1136/jnnp-2023-331603 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 6. Devanand  DP, Lee  S, Huey  ED. et al.  Associations between neuropsychiatric symptoms and neuropathological diagnoses of Alzheimer disease and related dementias. JAMA Psychiatry  2022;79:359–67. 10.1001/jamapsychiatry.2021.4363 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 7. Pillai  JA, Bena  J, Rothenberg  K. et al.  Association of variation in behavioral symptoms with initial cognitive phenotype in adults with dementia confirmed by neuropathology. JAMA Netw Open  2022;5:e220729–9. 10.1001/jamanetworkopen.2022.0729 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 8. Hampel  H, Hardy  J, Blennow  K. et al.  The amyloid-β pathway in Alzheimer’s disease. Mol Psychiatry  2021;26:5481–503. 10.1038/s41380-021-01249-0 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 9. Gallingani  C, Carbone  C, Tondelli  M. et al.  Neurofilaments light chain in neurodegenerative dementias: a review of imaging correlates. Brain Sci  2024;14:272. 10.3390/brainsci14030272 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 10. Jaime Garcia  D, Clancy  U, Arteaga  C. et al.  Blood biomarkers of vascular dysfunction in small vessel disease progression: insights from a longitudinal neuroimaging study. Alzheimers Dement  2025;21:e70152. 10.1002/alz.70152 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 11. Jack  CR  Jr, Bennett  DA, Blennow  K. et al.  NIA-AA research framework: toward a biological definition of Alzheimer’s disease. Alzheimers Dement  2018;14:535–62. 10.1016/j.jalz.2018.02.018 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 12. Shen  J, Chiew  H, Ng  A. et al.  White matter hyperintensity as a vascular contribution to the AT (N) framework. J Prev Alzheimers Dis  2023;10:387–400. 10.14283/jpad.2023.53 [DOI] [PubMed] [Google Scholar]
  • 13. Therriault  J, Vermeiren  M, Servaes  S. et al.  Association of phosphorylated tau biomarkers with amyloid positron emission tomography vs tau positron emission tomography. JAMA Neurol  2023;80:188–99. 10.1001/jamaneurol.2022.4485 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 14. Palmqvist  S, Mattsson  N, Hansson  O. et al.  Cerebrospinal fluid analysis detects cerebral amyloid-β accumulation earlier than positron emission tomography. Brain  2016;139:1226–36. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 15. Schindler  SE, Gray  JD, Gordon  BA. et al.  Cerebrospinal fluid biomarkers measured by Elecsys assays compared to amyloid imaging. Alzheimers Dement  2018;14:1460–9. 10.1016/j.jalz.2018.01.013 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 16. Guo  Y, You  J, Zhang  Y. et al.  Plasma proteomic profiles predict future dementia in healthy adults. Nature Aging  2024;4:247–60. 10.1038/s43587-023-00565-0 [DOI] [PubMed] [Google Scholar]
  • 17. Imbimbo  BP, Watling  M, Imbimbo  C. et al.  Plasma ATN (I) classification and precision pharmacology in Alzheimer’s disease. Alzheimers Dement  2023;19:4729–34. 10.1002/alz.13084 [DOI] [PubMed] [Google Scholar]
  • 18. Jin  Y, Xu  Z, Zhang  Y. et al.  Serum/plasma biomarkers and the progression of cardiometabolic multimorbidity: a systematic review and meta-analysis. Front Public Health  2023;11:1280185. 10.3389/fpubh.2023.1280185 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 19. Soni  RK. Frontiers in plasma proteome profiling platforms: innovations and applications. Clin Proteomics  2024;21:43. 10.1186/s12014-024-09497-2 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 20. Nakamura  A, Kaneko  N, Villemagne  VL. et al.  High performance plasma amyloid-β biomarkers for Alzheimer’s disease. Nature  2018;554:249–54. 10.1038/nature25456 [DOI] [PubMed] [Google Scholar]
  • 21. Morgan  AR, Touchard  S, Leckey  C. et al.  Inflammatory biomarkers in Alzheimer's disease plasma. Alzheimers Dement  2019;15:776–87. 10.1016/j.jalz.2019.03.007 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 22. Palmqvist  S, Tideman  P, Cullen  N. et al.  Prediction of future Alzheimer’s disease dementia using plasma phospho-tau combined with other accessible measures. Nat Med  2021;27:1034–42. 10.1038/s41591-021-01348-z [DOI] [PubMed] [Google Scholar]
  • 23. Mattsson  N, Cullen  NC, Andreasson  U. et al.  Association between longitudinal plasma neurofilament light and neurodegeneration in patients with Alzheimer disease. JAMA Neurol  2019;76:791–9. 10.1001/jamaneurol.2019.0765 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 24. Zhu  N, Santos-Santos  M, Illán-Gala  I. et al.  Plasma glial fibrillary acidic protein and neurofilament light chain for the diagnostic and prognostic evaluation of frontotemporal dementia. Transl Neurodegener  2021;10:50. 10.1186/s40035-021-00275-w [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 25. Heneka  MT, Gauthier  S, Chandekar  SA. et al.  Neuroinflammatory fluid biomarkers in patients with Alzheimer’s disease: a systematic literature review. Mol Psychiatry  2025;30:2783–98. 10.1038/s41380-025-02939-9 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 26. Jaime Garcia  D, Chagnot  A, Wardlaw  JM. et al.  A scoping review on biomarkers of endothelial dysfunction in small vessel disease: molecular insights from human studies. Int J Mol Sci  2023;24:13114. 10.3390/ijms241713114 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 27. Wan  S, Dandu  C, Han  G. et al.  Plasma inflammatory biomarkers in cerebral small vessel disease: a review. CNS Neurosci Ther  2023;29:498–515. 10.1111/cns.14047 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 28. Chen  JXY, Vipin  A, Sandhu  GK. et al.  Blood-brain barrier integrity disruption is associated with both chronic vascular risk factors and white matter hyperintensities. J Prev Alzheimers Dis  2025;12:100029. 10.1016/j.tjpad.2024.100029 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 29. Gao  F, Lv  X, Dai  L. et al.  A combination model of AD biomarkers revealed by machine learning precisely predicts Alzheimer’s dementia: China aging and neurodegenerative initiative (CANDI) study. Alzheimers Dement  2023;19:749–60. 10.1002/alz.12700 [DOI] [PubMed] [Google Scholar]
  • 30. Winchester  LM, Harshfield  EL, Shi  L. et al.  Artificial intelligence for biomarker discovery in Alzheimer’s disease and dementia. Alzheimers Dement  2023;19:5860–71. 10.1002/alz.13390 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 31. Park  S, Hong  CH, Son  SJ. et al.  Identification of molecular subtypes of dementia by using blood-proteins interaction-aware graph propagational network. Brief Bioinform  2024;25:bbae428. 10.1093/bib/bbae428 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 32. Cai  Y, Fan  X, Zhao  L. et al.  Comparing machine learning-derived MRI-based and blood-based neurodegeneration biomarkers in predicting syndromal conversion in early AD. Alzheimers Dement  2023;19:4987–98. 10.1002/alz.13083 [DOI] [PubMed] [Google Scholar]
  • 33. Chiu  S-I, Fan  LY, Lin  CH. et al.  Machine learning-based classification of subjective cognitive decline, mild cognitive impairment, and Alzheimer’s dementia using neuroimage and plasma biomarkers. ACS Chem Nerosci  2022;13:3263–70. 10.1021/acschemneuro.2c00255 [DOI] [PubMed] [Google Scholar]
  • 34. Wu  J, Chen  Z, Xiao  S. et al.  DeepMoIC: multi-omics data integration via deep graph convolutional networks for cancer subtype classification. BMC Genomics  2024;25:1–13. 10.1186/s12864-024-11112-5 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 35. Li  B, Wang  T, Nabavi  S. Cancer molecular subtype classification by graph convolutional networks on multi-omics data. In: Proceedings of the 12th ACM International Conference on Bioinformatics, Computational Biology, and Health Informatics, pp. 1–9. ACM, 2021. 10.1145/3459930.3469542 [DOI]
  • 36. Jiang  W, Ye  W, Tan  X. et al.  Network-based multi-omics integrative analysis methods in drug discovery: a systematic review. BioData Mining  2025;18:27. 10.1186/s13040-025-00442-z [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 37. Smirnov  DS, Ashton  NJ, Blennow  K. et al.  Plasma biomarkers for Alzheimer’s disease in relation to neuropathology and cognitive change. Acta Neuropathol  2022;143:487–503. 10.1007/s00401-022-02408-5 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 38. Vrillon  A, Bousiges  O, Götze  K. et al.  Plasma biomarkers of amyloid, tau, axonal, and neuroinflammation pathologies in dementia with Lewy bodies. Alzheimer's Research & Therapy  2024;16:146. 10.1186/s13195-024-01502-y [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 39. Ersoz  NS, Bakir-Gungor  B, Yousef  M. GeNetOntology: identifying affected gene ontology terms via grouping, scoring, and modeling of gene expression data utilizing biological knowledge-based machine learning. Front Genet  2023;14:1139082. 10.3389/fgene.2023.1139082 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 40. Sartori  F, Codicè  F, Caranzano  I. et al.  A comprehensive review of deep learning applications with multi-omics data in cancer research. Genes  2025;16:648. 10.3390/genes16060648 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 41. Nguyen  Q-H, Nguyen  H, Oh  EC. et al.  Current approaches and outstanding challenges of functional annotation of metabolites: a comprehensive review. Brief Bioinform  2024;25:bbae498. 10.1093/bib/bbae498 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 42. Ritchie  ME, Phipson  B, Wu  D. et al.  Limma powers differential expression analyses for RNA-sequencing and microarray studies. Nucleic Acids Res  2015;43:e47–7. 10.1093/nar/gkv007 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 43. Bottou  L, Curtis  FE, Nocedal  J. Optimization methods for large-scale machine learning. Siam Review  2018;60:223–311. 10.1137/16M1080173 [DOI] [Google Scholar]
  • 44. Roh  HW, Kim  NR, Lee  DG. et al.  Baseline clinical and biomarker characteristics of biobank innovations for chronic cerebrovascular disease with Alzheimer’s disease study: BICWALZS. Psychiatry Investig  2022;19:100–9. 10.30773/pi.2021.0335 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 45. Hwang  J, Jeong  JH, Yoon  SJ. et al.  Clinical and biomarker characteristics according to clinical spectrum of Alzheimer’s disease (AD) in the validation cohort of Korean brain aging study for the early diagnosis and prediction of AD. J Clin Med  2019;8:341. 10.3390/jcm8030341 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 46. Scheltens  P, Barkhof  F, Leys  D. et al.  A semiquantative rating scale for the assessment of signal hyperintensities on magnetic resonance imaging. J Neurol Sci  1993;114:7–12. 10.1016/0022-510X(93)90041-V [DOI] [PubMed] [Google Scholar]
  • 47. Fazekas  F, Chawluk  JB, Alavi  A. et al.  MR signal abnormalities at 1.5 T in Alzheimer's dementia and normal aging. Am J Roentgenol  1987;149:351–6. 10.2214/ajr.149.2.351 [DOI] [PubMed] [Google Scholar]
  • 48. Harris  SE, Cox  SR, Bell  S. et al.  Neurology-related protein biomarkers are associated with cognitive ability and brain volume in older age. Nat Commun  2020;11:800. 10.1038/s41467-019-14161-7 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 49. Cano  A, Esteban-de-Antonio  E, Bernuz  M. et al.  Plasma extracellular vesicles reveal early molecular differences in amyloid positive patients with early-onset mild cognitive impairment. J Nanobiotechnol  2023;21:54. 10.1186/s12951-023-01793-7 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 50. Latif-Hernandez  A, Yang  T, Butler  RR  III. et al.  A TrkB and TrkC partial agonist restores deficits in synaptic function and promotes activity-dependent synaptic and microglial transcriptomic changes in a late-stage Alzheimer's mouse model. Alzheimers Dement  2024;20:4434–60. 10.1002/alz.13857 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 51. Ledda  F, Paratcha  G, Sandoval-Guzmán  T. et al.  GDNF and GFRα1 promote formation of neuronal synapses by ligand-induced cell adhesion. Nat Neurosci  2007;10:293–300. 10.1038/nn1855 [DOI] [PubMed] [Google Scholar]
  • 52. Chiavellini  P, Canatelli-Mallat  M, Lehmann  M. et al.  Therapeutic potential of glial cell line-derived neurotrophic factor and cell reprogramming for hippocampal-related neurological disorders. Neural Regen Res  2022;17:469–76. 10.4103/1673-5374.320966 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 53. Yi  L, Xu  Y, Tong  S. Serum glial cell line-derived neurotrophic factor: a potential biomarker for white matter alteration in Parkinson’s disease with mild cognitive impairment. Front Neurosci  2024;18:1370787. [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 54. Calsolaro  V, Edison  P. Neuroinflammation in Alzheimer’s disease: current evidence and future directions. Alzheimers Dement  2016;12:719–32. 10.1016/j.jalz.2016.02.010 [DOI] [PubMed] [Google Scholar]
  • 55. Butterfield  DA, Halliwell  B. Oxidative stress, dysfunctional glucose metabolism and Alzheimer disease. Nat Rev Neurosci  2019;20:148–60. 10.1038/s41583-019-0132-6 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 56. Kipf  TN, Welling  M. Semi-supervised classification with graph convolutional networks. arXiv preprint arXiv  2016;1609.02907, preprint: not peer reviewed. [Google Scholar]
  • 57. Wu  F, Souza  A, Zhang  T. et al.  Simplifying graph convolutional networks. In: International Conference on Machine Learning, pp. 6861–71. PMLR, 2019. https://proceedings.mlr.press/v97/wu19e [Google Scholar]
  • 58. Pasa  L, Navarin  N, Erb  W. et al.  Empowering simple graph convolutional networks. IEEE Trans Neural Networks Learn Syst  2023;35:4385–99. 10.1038/s41583-019-0132-6 [DOI] [PubMed] [Google Scholar]
  • 59. Abu-El-Haija  S, Perozzi  B, Kapoor  A. et al.  Mixhop: Higher-order graph convolutional architectures via sparsified neighborhood mixing. In: International Conference on Machine Learning, pp. 21–9. PMLR, 2019. https://proceedings.mlr.press/v97/abu-el-haija19a.html [Google Scholar]
  • 60. Jin  D. et al.  Universal graph convolutional networks. Adv Neural Inform Process Syst  2021;34:10654–64. [Google Scholar]
  • 61. Wang  J, Liang  J, Cui  J. et al.  Semi-supervised learning with mixed-order graph convolutional networks. Inform Sci  2021;573:171–81. 10.1016/j.ins.2021.05.057 [DOI] [Google Scholar]
  • 62. Lundberg  SM, Lee  S-I. A unified approach to interpreting model predictions. Adv Neural Inform Process Syst  2017;30:1–10. [Google Scholar]
  • 63. Walker  KA, Chen  J, Shi  L. et al.  Proteomics analysis of plasma from middle-aged adults identifies protein markers of dementia risk in later life. Sci Transl Med  2023;15:eadf5681. 10.1126/scitranslmed.adf5681 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 64. Gong  J, Williams  DM, Scholes  S. et al.  Unraveling the role of plasma proteins in dementia: insights from two cohort studies in the UK, with causal evidence from Mendelian randomization. medRxiv  2024;2024.06. 04.24308415, preprint: not peer reviewed. [Google Scholar]
  • 65. Elmarakeby  HA, Hwang  J, Arafeh  R. et al.  Biologically informed deep neural network for prostate cancer discovery. Nature  2021;598:348–52. 10.1038/s41586-021-03922-4 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 66. Hartman  E, Scott  AM, Karlsson  C. et al.  Interpreting biologically informed neural networks for enhanced proteomic biomarker discovery and pathway analysis. Nat Commun  2023;14:5359. 10.1038/s41467-023-41146-4 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 67. Rossor  MN, Fox  NC, Mummery  CJ. et al.  The diagnosis of young-onset dementia. Lancet Neurol  2010;9:793–806. 10.1016/S1474-4422(10)70159-9 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 68. Ng  TP, Feng  L, Nyunt  MSZ. et al.  Metabolic syndrome and the risk of mild cognitive impairment and progression to dementia: follow-up of the Singapore longitudinal ageing study cohort. JAMA Neurol  2016;73:456–63. 10.1001/jamaneurol.2015.4899 [DOI] [PubMed] [Google Scholar]
  • 69. Marzola  P, Melzer  T, Pavesi  E. et al.  Exploring the role of neuroplasticity in development, aging, and neurodegeneration. Brain Sci  2023;13:1610. 10.3390/brainsci13121610 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 70. Nguyen  PT, Dorman  LC, Pan  S. et al.  Microglial remodeling of the extracellular matrix promotes synapse plasticity. Cell  2020;182:388–403.e15. 10.1016/j.cell.2020.05.050 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 71. Sun  Y, Xu  S, Jiang  M. et al.  Role of the extracellular matrix in Alzheimer’s disease. Front Aging Neurosci  2021;13:707466. 10.3389/fnagi.2021.707466 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 72. Ransohoff  RM. How neuroinflammation contributes to neurodegeneration. Science  2016;353:777–83. 10.1126/science.aag2590 [DOI] [PubMed] [Google Scholar]
  • 73. Su  J, Song  Y, Zhu  Z. et al.  Cell–cell communication: new insights and clinical implications. Signal Transduct Target Ther  2024;9:196. 10.1038/s41392-024-01888-z [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 74. Otnæss  MK, Djurovic  S, Rimol  LM. et al.  Evidence for a possible association of neurotrophin receptor (NTRK-3) gene polymorphisms with hippocampal function and schizophrenia. Neurobiol Dis  2009;34:518–24. 10.1016/j.nbd.2009.03.011 [DOI] [PubMed] [Google Scholar]
  • 75. Braskie  MN, Kohannim  O, Jahanshad  N. et al.  Relation between variants in the neurotrophin receptor gene, NTRK3, and white matter integrity in healthy young adults. Neuroimage  2013;82:146–53. 10.1016/j.neuroimage.2013.05.095 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 76. Konishi  Y, Yang  LB, He  P. et al.  Deficiency of GDNF receptor GFRα1 in Alzheimer's neurons results in neuronal death. J Neurosci  2014;34:13127–38. 10.1523/JNEUROSCI.2582-13.2014 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 77. Bonafina  A, Trinchero  MF, Ríos  AS. et al.  GDNF and GFRα1 are required for proper integration of adult-born hippocampal neurons. Cell Rep  2019;29:4308–4319.e4. 10.1016/j.celrep.2019.11.100 [DOI] [PubMed] [Google Scholar]
  • 78. Cocco  E, Scaltriti  M, Drilon  A. NTRK fusion-positive cancers and TRK inhibitor therapy. Nat Rev Clin Oncol  2018;15:731–47. 10.1038/s41571-018-0113-0 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 79. Airaksinen  MS, Saarma  M. The GDNF family: signalling, biological functions and therapeutic value. Nat Rev Neurosci  2002;3:383–94. 10.1038/nrn812 [DOI] [PubMed] [Google Scholar]
  • 80. Procaccini  C, Santopaolo  M, Faicchia  D. et al.  Role of metabolism in neurodegenerative disorders. Metabolism  2016;65:1376–90. 10.1016/j.metabol.2016.05.018 [DOI] [PubMed] [Google Scholar]
  • 81. Minta  K, Brinkmalm  G, Portelius  E. et al.  Brevican and neurocan peptides as potential cerebrospinal fluid biomarkers for differentiation between vascular dementia and Alzheimer’s disease. J Alzheimers Dis  2021;79:729–41. 10.3233/JAD-201039 [DOI] [PubMed] [Google Scholar]
  • 82. Dauar  MT, Labonté  A, Picard  C. et al.  Characterization of the contactin 5 protein and its risk-associated polymorphic variant throughout the Alzheimer’s disease spectrum. Alzheimers Dement  2023;19:2816–30. 10.1002/alz.12868 [DOI] [PubMed] [Google Scholar]
  • 83. Hosoki  S, Hansra  GK, Jayasena  T. et al.  Molecular biomarkers for vascular cognitive impairment and dementia. Nat Rev Neurol  2023;19:737–53. 10.1038/s41582-023-00884-1 [DOI] [PubMed] [Google Scholar]
  • 84. Chu  M, Wen  L, Jiang  D. et al.  Peripheral inflammation in behavioural variant frontotemporal dementia: associations with central degeneration and clinical measures. J Neuroinflammation  2023;20:65. 10.1186/s12974-023-02746-5 [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 85. Chen  G, Kang  SS, Wang  Z. et al.  Netrin-1 receptor UNC5C cleavage by active δ-secretase enhances neurodegeneration, promoting Alzheimer’s disease pathologies. Sci Adv  2021;7:eabe4499. 10.1126/sciadv.abe4499 [DOI] [PMC free article] [PubMed] [Google Scholar]

Associated Data

This section collects any data citations, data availability statements, or supplementary materials included in this article.

Supplementary Materials

NeuroFANN_Supplement_bbaf366

Data Availability Statement

The data used in this study are available upon request from the BICWALZS consortium biobank (http://www.bicwalzs.com) as well as the National Biobank of Korea (https://biobank.nih.go.kr) which is a central biobank supporting for the Korea Biobank Project. The source codes of this study are available at https://github.com/sunghongpark-ai/NeuroFANN.


Articles from Briefings in Bioinformatics are provided here courtesy of Oxford University Press

RESOURCES