Skip to main content
Scientific Reports logoLink to Scientific Reports
. 2025 Dec 15;16:1900. doi: 10.1038/s41598-025-31486-0

Developing a QSPR model for Alzheimer’s drugs using topological indices and M-polynomial: A computational study

Mohammad Hadi Akhbari 1,✉, Fateme Movahedi 2, Mahsa Zameni 3, Mohammad Hassan Shahavi 3,✉
PMCID: PMC12804841  PMID: 41392302

Abstract

Topological indices, which are numerical descriptors that encode molecular structure, are widely used in computational drug discovery due to their efficiency and interpretability. In this study, we developed a robust quantitative structure–property relationship (QSPR) framework to predict the core physicochemical properties of nine clinically relevant Alzheimer’s disease drugs, including Donepezil, Galantamine, and Memantine.​ We employed a streamlined computational approach, using MATLAB and the M-polynomial method, to efficiently calculate a series of degree-based topological indices. Through comprehensive regression analyses, we identified strong correlations between degree-based topological indices and key physicochemical properties, including boiling point and molar refractivity. While linear models provided a reasonable baseline, nonlinear models, particularly cubic and power equations, delivered significantly improved predictive accuracy. The analysis highlighted the critical interplay between the choice of the index and the regression model. For instance, the cubic model was frequently the most effective for predicting properties such as boiling point and flash point, while the power model performed best for molar refractivity and polarizability. Notably, the redefined first Zagreb index and the modified first Zagreb index exhibited exceptional predictive capacity, reflecting their sensitivity to structural features that govern physicochemical behavior. The strong performance of these QSPR models underscores their potential to accelerate the rational design of Alzheimer’s therapeutics. By enabling rapid, cost-effective, and reliable property prediction prior to synthesis, this framework offers a valuable tool for future drug development efforts.

Supplementary Information

The online version contains supplementary material available at 10.1038/s41598-025-31486-0.

Keywords: QSPR model, M-polynomial, Topological index, Regression model, Physicochemical property

Subject terms: Applied mathematics, Molecular biology, Mathematics and computing, Molecular medicine

Introduction

Alzheimer’s disease is a neurodegenerative disorder with an insidious onset and a progressively worsening course. It accounts for roughly 60–70% of dementia cases worldwide. The most rampant initial symptoms are forgetting recent events and short-term memory problems, as the disease affects the brain’s nerve function and communication, leading to the gradual destruction of brain function1,2. In the advanced stages, the disease’s impact becomes more severe. Patients often become fully dependent on caregivers for activities of daily living and may develop aphasia, akinesia, or severe motor impairment. Accompanying these physical declines are profound psychological effects, including depression, apathy, and fatigue3,4. Alzheimer’s is more likely to occur in women and people over the age of 655. Current statistics indicate that approximately 50 million people worldwide suffer from Alzheimer’s and other forms of dementia, with projections suggesting this number could exceed 150 million by 20506,7.

The cause of most Alzheimer’s cases is still largely unknown, except for the 1 to 2% of cases in which determinant genetic differences have been identified. At present, the treatment of Alzheimer’s disease is focused on managing symptoms rather than providing a cure. Therapeutic strategies include symptomatic treatments, management of behavioral disorders, and drugs designed to slow the progression of the disease. While no curative therapy is yet available, several approved drugs can effectively delay disease progression and attenuate memory impairment and behavioral disturbances in some patients8.

Computational chemistry has emerged as a cornerstone of modern drug discovery, providing systematic frameworks to address the complexities of multifactorial pathologies such as Alzheimer’s disease. Within this domain, Quantitative Structure-Activity/Property Relationship (QSAR/QSPR) modeling represents a compelling strategy. These models establish a predictive link between a molecule’s structure and its biological or physicochemical function. By using mathematical descriptors to encode chemical information numerically, QSAR/QSPR enables the rapid estimation of a compound’s properties, thereby accelerating the identification and optimization of potential new drugs9,10.

Topological indices are numerical descriptors derived from a molecule’s graph representation and encode critical structural information, such as branching, size, and atom connectivity. By distilling these complex structural features into single, quantitative values, they provide a powerful basis for predicting a compound’s physicochemical behavior and biological activity within QSAR and QSPR frameworks11. This approach is particularly valuable in Alzheimer’s drug discovery. The process of screening vast chemical libraries for potential therapeutic candidates is a significant logistical and financial bottleneck in the development of new treatments. By enabling rapid, computational prediction of a compound’s properties, topological indices offer an efficient strategy to prioritize candidates and streamline the discovery process.

By establishing a mathematical correlation between a molecule’s structure and its properties, QSPR models based on topological indices can predict key attributes of potential Alzheimer’s disease drugs before they are synthesized and tested experimentally. This predictive capability extends to a range of critical physicochemical properties, including molecular weight, boiling point, and polar surface area, as well as to indicators of therapeutic efficacy. The application of this predictive approach offers significant advantages for drug discovery. It not only accelerates the identification of promising drug candidates but also facilitates the rational design and optimization of novel therapeutics for Alzheimer’s disease. Ultimately, integrating these computational models into the development pipeline saves considerable time and resources, streamlining the path from initial concept to clinical application12,13.

Topological indices are established as cost-effective and powerful descriptors in chemoinformatics, particularly for estimating the physicochemical properties of small-molecule drugs14–16. Indeed, QSPR models built upon these indices have been successfully applied to predict key molecular attributes for a wide range of therapeutic agents, including those targeting cancer, hypertension, and cardiovascular diseases. This proven utility motivates the application of topological indices to characterize the structurally diverse compounds used in the management of Alzheimer’s disease17,18.

The utility of topological indices in QSPR modeling is well documented across diverse scientific domains. These numerical descriptors have been successfully employed to predict key properties in fields from materials science, such as characterizing boron-based nanomaterials, to chemoinformatics and nanotechnology19–21. Furthermore, within pharmacology, topological indices have proven instrumental for modeling compounds targeting a wide array of diseases, including viral infections and various forms of cancer22–24. This broad applicability underscores the immense potential of QSPR models to transform and expedite preclinical research. By providing a cost-effective, rapid framework for screening and optimizing chemical structures, these computational models reduce reliance on time-consuming, expensive laboratory experiments. This acceleration of the research and development pipeline is a critical advantage in the quest for novel therapeutics25–27.

Recent advancements in the field have demonstrated that integrating topological indices with machine learning models, such as neural networks and random forests, can significantly enhance the predictive accuracy for key molecular properties28. This synergy between classic descriptors and modern algorithms represents a promising frontier for developing more sophisticated and precise QSPR models.

Concurrently, other lines of research have focused on developing novel descriptors. For instance, newly proposed topological indices have been used to effectively model the structure-activity relationships of synthesized amide derivatives, demonstrating their potential as valuable tools in the search for anti-Alzheimer’s agents29.

Alongside these advanced computational methods, studies continue to affirm the value of traditional approaches. Recent work has shown that even conventional linear QSPR models can provide reliable and useful estimates of physicochemical and biological attributes for novel compounds being investigated for the treatment of Alzheimer’s disease30.

This study employs the M-polynomial formalism, in conjunction with edge-partitioning techniques, to systematically compute a panel of degree-based topological indices for the selected Alzheimer’s drugs. The M-polynomial serves as a powerful algebraic tool that elegantly encodes degree-based structural information of a molecular graph into a single bivariate polynomial function.

The primary advantage of this approach lies in its computational efficiency. Once the M-polynomial for a given molecular structure is derived, a wide array of degree-based topological indices can be obtained directly through simple algebraic and differential operations. This method thereby obviates the need for redundant, index-by-index calculations, creating a more streamlined and efficient workflow for characterizing chemical structures31.

In pharmacology, M-polynomial-based approaches have been instrumental in analyzing the physicochemical properties and structural features of various therapeutic agents. This method has been successfully used to model antiviral drugs targeting pathogens such as SARS-CoV-2, as well as anti-tumor agents such as cyclodextrin derivatives32–34. Furthermore, polynomial mathematical models are widely used to optimize the development of nanotechnology-based products, including nanoemulsions and nanoparticles35–40.

The broad utility of these techniques for predicting the properties of drugs and nanostructures is further evidenced41–43. This established precedent across multiple therapeutic and formulation domains motivates our application of the M-polynomial method to systematically characterize the physicochemical attributes of drugs used in Alzheimer’s disease therapy. Therefore, this study aims to systematically investigate the relationship between molecular structure and key physicochemical properties for a curated set of nine Alzheimer’s disease drugs.

The specific compounds selected for this analysis are ABT-089, ABT-126, Donepezil, Galantamine, Memantine, Metrifonate, Phencyclidine, Rivastigmine, and Tacrine44. The molecular structures of these compounds are depicted in Fig. 1.

Fig. 1.

Fig. 1

Chemical structures of Alzheimer’s drugs from ChemSpider45.

By leveraging a panel of degree-based topological indices (summarized in Table 1) derived via the M-polynomial method, we develop and validate a series of QSPR models. The central goal is to elucidate which structural features and topological descriptors are most predictive of key physicochemical attributes, thereby providing a robust computational framework to guide the rational design of future anti- Alzheimer’s disease therapeutic agents.

Table 1.

Relations of M-polynomial with some zagreb-type topological indices47–52.

Topological index Inline graphic Derivation from Inline graphic
First Zagreb Inline graphic Inline graphic Inline graphic
Second Zagreb Inline graphic Inline graphic Inline graphic
Third Zagreb Inline graphic Inline graphic Inline graphic
Redefined first Zagreb Inline graphic Inline graphic Inline graphic
Redefined second Zagreb Inline graphic Inline graphic Inline graphic
Redefined third Zagreb Inline graphic Inline graphic Inline graphic
Modified first Zagreb Inline graphic Inline graphic Inline graphic
Modified second Zagreb Inline graphic Inline graphic Inline graphic
Hyper first Zagreb Inline graphic Inline graphic Inline graphic
Hyper second Zagreb Inline graphic Inline graphic Inline graphic
Nano Zagreb Inline graphic Inline graphic Inline graphic
Augmented Zagreb Inline graphic Inline graphic Inline graphic

Methodology

This study centers on a set of nine previously mentioned drugs relevant to Alzheimer’s disease. The two-dimensional molecular graphs for these compounds, shown in Fig. 1, were used as the basis for all subsequent graph-theoretical analyses. Suppose that Inline graphic is a simple connected graph representing a molecular structure where Inline graphic is the set of vertices (atoms) and Inline graphic is the set of edges (bonds).

Two vertices Inline graphic and Inline graphic in Inline graphic are called adjacent if Inline graphic is a member of the set of edges of Inline graphic. The number of vertices adjacent to a vertex Inline graphic is called the degree of Inline graphic, which is denoted by Inline graphic The foundation for constructing the M-polynomial is the edge-partitioning method. This technique involves classifying the edges of the molecular graph into distinct sets based on the degrees of their endpoint vertices. For a graph Inline graphic, an edge Inline graphic belongs to the partition Inline graphic if the degree of the vertex Inline graphics Inline graphic and the degree of the vertex Inline graphic is Inline graphic.

By systematically partitioning all edges in this manner, we can accurately construct the M-polynomial, which then acts as a generator for the desired topological indices46. The M-polynomial of a graph is formally defined as:

graphic file with name d33e465.gif 1

where Inline graphic is the number of edges Inline graphic such that Inline graphic47.

This combined approach provides a rigorous, efficient, and transparent framework for analyzing the structural properties of complex molecules. Its systematic nature makes it particularly well-suited for the comparative analysis of the Alzheimer’s drugs investigated in this paper.

Table 1 shows Zagreb-type topological indices and their derivation from the M-polynomial Inline graphic, in which

graphic file with name d33e497.gif
graphic file with name d33e500.gif

The following 5-step algorithm was implemented to compute and analyze the topological indices defined in Table 1 using the M-polynomial method. Subsequently, this framework is used to develop and present a QSPR model that examines the correlation between these indices and the physicochemical properties of the selected Alzheimer’s drugs.

Step 1: Draw Inline graphic and calculate its adjacency matrix Inline graphic using TopoCluj.

Step 2: Determine the cardinality matrix Inline graphicfor the edge-partition method using our MATLAB program to input the matrix Inline graphic.

Step 3: Use our MATLAB program to input a matrix Inline graphic and compute M-polynomial Inline graphic for Alzheimer’s drugs.

Step 4: Compute Zagreb-type topological indices in Table 1 using our MATLAB code.

Step 5: Employ SPSS software to generate regression models and select the most appropriate model.

The pseudo-code of the proposed algorithm is presented in Appendix A.

Calculation of the M-polynomial of alzheimer’s drugs using MATLAB programming

This section details our main computational results. We utilize the M-polynomials to compute several topological indices. We analyze the molecular structures of Alzheimer’s drugs, represented as graphs in Fig. 1. By examining these molecular graphs, we extract the necessary information about the structural characteristics of these drugs to input into our MATLAB programming.

  • ABT-089: Let Inline graphic be the molecular graph of the ABT-089 drug with the molecular formulaInline graphic. This graph has order 14 and size 15 with Inline graphic and Inline graphic. We input the edge-partition matrix of the graph Inline graphic as Inline graphic in our MATLAB.

  • ABT-126: Suppose that Inline graphic is the graph of the molecular structure of the ABT-126 drug with the molecular formulaInline graphic. This graph contains 11 vertices, 10 edges, Inline graphic and Inline graphic. The edge-partition matrix of the graph Inline graphic is as Inline graphic.

  • Donepezil: The molecular graph of the drug Donepezil is denoted as Inline graphic, and it has the molecular formulaInline graphic. This graph contains 28 vertices and 31 edges with Inline graphic and Inline graphic. Also, we have Inline graphic.

  • Galantamine: Let Inline graphic be the molecular graph of the Galantamine drug with the molecular formulaInline graphic. This graph has 22 vertices and 25 edges with Inline graphic and Inline graphic. We obtain the edge-partition matrix of the graph Inline graphic as Inline graphic.

  • Memantine: We suppose that Inline graphic is the molecular graph of the Memantine drug with the molecular formulaInline graphic. This graph has 13 vertices and 15 edges with Inline graphic and Inline graphic. Furthermore, we obtain Inline graphic.

  • Metrifonate: Let Inline graphic be the molecular graph of the Metrifonate drug with the molecular formulaInline graphic. This graph has 12 vertices and 11 edges with Inline graphic and Inline graphic. We input the matrix Inline graphic in our MATLAB programming.

  • Phencyclidine: Let Inline graphic be the molecular graph of the Phencyclidine drug with the molecular formulaInline graphic. This graph has 18 vertices and 20 edges with Inline graphic and Inline graphic. Also, we have Inline graphic.

  • Rivastigmine: Let Inline graphic be the molecular graph of the Rivastigmine drug with the molecular formulaInline graphic. The graph Inline graphic has 17 vertices and 17 edges with Inline graphic and Inline graphic. The edge-partition matrix of the graph Inline graphic is as Inline graphic.

  • Tacrine: Let Inline graphic be the molecular graph of the Tacrine drug with the molecular formulaInline graphic. This graph has 15 vertices and 17 edges, which Inline graphic and Inline graphic. We input the edge-partition matrix of the graph Inline graphic as Inline graphic in the MATLAB code.

The following theorem obtains the M-polynomial expression for these drugs.

Theorem 1

For the molecular graphs of the Alzheimer’s drugs described above,

i)

graphic file with name d33e1113.gif

ii)

graphic file with name d33e1117.gif

iii)

graphic file with name d33e1121.gif

iv)

graphic file with name d33e1125.gif

v)

graphic file with name d33e1129.gif

vi)

graphic file with name d33e1133.gif

vii)

graphic file with name d33e1138.gif

viii)

graphic file with name d33e1142.gif

ix)

graphic file with name d33e1146.gif

Proof

i) Since Inline graphic, for the molecular graph of the ABT-089 drug, we have Inline graphic, Inline graphic, Inline graphic, and Inline graphic. Therefore, using the definition (1), we get

graphic file with name d33e1174.gif
graphic file with name d33e1177.gif
graphic file with name d33e1180.gif

Similar to the proof of the case Inline graphic, the remaining cases are proved using relation (1) and the obtained edge-partition matrix Inline graphic. ■

Illustrative example: calculation for the ABT-089 drug

In this section, we provide a detailed, step-by-step manual calculation of the twelve Zagreb-type topological indices listed in Table 1 of the ABT-089 drug. These calculations verify the results from our MATLAB code and illustrate the practical application of the M-polynomial method. The molecular graph of ABT-089, denoted as Inline graphic, has the following M-polynomial, which was derived in Theorem 1. The M-polynomial for the molecular graph of ABT-089 (Inline graphic) is:

graphic file with name d33e1211.gif

To compute the topological indices, we use the operators defined in Table 1, where Inline graphic. The key operators are:

graphic file with name d33e1223.gif
graphic file with name d33e1226.gif
graphic file with name d33e1231.gif
graphic file with name d33e1235.gif

Below are the manual calculations for each of the 12 topological indices.

  • First Zagreb index:

The formula is Inline graphic. Therefore, we get.

graphic file with name d33e1258.gif

Evaluating at Inline graphic:

graphic file with name d33e1268.gif
  • Second Zagreb Index: The formula is Inline graphic. Therefore, we get
    graphic file with name d33e1283.gif
    graphic file with name d33e1286.gif
    graphic file with name d33e1289.gif

Evaluating at Inline graphic:

graphic file with name d33e1299.gif
  • Third Zagreb Index: The formula is Inline graphic. Therefore, we get
    graphic file with name d33e1314.gif

Evaluating at Inline graphic:

graphic file with name d33e1324.gif
  • Redefined First Zagreb Index: The formula is Inline graphic. Therefore, we have
    graphic file with name d33e1340.gif

Evaluating at Inline graphic:

graphic file with name d33e1350.gif
  • Redefined Second Zagreb Index:

  • The formula is Inline graphic. Consider,

graphic file with name d33e1369.gif
graphic file with name d33e1379.gif
graphic file with name d33e1383.gif

Evaluating at Inline graphic:Inline graphic

  • Redefined Third Zagreb Index:

The formula is Inline graphic. Let.

graphic file with name d33e1413.gif

Therefore, we get

graphic file with name d33e1419.gif
graphic file with name d33e1422.gif

Evaluating at Inline graphic:

graphic file with name d33e1431.gif
  • Modified First Zagreb Index:

The formula is Inline graphic.

graphic file with name d33e1449.gif
graphic file with name d33e1452.gif

Evaluating at Inline graphic:

graphic file with name d33e1461.gif
  • Modified Second Zagreb Index:

The formula is Inline graphic. We have

graphic file with name d33e1480.gif
graphic file with name d33e1483.gif

Evaluating at Inline graphic:

graphic file with name d33e1492.gif
  • Hyper First Zagreb Index:

The formula is Inline graphic. Suppose that

graphic file with name d33e1510.gif
graphic file with name d33e1519.gif
graphic file with name d33e1523.gif

Evaluating at x = 1:

graphic file with name d33e1529.gif
  • Hyper Second Zagreb Index:

The formula is Inline graphic. We have

graphic file with name d33e1547.gif
graphic file with name d33e1550.gif

And consequently,

graphic file with name d33e1555.gif

Evaluating at Inline graphic:

graphic file with name d33e1564.gif
  • Nano Zagreb Index:

The formula is Inline graphic. We get

graphic file with name d33e1583.gif
graphic file with name d33e1586.gif

hence

graphic file with name d33e1591.gif

Evaluating at Inline graphic:

graphic file with name d33e1600.gif
  • Augmented Zagreb Index:

The formula is Inline graphic. Consider,

graphic file with name d33e1618.gif

Therefore, we get

graphic file with name d33e1624.gif
graphic file with name d33e1627.gif
graphic file with name d33e1630.gif
graphic file with name d33e1633.gif

Consequently,

graphic file with name d33e1638.gif

Evaluating at Inline graphic:

graphic file with name d33e1647.gif

The calculated values match the numerical results presented in the manuscript, confirming the accuracy of both the derived formulas and the computational algorithm.

Numerical results and regression models

We employ the provided MATLAB code to compute the Zagreb-type topological indices defined in Table 1 using Theorem 1. As an example, we calculate the topological indices of the ABT-089 graph using our MATLAB code. The resulting values are presented below.

Output of our MATLAB programming for the ABT-089 drug:

The order = 14.

The size = 15.

The maximum degree = 3.

The minimum degree = 1.

The first Zagreb index is 28*x^2*y^2 + 30*x^2*y^3 + 6*x^3*y^3 + 4*x*y^3.

Z1(G) = 68.

The second Zagreb index is 28*x^2*y^2 + 36*x^2*y^3 + 9*x^3*y^3 + 3*x*y^3.

Z2(G) = 76.

The third Zagreb index is 6*x^2*y^3 + 2*x*y^3.

Z3(G) = 8.

The redefined first Zagreb index is 7*x^2*y^2 + 5*x^2*y^3 + (2*x^3*y^3)/3 + (4*x*y^3)/3.

Rz1(G) = 14.

The redefined second Zagreb index is (31*x^4)/4 + (36*x^5)/5 + (3*x^6)/2.

Rz2(G) = 329/20.

The redefined third Zagreb index is 112*x^2*y^2 + 180*x^2*y^3 + 54*x^3*y^3 + 12*x*y^3.

Rz3(G) = 358.

The modified first Zagreb index is 2*x^4 + (6*x^5)/5 + x^6/6.

M1*(G) = 101/30.

The modified second Zagreb index is (7*x^2*y^2)/4 + x^2*y^3 + (x^3*y^3)/9 + (x*y^3)/3.

M2*(G) = 115/36.

The Hyper first Zagreb index is 128*x^4 + 150*x^5 + 36*x^6.

HM1(G) = 314.

The Hyper second Zagreb index is 112*x^2*y^2 + 216*x^2*y^3 + 81*x^3*y^3 + 9*x*y^3.

HM2(G) = 418.

The Nano Zagreb index is 30*x^2*y^3 + 8*x*y^3.

NZ(G) = 38.

The Augmented Zagreb index is Az(G)=:(475*x^2)/8 + 48*x^3 + (729*x^4)/64.

Az(G) = 7601/64.

M-polynomial is 7*x^2*y^2 + 6*x^2*y^3 + x^3*y^3 + x*y^3.

The numerical results for the topological indices of Alzheimer’s drugs are stated in Table 2. We employ various regression models to analyze the QSPR model, focusing on the topological indices computed from the molecular structures of Alzheimer’s drugs. The regression equations considered in this study are as follows.

The linear equation: Inline graphic

Quadratic equation: Inline graphic

Cubic equation: Inline graphic

Logarithmic equation: Inline graphic

Inverse equation: Inline graphic

Power equation: Inline graphic

S-Curve equation: Inline graphic

Exponential equation: Inline graphic

Compound equation: Inline graphic

Growth equation: Inline graphic

where Inline graphic and Inline graphic denote the topological index and the physicochemical property, respectively. The constant and the coefficient of the regression model are represented by Inline graphic, Inline graphic, Inline graphic, and Inline graphic53.

To evaluate the predictive performance of the regression models, three error metrics were employed: Mean Absolute Error (MAE), Mean Squared Error (MSE), and Root Mean Squared Error (RMSE). These are defined as

graphic file with name d33e1812.gif
graphic file with name d33e1815.gif
graphic file with name d33e1818.gif

where Inline graphic denotes the observed value, Inline graphic the predicted value, and Inline graphic the number of observations. MAE captures the average absolute deviation, MSE penalizes larger deviations more strongly, and RMSE expresses error in the same units as the predicted property.

In addition to error metrics, the coefficient of determination (Inline graphic) was used to assess model performance. Inline graphic is defined as.

graphic file with name d33e1847.gif

where Inline graphic​ is the mean of the observed values. Inline graphic quantifies the proportion of variance in the observed data explained by the model, with values closer to 1 indicating a better fit54.

Data analysis and discussion

This section evaluates the predictive power of the topological indices listed in Table 1 for a set of Alzheimer’s drugs. The goal is to build a QSPR model. To do this, we correlate the calculated topological indices (Table 2) with key experimental physicochemical properties (Table 3). These properties, including boiling point (BP), enthalpy of vaporization (E), flash point (FP), molar refractivity (MR), molar volume (MV), polarizability (P), and polar surface area (PSA) were selected because they are critical for determining a drug’s pharmacokinetic behavior. For instance, PSA is a crucial factor for predicting a drug’s ability to cross the blood-brain barrier, an essential requirement for any Alzheimer’s therapeutic. Similarly, MR and MV relate to the molecule’s size and binding potential. By establishing strong regression models between these properties and our indices, we aim to demonstrate that molecular structure, as captured by topological indices, can effectively predict a compound’s real-world behavior. The predictive performance of the regression models was assessed using Inline graphic and three error metrics: MAE, MSE, and RMSE. Lower values of MAE, MSE, and RMSE indicate better predictive accuracy, while higher Inline graphic values indicate a greater proportion of the observed variance that the model explains. The best-performing model is therefore identified as the one with the highest Inline graphic and lowest error metrics, with a significance level of less than Inline graphic.

Table 2.

The numerical values of Zagreb indices of Alzheimer’s drugs.

Drugs Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
ABT-089 Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
ABT-126 Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Donepezil Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Galantamine Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Memantine Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Metrifonate Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Phencyclidine Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Rivastigmine Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Tacrine Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic

Table 3.

The physicochemical features of alzheimer’s drugs.

Drugs Flash point
(FP)
Boiling point
(BP)
Polarizability
(P)
Enthalpy of
vaporization (E)
Polar surface
area (PSA)
Molar Refractivity
(MR)
Molar Volume
(MV)
ABT-089 Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
ABT-126 Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Donepezil Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Galantamine Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Memantine Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Metrifonate Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Phencyclidine Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Rivastigmine Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Tacrine Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic

Interpretation of key correlations

Table 4. The correlation coefficient calculated by linear regression models between Zagreb indices and the physicochemical properties of Alzheimer’s drugs∙.

Table 4.

Presents the correlation coefficients (R) between the topological descriptors and the physicochemical properties of the drugs, derived from linear regression equations. The results indicate that some indices, particularlyInline graphic, Inline graphic, and Inline graphic, show a very high correlation with most of the studied properties (except for PSA). In contrast, indices such as Inline graphic, Inline graphic, and Inline graphic do not show significant correlations. Notably, the PSA property shows no significant relationship with any of the topological indices in the linear model.

Topological index Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic

The detailed linear regression models that generated these high correlation values, including coefficients, R-squared values, and statistical significance, are presented in Tables 5 and 6.

Table 5.

The linear regression equations that give the best approximation for the physicochemical properties FP, BP, and P of alzheimer’s drugs.

Regression equations Inline graphic Std. Error MAE MSE RMSE P-value
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic

Table 6.

The linear regression equations that give the best approximation for the physicochemical properties E, MR, and MV of alzheimer’s drugs.

Regression equations Inline graphic Std. Error MAE MSE RMSE P-value
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic

A detailed examination of our regression models, presented in Tables 4 and 5, and 6, indicates that specific topological indices are particularly effective at predicting certain physicochemical properties. The analysis of correlation coefficients (R) in Table 4 reveals a striking pattern: Inline graphic demonstrates exceptionally strong correlations across a majority of the properties, including FP (Inline graphic), BP (Inline graphic), P (Inline graphic), and MR (Inline graphic).

The strength of these correlations is confirmed by the linear regression models detailed in Tables 5 and 6. For example, the linear regression model using Inline graphicdemonstrated superior performance in predicting MR. The model, defined by the equationInline graphic, achieved a coefficient of determination (Inline graphic) of Inline graphic, indicating it accounts for Inline graphic of the variance in MR. The model’s high accuracy is reinforced by low error metrics, including an MAE of 3.18, an MSE of 13.01, and an RMSE of Inline graphic. Similarly, the linear model for P, defined by the equationInline graphic, yielded a high coefficient of determination (Inline graphic) and minimal errors (Inline graphic, Inline graphic, Inline graphic). The superior performance of the Inline graphic index is mechanistically significant. As a descriptor defined by the sum of connectivity ratios, it is susceptible to branching in the molecular structure. It is a strong correlation with properties like Molar Volume (Inline graphic) and Enthalpy of vaporization (Inline graphic) suggests that intermolecular forces are heavily influenced by the molecule’s specific topology, not just its size. Furthermore, other indices show strong predictive power. As shown in Table 4, the modified Zagreb indices, Inline graphic, and Inline graphic, also exhibit excellent correlations, particularly with BP (Inline graphic for Inline graphic) and MR (Inline graphic for Inline graphic). The corresponding linear model for BP using Inline graphic (Inline graphic) is one of the most robust, with an Inline graphic of Inline graphic and low error metrics (Inline graphic, Inline graphic, Inline graphic) (Table 5). Also, the linear model for FP using Inline graphic (Inline graphic) is highly significant, as its Inline graphic value of Inline graphic indicates that this index alone can explain over Inline graphic of the variance in flash point supported by low errors (Inline graphic Inline graphic).

This underscores the importance of inverse-degree-based descriptors, which effectively capture the contributions of atoms with lower connectivity, often found in peripheral functional groups.

In contrast, no significant correlations were found for Inline graphic, Inline graphic, and Inline graphic indices. This suggests that descriptors based solely on the difference in vertex degrees, such as Inline graphic and Inline graphic, are less informative for predicting these bulk physicochemical properties.

Critically, PSA, a key parameter for drug absorption and blood-brain barrier penetration, shows no significant linear correlation with any of the studied indices (Table 4). This highlights a limitation of these 2D topological descriptors: while excellent at capturing properties related to overall molecular size and branching, they may be insufficient for predicting properties governed by specific polar-atom arrangements.

Moving beyond linear relationships, our analysis of non-linear models revealed even stronger and more nuanced predictive capabilities. The detailed results for all non-linear regression models, including quadratic, cubic, logarithmic, inverse, power, S-curve, exponential, compound, and growth equations, are provided in Appendix B (Tables B1-B9).

A key finding is the remarkable performance of Inline graphic, which consistently emerged as one of the most powerful predictors across multiple nonlinear models for properties such as MR and P.

For instance, using a power model, Inline graphic predicts MR with an Inline graphic of Inline graphic (Table B5), and using a cubic model, it predicts P with an Inline graphic of Inline graphic (Table B2). The widespread success of Inline graphicis particularly noteworthy. As a reformulated Zagreb-type index, it is susceptible to the distribution of edge degrees across the molecular graph, effectively capturing complex structural features, such as the presence of heteroatoms and branching patterns. Its dominant performance suggests that these specific structural characteristics are the primary drivers for properties that depend on electron distribution and molecular volume.

However, our analysis also identified specific cases where other indices provided superior predictions, highlighting the importance of selecting the right descriptor for the right property. For instance, in the cubic regression model, the Inline graphic index for predicting FP and BP, and the Inline graphic index for predicting E, were exceptionally accurate. For FP and BP, the cubic model based on Inline graphic yielded the highest coefficient of determination with Inline graphic and Inline graphic, respectively (Table B2). For E, the cubic model using the Inline graphic index was the most powerful predictor, with an Inline graphic value of Inline graphic (Table B2).

The success of different indices in predicting specific properties is chemically intuitive. Properties like FP and E are highly dependent on complex intermolecular forces that are better captured by the intricate formulations of the Inline graphic and Inline graphic indices, whereas properties like MR and P are more directly related to the overall connectivity and branching captured by Inline graphic

Furthermore, logarithmic models revealed unique relationships for certain properties. While not always the top-performing model, they offer insight into non-linear saturation effects. The Inline graphic index proved to be the best predictor for E within a logarithmic framework (Inline graphic, Table B3), while the Inline graphic index was superior for estimating MV (Inline graphic, Table B3). The strong performance of a logarithmic model for these properties suggests that, as molecular complexity increases, their values do not grow linearly but rather approach a plateau. This implies that beyond a certain molecular size or complexity, adding more atoms yields diminishing returns in increasing molar volume or enthalpy of vaporization, a phenomenon well captured by a logarithmic function.

Significance of model type and performance

A central finding of our comparative analysis, summarized in Table 7., is that the choice of regression model is as critical as the selection of the topological index itself. While linear models provide a helpful baseline, our results consistently show that non-linear models offer superior accuracy for predicting the physicochemical properties of the studied Alzheimer’s drugs. This is evidenced not only by higher Inline graphic values but also by correspondingly lower error metrics (MAE, MSE, and RMSE) across the non-linear models. To visually supplement these statistical findings, Fig. 2 presents the regression plots for the most significant correlations. These graphs are essential for comparing the goodness-of-fit of various model types. By illustrating the distribution of data points around the regression curves, the plots provide an intuitive understanding of each model’s predictive power. Specifically, they highlight instances where non-linear models (e.g., cubic, power) more accurately capture the underlying structure-property relationships than simple linear models, thereby visually reinforcing our conclusion that non-linear approaches offer superior accuracy for this dataset.

Table 7.

Choosing the best equation to estimate the physicochemical properties of alzheimer’s drugs with topological indices using max Inline graphic.

Property Indices Equations
Linear Quadratic Cubic Logarithmic Inverse Power S-Curve Exponential Compound Growth
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic

Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic
Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic Inline graphic

Fig. 2.

Fig. 2

Fig. 2

Fig. 2

Fig. 2

Graphical representation of regression models between topological indices and physicochemical properties of drugs.

The cubic model, while often more complex, demonstrated a unique and widespread advantage in our study. It emerged as the best predictor for FP and BP when using indices such as Inline graphic, Inline graphic, Inline graphic, Inline graphic, Inline graphic, Inline graphic, Inline graphic, Inline graphic, and Inline graphic. It was also the top-performing model for predicting the Enthalpy of E with all indices except Inline graphic and Inline graphic. This suggests that these specific properties are governed by highly complex, non-linear interactions that are only adequately captured by a higher-order polynomial.

The power model also demonstrated exceptional utility, emerging as the best predictor for properties such as MR and P across several topological indices, includingInline graphic, Inline graphic, Inline graphic, Inline graphic, and Inline graphic. This strong performance suggests that the relationship between these structural descriptors and specific properties is not merely additive but follows a multiplicative or power-law dependency. Such behavior is often characteristic of complex physicochemical phenomena, where properties scale non-linearly with molecular attributes, such as size and electron distribution.

Interestingly, for several properties, different model types yielded nearly identical high-accuracy predictions. For instance, for FP and BP prediction using the Inline graphic index, both quadratic and cubic models achieved the same highest Inline graphic value (Inline graphic). This redundancy implies that, for these specific structure-property relationships, adding a third-order (cubic) polynomial term does not capture significantly more variance than a simpler second-order (quadratic) model. In such cases, the quadratic model would be preferred for its greater parsimony. A similar pattern of shared high performance was observed among the exponential, compound, and growth models for predicting MV, with theInline graphic,Inline graphic,Inline graphic and Inline graphic indices, indicating that an exponential growth function robustly describes this particular relationship. Given their mathematical equivalence, these models inherently produced identical fits and R-squared values, confirming the suitability of a single exponential model for reporting.

In contrast to other models, the logarithmic model did not emerge as the top predictor for any property in this study. While such models can be insightful for phenomena involving saturation, our findings indicate that for this dataset, other non-linear functions provide a better fit.

In conclusion, our analysis extends beyond a simple ranking of indices, demonstrating that a nuanced, property-specific approach to model selection is essential. The findings provide a clear roadmap: for properties governed by complex interactions, such as FP, BP, and E, higher-order polynomials, like the cubic model, are often necessary. For properties governed by scaling laws such as MR and P, power models are superior. This understanding is critical for building truly predictive QSPR frameworks in the future.

A primary limitation of this study is the small sample size (Inline graphic) available for the regression analysis. This constrains the statistical robustness of our models and is a common challenge in QSPR studies of specific drug classes, such as those for Alzheimer’s disease, where the number of approved or late-stage compounds is inherently limited9,12,55. Consequently, our findings should be interpreted as exploratory and hypothesis-generating. Although the chemical homogeneity of our curated dataset may partially mitigate variance, we recommend that these models be validated with larger and more diverse datasets in future studies to confirm their predictive power and generalizability.

Conclusion

This study successfully developed robust QSPR models to predict the physicochemical properties of nine key Alzheimer’s drugs. The significant findings are summarized as follows:

  • Zagreb-type topological indices, efficiently calculated using the M-polynomial method, served as effective molecular descriptors for building the predictive models.

  • Nonlinear regression models, particularly cubic and power equations, demonstrated significantly greater predictive accuracy than standard linear models.

  • The selection of an optimal regression model proved to be property-specific, highlighting that the predictive power of a topological index is maximized only when paired with the appropriate mathematical model.

  • The cubic model was most effective for properties governed by complex intermolecular forces, such as BP and E. Indices like Inline graphic and Inline graphic yielded outstanding coefficients of determination (Inline graphic) up to 0.947 and 0.922, respectively.

  • The power model was the ideal choice for properties related to electron distribution and molecular volume, such as MR and P. Within this model, the Inline graphic index achieved an exceptional Inline graphic value of 0.963 for MR.

While our models show strong predictive capability, this work opens several promising avenues for future research. A primary limitation of this study is the small sample size (Inline graphic), which is a common challenge for specific drug classes. Therefore, a critical next step is to validate these models on larger, more diverse datasets to ensure their generalizability. Future work should also extend this QSPR framework to a QSAR analysis to correlate these indices with biological activities, incorporate 3D descriptors to capture spatial features more effectively, and apply machine learning algorithms to uncover more complex patterns for the design of next-generation Alzheimer’s therapeutics.

Supplementary Information

Below is the link to the electronic supplementary material.

Supplementary Material 1 (164.8KB, docx)

Author contributions

M.H.A. and M.H.S. conceptualized the study and supervised the research process. F.M. and M.Z. performed the computational analyses, including MATLAB coding and SPSS modeling. F.M. also contributed to data collection and graph analysis. All authors participated in the interpretation of results, contributed to the writing and critical revision of the manuscript, and approved the final version.

Data availability

Data, including all molecular structures and properties, is available at http://www.chemspider.com. The authors provide the MATLAB code upon request.

Declarations

Competing interests

The authors declare no competing interests.

Footnotes

Publisher’s note

Springer Nature remains neutral with regard to jurisdictional claims in published maps and institutional affiliations.

Contributor Information

Mohammad Hadi Akhbari, Email: hadi.akhbari@iau.ac.ir.

Mohammad Hassan Shahavi, Email: m.shahavi@ausmt.ac.ir.

References

  • 1.Liu, X., Yang, B., Liu, Q., Gao, M. & Luo, M. The long-term neuroprotective effect of MIND and mediterranean diet on patients with alzheimer’s disease. Sci. Rep.15, 32725. 10.1038/s41598-025-17055-5 (2025). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 2.Van Rossem, D. et al. Assessing the link between cerebellar volume and cognitive function in alzheimer’s disease: a pilot study. Sci. Rep.15, 32943. 10.1038/s41598-025-12975-8 (2025). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 3.Aghaei, A. & Moghaddam, M. E. An integrated predictive model for alzheimer’s disease progression from cognitively normal subjects using generated MRI and interpretable AI. Sci. Rep.15, 28340. 10.1038/s41598-025-13478-2 (2025). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 4.Ahamed, M. K. U. et al. A hybrid filtering and deep learning approach for early alzheimer’s disease identification. Sci. Rep.1510.1038/s41598-025-03472-z (2025). [DOI] [PMC free article] [PubMed]
  • 5.Liu, W. et al. Global burden of alzheimer’s disease and other dementias in adults aged 65 years and over, and health inequality related to SDI, 1990–2021: analysis of data from GBD 2021. BMC Public. Health. 25, 1256. 10.1186/s12889-025-22378-z (2025). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 6.Hassan, N., Miah, A. S. M., Suzuki, K., Okuyama, Y. & Shin, J. Stacked CNN-based multichannel attention networks for alzheimer disease detection. Sci. Rep.15, 5815. 10.1038/s41598-025-85703-x (2025). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 7.Liu, S. & Geng, D. A systematic analysis for disease burden, risk factors, and trend projection of alzheimer’s disease and other dementias in China and globally. PLoS One. 20, e0322574. 10.1371/journal.pone.0322574 (2025). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 8.Anitha, K. et al. Recent insights into the neurobiology of alzheimer’s disease and advanced treatment strategies. Mol. Neurobiol.62, 2314–2332. 10.1007/s12035-024-04384-1 (2025). [DOI] [PubMed] [Google Scholar]
  • 9.Ashraf, T., Idrees, N. & Belay, M. B. Regression analysis of topological indices for predicting efficacy of alzheimer’s drugs. PLoS One. 19, e0309477. 10.1371/journal.pone.0309477 (2024). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 10.Ravi, V. & Chidambaram, N. QSPR analysis of physico-chemical and Pharmacological properties of medications for parkinson’s treatment utilizing neighborhood degree-based topological descriptors. Sci. Rep.15, 16941. 10.1038/s41598-025-00898-3 (2025). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 11.Movahedi, F., Zameni, M., Akhbari, M. H., Shahavi, M. H. & Rajabnezhad, S. Topological indices and data analysis techniques modeling to predict the physicochemical properties of Tetracycline antibiotics. J. Pharm. Sci.114, 103871. 10.1016/j.xphs.2025.103871 (2025). [DOI] [PubMed] [Google Scholar]
  • 12.Sardar, M. S. & Hakami, K. H. QSPR Analysis of Some Alzheimer’s Compounds via Topological Indices and Regression Models. Journal of Chemistry, 2024, 5520607, 10.1155/2024/5520607 (2024).
  • 13.Ahmed, W. E., Hanif, M. F., Alzahrani, E. & Fiidow, O. A. Predicting bone cancer drugs properties through topological indices and machine learning. Sci. Rep.15, 31150. 10.1038/s41598-025-16497-1 (2025). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 14.Arockiaraj, M., Greeni, A. B., Kalaam, A. R. A., Aziz, T. & Alharbi, M. Mathematical modeling for prediction of physicochemical characteristics of cardiovascular drugs via modified reverse degree topological indices. Eur. Phys. J. E. 47, 53. 10.1140/epje/s10189-024-00446-3 (2024). [DOI] [PubMed] [Google Scholar]
  • 15.Hasani, M. & Ghods, M. Predicting the physicochemical properties of drugs for the treatment of parkinson’s disease using topological indices and MATLAB programming. Mol. Phys.122, e2270082. 10.1080/00268976.2023.2270082 (2024). [Google Scholar]
  • 16.Shahavi, M. H., Jahanshahi, M., Najafpour, G., Ebrahimpour, M. & Hosenian, A. Expanded bed adsorption of biomolecules by NBG contactor: experimental and mathematical investigation. World Appl. Sci. J.13, 181–187 (2011). https://www.idosi.org/wasj/wasj13(2)2011.htm [Google Scholar]
  • 17.Mahboob, A., Rasheed, M. W., Hanif, I., Amin, L. & Alameri, A. Role of molecular descriptors in quantitative structure-property relationship analysis of kidney cancer therapeutics. Int. J. Quantum Chem.124, e27241. 10.1002/qua.27241 (2024). [Google Scholar]
  • 18.P, U. P., Suresh, M., Tolasa, F. T. & Bonyah, E. QSPR/QSAR study of antiviral drugs modeled as multigraphs by using ti’s and MLR method to treat COVID-19 disease. Sci. Rep.14, 13150. 10.1038/s41598-024-63007-w (2024). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 19.Movahedi, F. Matching polynomials for some Nanostar dendrimers. Asian-European J. Math.14, 2150188. 10.1142/S1793557121501886 (2021). [Google Scholar]
  • 20.Tawhari, Q. M., Naeem, M., Maqbool, S., Rauf, A. & Oyelakin, O. Mathematical modeling and statistical analysis of breast cancer drugs using M-polynomial indices for the physical properties. Sci. Rep.15, 30365. 10.1038/s41598-025-07067-6 (2025). [DOI] [PMC free article] [PubMed] [Google Scholar] [Retracted]
  • 21.Xavier, D. A., Julietraja, K., Alsinai, A. & Akhila, S. Prediction of properties of Boron alpha-icosahedral nanosheet by bond-addictive M-polynomial. Sci. Rep.14, 1197. 10.1038/s41598-024-51642-2 (2024). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 22.Samiei, Z. & Movahedi, F. Investigating graph invariants for predicting properties of chemical structures of antiviral drugs. Polycycl. Aromat. Compd.44, 6696–6713. 10.1080/10406638.2023.2283625 (2024). [Google Scholar]
  • 23.Movahedi, F. & Akhbari, M. H. Degree-based topological indices of the molecular structure of hyaluronic acid–methotrexate conjugates in cancer treatment. Int. J. Quantum Chem.123, e27106. 10.1002/qua.27106 (2023). [Google Scholar]
  • 24.Abbasi, N., Shourian, M. & Shahavi, M. H. Synthesis and characterization of Q0-Eugenol nanoemulsion for drug delivery to breast and hepatocellular cancer cell lines. ACS Omega. 10, 25299–25312. 10.1021/acsomega.4c11261 (2025). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 25.Shi, X., Kosari, S., Ghods, M. & Kheirkhahan, N. Innovative approaches in QSPR modelling using topological indices for the development of cancer treatments. PLoS One. 20, e0317507. 10.1371/journal.pone.0317507 (2025). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 26.Zaman, S., Majeed, M. U., Ahmed, W. & Saleem, M. T. On neighborhood Eccentricity-Based topological indices with QSPR analysis of PAHs drugs. Measurement: Interdisciplinary Res. Perspect.23, 199–212. 10.1080/15366367.2024.2329950 (2025). [Google Scholar]
  • 27.Mehta, K. et al. Modernizing preclinical drug development: the role of new approach methodologies. ACS Pharmacol. Translational Sci.8, 1513–1525. 10.1021/acsptsci.5c00162 (2025). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 28.Ahmed, W. et al. Harnessing topological descriptors: A comparative analysis of artificial neural networks and random forest for predicting Anti-Alzheimer drug properties. Nano Online. 255008510.1142/s1793292025500857 (2025).
  • 29.Öztürk Sözen, E., Eryaşar, E. & Çakmak, Ş. Szeged-like topological descriptors and COM-polynomials for graphs of some alzheimer’s agents. Mol. Phys.122, e2305853. 10.1080/00268976.2024.2305853 (2024). [Google Scholar]
  • 30.Ahmed, W., Ali, K., Zaman, S. & Raza, A. Molecular insights into anti-Alzheimer’s drugs through predictive modeling using linear regression and QSPR analysis. 38, 2450260, (2024). 10.1142/s0217984924502609
  • 31.Kumar, V. & Das, S. A novel approach to determine the Sombor-type indices via M-polynomial. J. Appl. Math. Comput.71, 983–1007. 10.1007/s12190-024-02272-4 (2025). [Google Scholar]
  • 32.Zaman, S., Rasheed, S. & Alamer, A. A quadratic regression model to quantify certain latest Corona treatment drug molecules based on coindices of M-polynomial. J. Supercomputing. 80, 26805–26830. 10.1007/s11227-024-06434-w (2024). [Google Scholar]
  • 33.Rauf, A., Naeem, M., Ramzan, R. & Cham, A. Exploring physicochemical characteristics of cyclodextrin through M-polynomial indices. Sci. Rep.14, 20029. 10.1038/s41598-024-68775-z (2024). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 34.Usha, A., Shanmukha, M. C., Shilpa, K. C. & Praveen, B. M. Comparative study of degree-based molecular descriptors of cyclodextrins through M-polynomial and NM-polynomial. J. Indian Chem. Soc.100, 100999. 10.1016/j.jics.2023.100999 (2023). [Google Scholar]
  • 35.Yavari, K., Karami, C., Bijari, S., Adami, D. & Shahavi, M. H. Removal of amoxicillin from hospital waste using Fe2O3-Ag adsorbent and optimization by response surface methodology and machine learning prediction. Curr. Res. Green. Sustainable Chem.10048310.1016/j.crgsc.2025.100483 (2025).
  • 36.Gorji, N., Jahanshahi, M., Shahavi, M. H. & Ayrilmis, N. Ethylcellulose microparticles as green encapsulation for slow release of Microspherical abamectin pesticide for agricultural applications: improvement of process parameters. Int. J. Biol. Macromol.321, 146336. 10.1016/j.ijbiomac.2025.146336 (2025). [DOI] [PubMed] [Google Scholar]
  • 37.Gholami, M., Hosseini, M., Shahavi, M. H. & Jahanshahi, M. Process optimization of corn starch nanoparticles containing Linalyl acetate: characterization and antibacterial properties. Int. J. Industrial Chem.14, 142301–142311. 10.57647/j.ijic.2023.1402.04 (2023). [Google Scholar]
  • 38.Shahavi, M. H., Hosseini, M., Jahanshahi, M., Meyer, R. L. & Darzi, G. N. Evaluation of critical parameters for Preparation of stable clove oil nanoemulsion. Arab. J. Chem.12, 3225–3230. 10.1016/j.arabjc.2015.08.024 (2019). [Google Scholar]
  • 39.Mofidian, R., Barati, A., Jahanshahi, M. & Shahavi, M. H. Optimization on thermal treatment synthesis of lactoferrin nanoparticles via Taguchi design method. SN Appl. Sci.1, 1339. 10.1007/s42452-019-1353-z (2019). [Google Scholar]
  • 40.Shahavi, M. H., Hosseini, M., Jahanshahi, M., Meyer, R. L. & Darzi, G. N. Clove oil nanoemulsion as an effective antibacterial agent: Taguchi optimization method. Desalination Water Treat.57, 18379–18390. 10.1080/19443994.2015.1092893 (2016). [Google Scholar]
  • 41.Koam, A. N. A. et al. Hosoya entropy analysis of some fullerene structures. Discover Nano. 20, 100. 10.1186/s11671-025-04255-1 (2025). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 42.Youssefi, M. R.et al. Dietary Supplementation with Eugenol Nanoemulsion Alleviates the Negative Effects of Experimental Coccidiosis on Broiler Chicken’s Health and Growth Performance.Molecules 28, 2200, doi:https://doi.org/10.3390/molecules28052200 (2023). [DOI] [PMC free article] [PubMed]
  • 43.Xavier, D. A. et al. Comparative study of molecular descriptors of Pent-Heptagonal nanostructures using neighborhood M-Polynomial approach. Molecules28, 2518. 10.3390/molecules28062518 (2023). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 44.Abdallah, A. E. Review on anti-alzheimer drug development: approaches, challenges and perspectives. RSC Adv.14, 11057–11088. 10.1039/d3ra08333k (2024). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 45.ChemSpider, Royal Society of Chemistry, (2025). https://www.chemspider.com/
  • 46.Ali, P., Kirmani, S. A. K., Rugaie, A., Azam, F. & O. & Degree-based topological indices and polynomials of hyaluronic acid-curcumin conjugates. Saudi Pharm. J.28, 1093–1100. 10.1016/j.jsps.2020.07.010 (2020). [DOI] [PMC free article] [PubMed] [Google Scholar]
  • 47.Gutman, I. & Trinajstić, N. Graph theory and molecular orbitals. Total φ-electron energy of alternant hydrocarbons. Chem. Phys. Lett.17, 535–538. 10.1016/0009-2614(72)85099-1 (1972). [Google Scholar]
  • 48.Raza, Z. The harmonic and second Zagreb indices in random polyphenyl and Spiro chains. Polycycl. Aromat. Compd.42, 671–680. 10.1080/10406638.2020.1749089 (2022). [Google Scholar]
  • 49.Ayache, A., Alameri, A., Alsharafi, M. & Ahmed, H. The Second Hyper-Zagreb Coindex of Chemical Graphs and Some Applications. J. Chem.2021, 3687533. 10.1155/2021/3687533 (2021). [Google Scholar]
  • 50.Islam, S. R., Mohsin, B. B. & Pal, M. Hyper-Zagreb index in fuzzy environment and its application. Heliyon1010.1016/j.heliyon.2024.e36110 (2024). [DOI] [PMC free article] [PubMed]
  • 51.Hayat, S. & Asmat, F. Sharp bounds on the generalized multiplicative first Zagreb index of graphs with application to QSPR modeling. Mathematics11, 2245. 10.3390/math11102245 (2023). [Google Scholar]
  • 52.Mondal, S. & Das, K. C. Complete solution to open problems on exponential augmented Zagreb index of chemical trees. Appl. Math. Comput.482, 128983. 10.1016/j.amc.2024.128983 (2024). [Google Scholar]
  • 53.Roy, K. Advances in QSAR Modeling 1 edn (Springer Cham, 2017).
  • 54.Brook, R. J. & Arnold, G. C. Applied Regression Analysis and Experimental Design (CRC, 2018).
  • 55.HH, M. Y., Suresh, M. & Bayati, J. H. H. Topological indices and QSPR/QSAR analysis of some drugs being investigated for the treatment of alzheimer’s disease patients. Baghdad Sci. J.22, 242–272. 10.21123/bsj.2024.10866 (2025). [Google Scholar]

Associated Data

This section collects any data citations, data availability statements, or supplementary materials included in this article.

Supplementary Materials

Supplementary Material 1 (164.8KB, docx)

Data Availability Statement

Data, including all molecular structures and properties, is available at http://www.chemspider.com. The authors provide the MATLAB code upon request.


Articles from Scientific Reports are provided here courtesy of Nature Publishing Group

RESOURCES