Abstract
Industrial effluents of textile, paper, and leather industries contain various toxic dyes as one of the waste material. It imparts major impact on human health as well as environment. The white rot fungus Pycnoporus cinnabarinus Laccase is generally used to degrade these toxic dyes. In order to decipher the mechanism of process by which Laccase degrade dyes, it is essential to know its 3D structure. Homology modeling was performed in presented work, by satisfying Spatial restrains using Modeller Program, which is considered as standard in this field, to generate 3D structure of Laccase in unison, SWISSMODEL web server was also utilized to generate and verify the alternative models. We observed that models created using Modeller stands better on structure evaluation tests. This study can further be used in molecular docking techniques, to understand the interaction of enzyme with its mediators like 2, 2‐azinobis (3‐ethylbenzthiazoline‐6‐sulfonate) (ABTS) and Vanillin that are known to enhance the Laccase activity.
Keywords: Homology modeling, Spatial restraints, Modeller, Laccase, QMEAN, Ml methods, Beale restart conjugate gradients method, Leap-frog verletintegrator
Background
In developing countries including India textile, leather and paper industries represent an important economic sector. Huge amount of capital and Human resource is engaged in these industries. These industries are one of the most important sources of Environmental pollution. Mills of these industries emancipate enormous amount of waste matter each year, that contain variety of chemicals such as formaldehyde, chlorine, heavy metals (such as lead and mercury) and toxic dyes, which lay noteworthy foundation of environmental degradation and human illnesses. Most -the dyes that are released from these industries are polymers possessing very complex structure and are very difficult to decompose biologically [1]. Many reactive dyes are not degraded in ordinary aerobic sewage treatment processes and that they can be discharged unaffected from the treatment plant [2]. An even very minute concentration of dyes in effluent is visible and is often carcinogenic [3]. Laccase belongs to group of enzyme named as large blue copper proteins or blue copper oxidases possessing polyphenol oxidase activity. It functions by generally reducing oxygen to water simultaneously oxidizing a polyphenolic substrate .Laccase has evolved with a remarkable property of non‐specificity of its reducing substrates and encompass vast range of substrates oxidized [4], making it a marvelous contrivance to oxidize toxic dyes which are generally polyphenols[5]. Knowledge of 3D structure of Laccase can aid us to uncover the mystery of how Laccase has attained such huge functional diversity. To experimentally discover functionality of any protein, the information of its 3D structure remains an indispensable fact, which is achieved using techniques like X‐Ray Crystallography or NMR spectroscopy. Experimental techniques are very tedious and prolonged and not always succeed in determining structure for all proteins especially membrane proteins [6]. Moreover, the rate at which protein sequence data is accumulating is far more than the structural information available, thus creating a gap between available sequences and experimentally solved structures. Computational methods like homology modeling can help reduce this gap. It is known that existing proteins are result of continuous evolution of previously existing ones, thus proteins can be grouped into families. Members in same family are similar and thus have similar folds; this fact allows predicting the structure of other members of family if structure of single member is known and the technique by which this task is achieved is termed as Homology Modeling. Modeller [7] is stand‐alonepackage for homology modeling that accomplishes the job by method called ‘Satisfaction of spatial restraints’ using a set of restraints derived from the alignment and expressed as Probability Density Functions, finally the model is obtained by minimizing the violations to these restraints. Studies have proved that Modeller outperforms most of other homology modeling suits, it's fast, reliable and freely available and hence we selected it in current study [8].
Methodology
Sequence Retrieval and Template selection
The sequence of laccase enzyme of “Pycnoporus cinnabarinus” was retrieved from SWISS‐PROT [9] database with accession number O59896. The complete sequence length of Laccase was reported as 518 amino acid residues which is synthesized as inactive precursor having first 21 residues as signaling peptide, the remaining 22 to 518 residues constitute the functional domains of the enzyme [10]. Hence while searching template,first 21 amino acid residues were removed. A template selection search was performed using BLAST‐P [11] against PDB [12] database from NCBI interface simultaneously ”Template Identification Tool” at Swiss‐ Model interface [13] provided by Swiss Institute of Bioinformatics was utilized for template selection. In results, 12 significant hits with E‐value zero were observed, the best of these comprise of model 3FPX‐A and 3DIV‐A from species Trametes hirsute and Cerrena maxima sharing 84 % and 82% sequence identity with query sequence. The remaining three, of which two models 2QT6‐B and 2VDZ‐A are Laccase from Lentinus tigrinus and Coriolopsis gallica respectively showing 77 % sequence identity with query sequence, while template 1GYC‐A from species Tremetes versicolor seems to have 78 % sequence identical to that of laccase from Pycnoporus cinnabarinus. Similar results were obtained from NCBI BLAST‐P server.
Sequence Analysis
Elementary domain analysis carried out on “InterPro Domain Scan” [14] revealed presence of one Multicopper oxidase domains of type I, II and III each from residue number 163 ‐ 305, 365 ‐ 492, 30 ‐ 152 respectively, in addition one copper binding site is predicted around residue number 125 to 145. Phylogenetic analysis was performed using traditional approach i.e. Multiple sequence alignment (MSA) was constructed using Clustal-Xversion 1.81 [15] followed by tree reconstruction using Phylip Package [16]. 1000 datasets of query sequence were generated by resampling the original MSA using Seqboot followed by estimation of unrooted phylogeny for each dataset by using Protpars. Subsequently using family of consensus tree methods called the Ml methods implemented in Consense program, a consensus tree was drawn following majority rule consensus. Laccase from species Coriolus zonatus was used as an outgroup in tree reconstruction process. Finally, a rooted tree without using branch length was generated (see Figure 1) using Drawgram program. Figure 1 clearly illustrates that Templates 3FPX and 3DIV are closest ancestor of query sequence and thus were used as templates for further modeling process. One interesting fact to note is that template 1GYC, which has higher sequence identity (78 %) to query, is distantly related as compared to templates 2VDZ and 2QT6 which shares less sequence identity (77 %).
Figure 1.

Evolutionary relationship of query sequence with templates
Modeling of the structure
The Align2D command from Modeller was utilized to align query sequence to template structures, followed by generation of models by Loop‐Model class. Five models were allowed to be build from each template and single loopmodel was generated after refinement process. Thus five models from each template were generated. At the same time SWISS‐MODEL server was also used to generate models using same template. Energy minimization was carried out by first implementing the Conjugate Gradient and subsequently Molecular Dynamics(MD) approach; Conjugate Gradient was allowed to perform 200 optimizing iterations using a modified version of the Beale restart conjugate gradients method, while Molecular Dynamics optimization was performed at 300 Kelvin for 500 MD steps using the leap‐frog Verlet integrator.
Model Evaluation
Models so produced were ranked on QMEAN server [17]. All the models created by means of Modeller using template 3FPX had Qmean score above 0.83 indicating good quality of each model produced, while only one model (Model 2.pdb) generated using template 3DIV had Qmean score above 0.83 (See Table 1 in supplementary material); also Models generated by using Swiss model Automated homology modeling server had score 0.85 and 0.84 for template 3FPX and 3DIV respectively. Models deduced by template 3DIV showed poor QMEAN score despite of the method used (Modeller or SwissModel server). Models were evaluated on basis of geometrical and stereochemical constraints using ProCheck [18] and factors like unfavorable atomic contacts, side chain planarity problems, connections to aromatic rings out of plane etc were assessed using WhatCheck (WhatIf) [19] utilities available at SIB server. Finally the models were visualized using Chimera software [20].
Discussion
It is evident from Table 2 (see supplementary material) and Figure 3a that best model created using template 3FPX employing Modeller program (Model1.pdb) has more than 91 % of its Non Proline and Non Glycine amino acid residues in core region, 7.7 % in allowed region, 0.5 % in generously allowed regions and only 0.5 % in disallowed region as compared to models created using SWISSMODEL (Model 3FPX swissmodel.pdb) have only 82 % of its Non Proline and Non Glycine amino acid residues in core region, 16.4 % in allowed region, 0.5 % in generously allowed regions and none of its amino acid residue lie in disallowed region, in other words approximately only 8 % of residues created using template 3FPX employing Modeller lie outside Core region as compared to 17.2 % of residues created using template 3FPX employing SWISSMODEL remain outside Core region indicating that models created using Modeller are better in terms of geometrical and stereo chemical properties; Similarly in case of best model created using template 3DIV by means of Modeller(Model1.pdb) has around 89 % of its Non Proline and Non Glycine amino acid residues in core region, 9.6 % in allowed region, 0.7 % in generously allowed regions and none of its amino acid residue lie in disallowed region as compared to models created using SWISSMODEL(Model 3DIV swissmodel.pdb) that have 83% of its Non Proline and Non Glycine amino acid residues in core region, 15.4% in allowed region, 0.5 % in generously allowed regions and 0.7% in disallowed, in other words approximately only 10.3 % of residues created using template 3FPX employing Modeller lie outside Core region as compared to 16.8 % of residues created using template 3FPX employing SWISSMODEL region again demonstrating that models created using Modeller are better in terms of geometrical and stereo chemical properties (Table 3 see supplementary material and Figure 3b).
Figure 3.

A Procheck results for best model created employing Modeller using 3FPX as template B Procheck results for best model created utilizing Modeller using 3DIV as template
The WhatIf report reveals 5 errors regarding Side chain planarity problems in models created using 3FPX as template utilizing SWISSMODEL in contrast only two out of five models created using Modeller showed that error containing single residue possessing side chain planarity problem; none of model possessed error of Connections to aromatic rings out of plane in case of MODELLER while 6 same errors were identified in model created at SWISSMODEL server.
In case of models created using 3DIV as template utilizing Modeller reported single side chain planarity problem error in only one model out of five, in contrast 18 side chain planarity problem errors were observed in model created using SWISSMODEL. None of aromatic rings had connection out of plane in any model created using MODELLER while 8 residues had aromatic rings that had connection out of plane in model generated at SWISSMODEL server.
Domain Analysis
As predicted in sequence, all the three domains are detected in structure. The type III copper binding domain containing Plastocyanin fold is shown in Ribbon representation in Figure 4a with copper binding site in Sticks representation among ASN 130 and PRO 132, the predicted copper binding residues in Magenta. Type I and Type II domains are depicted in Figure 4c and Figure 4d respectively, while Figure 4b represents the entire tertiary structure of Laccase with all the three domains.
Figure 4.

A Type III copper binding Domain (Sky Blue Ribbons) with side chains of residues of metal binding site (Red sticks); B All three domains with green coloured (Type III), orange coloured (Type I), dark magenta coloured (Type II); C Type I copper binding domain with β‐sheets in Magenta, Helices in Orange and Loops in Grey; D Type II copper binding domain with rainbow colouring scheme.
Conclusion
Finally to conclude, Model 1 constructed using template 3FPX employing MODELLER performs better in ProCheck and WhatIf structure validation test as compared to models made at SWISSMODEL server. The same model also outperforms the models created using 3DIV as template. Furthermore we can utilize the resulting models of this work in Molecular Docking studies to gain more insight of its interaction with its mediators like Vanillin and ABTS that are known to enhance its activity
Supplementary material
Figure 2.

Alignment of query sequence with 3FPX template
Footnotes
Citation:Meshram etal; Bioinformation 5(4): 150-154 (2010)
References
- 1.Neppolian B, et al. J. Environ. Sci. health. 1999;34:1829. [Google Scholar]
- 2.Carliell CM, et al. Water. 1996;22:225. [Google Scholar]
- 3.Kim S, et al. J. Biosci. Bioeng. 2003;95:102. doi: 10.1016/S1389-1723(03)80156-1. [DOI] [PubMed] [Google Scholar]
- 4.Wood DA. J. Gen. Microbiol. 1980;127:327. [Google Scholar]
- 5.Kersten PJ, et al. Biochem J. 1990;268:475. doi: 10.1042/bj2680475. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 6.Johnson MS, et al. Crit Rev Biochem Mol Biol. 1994;29:1. doi: 10.3109/10409239409086797. [DOI] [PubMed] [Google Scholar]
- 7.Sali A, et al. J. Mol. Biol. 1993;234:779. doi: 10.1006/jmbi.1993.1626. [DOI] [PubMed] [Google Scholar]
- 8.Wallner B, et al. Protein Sci. 2005;14:1315. doi: 10.1110/ps.041253405. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 9.Boeckmann B, et al. Nucleic Acids Res. 2003;31:365. doi: 10.1093/nar/gkg095. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 10.Eggert C, et al. Appl. Environ. Microbiol. 1996;62:1151. doi: 10.1128/aem.62.4.1151-1158.1996. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 11.Stephen F, et al. Nucleic Acids Res. 1997;25:3389. doi: 10.1093/nar/25.17.3389. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 12.Westbrook J, et al. Nucleic Acids Res. 2003;31:489. doi: 10.1093/nar/gkg068. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 13.Schwede T, et al. Nucleic Acids Res. 2003;31:3381. doi: 10.1093/nar/gkg520. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 14.Zdobnov EM, et al. Bioinformatics. 2001;17:847. doi: 10.1093/bioinformatics/17.9.847. [DOI] [PubMed] [Google Scholar]
- 15.Jeanmougin F, et al. Trends Biochem Sci. 1998;23:403. doi: 10.1016/s0968-0004(98)01285-7. [DOI] [PubMed] [Google Scholar]
- 16.Felsenstein J, et al. Cladistics. 1989;5:164. [Google Scholar]
- 17.Benkert P, et al. Nucleic Acids Res. 2009;1:37. doi: 10.1093/nar/gkp322. [DOI] [PMC free article] [PubMed] [Google Scholar]
- 18.Laskowski RA, et al. J. Appl. Cryst. 1993;26:283. [Google Scholar]
- 19.Hooft RW, et al. Nature. 1996;381:272. doi: 10.1038/381272a0. [DOI] [PubMed] [Google Scholar]
- 20.Pettersen EF, et al. J. Comput. Chem. 2004;25:1605. doi: 10.1002/jcc.20084. [DOI] [PubMed] [Google Scholar]
Associated Data
This section collects any data citations, data availability statements, or supplementary materials included in this article.
