{"id":19054,"date":"2018-03-25T11:38:00","date_gmt":"2018-03-25T11:38:00","guid":{"rendered":"http:\/\/biomedpharmajournal.org\/?p=19054"},"modified":"2020-04-23T04:42:58","modified_gmt":"2020-04-23T04:42:58","slug":"multilabel-classification-of-membrane-protein-in-human-by-decision-tree-dt-approach","status":"publish","type":"post","link":"https:\/\/biomedpharmajournal.org\/staging\/vol11no1\/multilabel-classification-of-membrane-protein-in-human-by-decision-tree-dt-approach\/","title":{"rendered":"Multilabel Classification of Membrane Protein in Human by Decision Tree (DT) Approach"},"content":{"rendered":"<p><strong>Int<\/strong><strong>r<\/strong><strong>oductio<\/strong><strong>n<\/strong><\/p>\n<p>Multilabel classification methods are progres- sively used in recent research works, protein function, protein type, semantic scene classification and music categorization. A general form of multi class classification is Multi-label classification. It is single-label problem of grouping instances into one of more than two classes. The main feature of multi-label problem is that the instance can be assigned to any number of classes. We proposes a multi label classification of different types of membrane proteins by implementing DT classifier algorithm. Membrane proteins play different roles in cellular biology. About 30% of human genomes have been encoded from membrane proteins. Information of a given membrane protein type helps to determine its function. Membrane proteins are refereed as membrane associated proteins or membrane-bound proteins. They are classified on the basis of their interaction modes with membranes, and cellular lo-cations. Membrane proteins play important role in-volved in various cellular processes.<sup>1<\/sup> The number of membrane proteins in humans is to 8000 as per the estimation of Gao et al.<sup>2<\/sup> According to Krogh et. al<sup>3<\/sup> 20-30% of genes are involved in encoding membrane proteins. The role of membrane proteins the discovery of new drugs as well as in the analyses of the mechanism of cellular activities is worth mentioning.<sup>4,5,6<\/sup> All membrane protein func-tions are usually related with its type.<sup>7<\/sup> Application of traditional biophysical methods<sup>8<\/sup> are time con- suming and costly while determining the types of uncharacterized membrane proteins. On the basis of the interactions between membrane proteins and membrane, H. Lodish et. al<sup>9<\/sup> membrane protein are divided in to two types intrinsic and extrinsic mem-brane proteins. (Fig.1)<\/p>\n<p>Transmembrane protein(Integral membrane pro-teins) are permanently bound to the biological mem-brane. Peripheral membrane proteins are temporar-ily attached to a membrane or integral membrane proteins. Integral membrane proteins are classified as Transmembrane proteins and Anchored mem- brane proteins. Transmembrane proteins are type putational method is an urgent need for the protein functional type prediction. This paper proposed a multi-label classification of membrane proteins in humans, using DT classifier algorithm. For that datasets S1 is constructed from UniProt database. It is reported from the performance of this method that it could be quite effective to classify membrane protein types.<\/p>\n<table style=\"width: 70%;\" border=\"1\" cellpadding=\"5\">\n<tbody>\n<tr>\n<td>\u00a0<img decoding=\"async\" class=\"alignnone size-thumbnail wp-image-19078\" src=\"https:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig12-150x150.jpg\" alt=\"Figure 1: Two types of Membrane Proteins structure\" width=\"150\" height=\"150\" srcset=\"https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig12-150x150.jpg 150w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig12-256x256.jpg 256w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig12.jpg 782w\" sizes=\"(max-width: 150px) 100vw, 150px\" \/><\/td>\n<td><strong>Figure 1: Two types of Membrane Proteins structure<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p><a href=\"http:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig12.jpg\" target=\"_blank\">Click here to View figure<\/a><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p>I,type II, and Multi-pass, whereas Anchored mem-brane proteins are Lipid and GPI. Based on the positions and intramolecular arrangements in a cell, membrane proteins are classified into six types,<sup>10<\/sup> shown in Fig.2.<\/p>\n<table style=\"width: 70%;\" border=\"1\" cellpadding=\"5\">\n<tbody>\n<tr>\n<td><img decoding=\"async\" class=\"alignnone size-thumbnail wp-image-19079\" src=\"https:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig22-150x150.jpg\" alt=\"Figure 2: Schematic illustration to show the six types of membrane proteins: (a) type I, (b) type II, (c) multipass transmembrane (MPT), (d) lipid chain-anchored membrane (LCM), (e) GPI-anchored membrane, and (f) peripheral membrane(PM). [11].\" width=\"150\" height=\"150\" srcset=\"https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig22-150x150.jpg 150w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig22-256x256.jpg 256w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig22.jpg 758w\" sizes=\"(max-width: 150px) 100vw, 150px\" \/><\/td>\n<td><strong>Figure 2: Schematic illustration to show the six types of membrane proteins: (a) type I, (b) type II, (c) multipass transmembrane (MPT), (d) lipid chain-anchored membrane (LCM), (e) GPI-anchored membrane, and (f) peripheral membrane(PM). [11].<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p><a href=\"http:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig22.jpg\" target=\"_blank\">Click here to View figure<\/a><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p>In MPT, the polypeptide crosses the lipid bilayer multiple times, i.e, spanning the membrane more than once. LCM are covalently linked to a lipid molecule and serve to anchor them to either the cytoplasmic or extracellular surface of a biological membrane. GPI-anchored membrane protein is also called \u00a0membrane-anchored\u00a0 proteins. It \u00a0is bound to the membrane by a glycosylphosphatidylinositol (GPI) anchor.<\/p>\n<p>Membrane proteins are a common type of pro-teins along with soluble globular proteins, fibrous proteins, and disordered\u00a0 proteins. They are tar- gets of over 50% of all modern medicinal drugs.<sup>12<\/sup> It is estimated that 20-30% of all genes in most genomes encode membrane proteins.<sup>3\u00a0<\/sup>Thus classification of membrane proteins into six types is a resource intensive and time consuming task. Therefore, developing a reliable and effective com-<\/p>\n<p><strong>Rel<\/strong><strong>a<\/strong><strong>te<\/strong><strong>d Work<\/strong><\/p>\n<p>The computational methods used for the clas- sification of membrane proteins include analyti-cal methods, mathematical modelling and simula- tion. The bioinformatics application generally use strategical analysis methods, like machine learning methods for the classification and prediction of membrane proteins.<\/p>\n<p>For \u00a0the \u00a0successful\u00a0 implementation of \u00a0machine learning techniques are equally important. both fea- ture extraction and learning algorithms are equally required. Feature predictions commonly used are: amino acid composition (AAC), position-specific scoring matrices (PSSMs), pseudo amino acid com- position (PseAAC), physicochemical properties of amino acids and functional domains. AAC was simplest and most efficient represention of protein sequence. Membrane proteins are classified accord- ing to two different schemes by Kuo et al<sup>13\u00a0<\/sup>which are based on protein types and its location. Their dataset was constructed from the SWISS PROT (release 35) database. The rate of correct prediction of membrane proteins type and cellular location revealed that 76-81% and 66-70 % by using the self consistency ,jackknife tests, as well as by an independent dataset test. This method was improved by using N-Terminal Amino Acid Sequence.<sup>14<\/sup> It also used the dataset from SWISS-PROT database, from which all sequences were extracted and some of the inappropriate sequences were removed before redundancy reduction. It was undertaken to avoid problems related to redundant data during Neural Networks training and testing. A \u00a0success rate of 85% (plant) or 90% \u00a0(non plant) on redundancy reduced test sets were observed.<\/p>\n<p>Garg et el.<sup>15<\/sup> introduced a systematic\u00a0 ap- proach for predicting subcellular localizations(SL) of human proteins. A set of human proteins with experimentally\u00a0 annotated \u00a0SL has been \u00a0retrieved from the SWISS-PROT database.<sup>16<\/sup> The final dataset consists of 3780 protein sequences that belong to 11 SL. The SVM-based modules for predicting SL using traditional amino acid and dipeptide (i+1) composition achieved accuracy of 76.6% and 77.8%. PSI-BLAST, when carried out using a similarity-based search against a nonredun- dant database of experimentally annotated proteins, yielded 73.3% accuracy. Yu-Dong at el.<sup>8<\/sup> pro- posed a new method for predicting the membrane protein types using the Nearest Neighbor Algorithm. They used manually constructed dataset from Swiss- Prot (http:\/\/cn.expasy.org\/, release 51.2)<sup>17<\/sup> mainly according to the annotation line stated as SL, to classify the six types of membrane proteins. The predictor achieved the accuracy of 87.02 by using the 56 most contributive features %.<\/p>\n<p>Lipeng at el.<sup>18<\/sup> proposed a new method in which, protein can be represented by a high di- mensional feature vector by using Dipeptide com- position method. They used only 2059 membrane protein \u00a0sequences from the dataset prepared by Chou and Elord. Based on the reduced low dimen- sional features KNN classifier was introduced to identify the membrane protein types, with predic- tion accuracy of 82.0%.<sup>13<\/sup> Jei Lein et el.<sup>19<\/sup> classified protein based on Chou\u2019s pseudo amino acid compostion with an Ensemble classifier. The protein locations are classified into 5 types. The testing and training dataset that they used originally was prepared by Cedano et al. (1997).<sup>20<\/sup> The com- posite KNN classifier predicted the proteins with location types (1) nuclear proteins, (2) intracellular proteins (non-nuclear), (3) extracellular proteins, (4) anchored membrane proteins, and (5) integral membrane proteins (M,A,E,I,N) with accuracy of 90.0%, 70.8%, 74.2%, 81.5%, 82.5% respectively.<\/p>\n<table style=\"width: 70%;\" border=\"1\" cellpadding=\"5\">\n<tbody>\n<tr>\n<td>\u00a0<img decoding=\"async\" class=\"alignnone size-thumbnail wp-image-19080\" src=\"https:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig31-150x150.jpg\" alt=\"Figure 3: Classification of Membrane Proteins\" width=\"150\" height=\"150\" srcset=\"https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig31-150x150.jpg 150w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig31-256x256.jpg 256w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig31.jpg 787w\" sizes=\"(max-width: 150px) 100vw, 150px\" \/><\/td>\n<td><strong>Figure 3: Classification of Membrane Proteins<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p><a href=\"http:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig31.jpg\" target=\"_blank\">Click here to View figure<\/a><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p>For classifying 6 types of membrane proteins 3 methods such as, BLAST\/PSI-BLAST Method, Network-Based Method, Shortest-Distance Method were introduced by Huang at el.<sup>21<\/sup> They proposed an integrated approach to predict multiple types of membrane proteins by employing sequence homology and protein-protein interaction network.<sup>22<\/sup> According to their positions and intramolecular arrangements in a cell, membrane proteins are \u00a0classified \u00a0into \u00a0six \u00a0types : (1) GPI (Glycosylphosphatidylinisotol)-anchor; (2) Lipid-anchor (LCM); (3) Multi-pass (MPT); (4) Peripheral(PM); (5) Single-pass type I; (6)Single- pass type II membrane proteins shown in Fig 3. To evaluate the performance of classification method, the sequence clustering program CD-HIT was employed (Cluster Database at High Identity with Tolerance)<sup>23<\/sup> to construct three datasets: S1, S2, S3 from 3789 proteins. S1 contained 2935 protein sequences in which protein had less than 70% sequence similarity. S2 contained 2120 protein sequences in which protein had sequence similarity lower than 40%. S3 contained 1475 protein sequences with sequence identity less than 25%. The BLAST\/PSI-BLAST method achieved the best performance with the highest accuracy 94.71%, 91.15% and 85.02% on datasets S1, S2 and S3, respectively. However, 481, 529 and 620 proteins cannot be annotated from data set. The network-based method achieved the second highest accuracy, i.e. 66.68%, 62.46%, 58.75% on the three datasets,S1,S2,S3 respectively. Since no interactive proteins can be found in the corresponding datasets, there were\u00a0 86, \u00a038, \u00a041 \u00a0proteins unannotated. The shortest distance method was capable of annotating all proteins, although it was least effective with lowest Accuracy achieved (54.97%, 48.75%, 44.99% on the three datasets, respectively). The proposed method is capable of annotating all proteins from the dataset S1. It uses 967 features from each of the membrane protein sequences.<\/p>\n<p><strong>Method<\/strong><strong>s and Methodology<\/strong><\/p>\n<p><strong>Dataset<\/strong><\/p>\n<p>A total of 3789 human membrane protein sequence were downloaded and verified from Uniprot Protein database (release 2012). To eval- uate the performance of the \u00a0prediction method, W. Li et al<sup>23<\/sup> use the sequence clustering pro-gram CD-HIT (Cluster Database at Height Identity Tolerance)<sup>24<\/sup>\u00a0 to prepare the benchmark set of data S1 from 3789, containing 2935 proteins sequences with sequence similarity less than 70%. In our pro-<\/p>\n<p>&nbsp;<\/p>\n<table style=\"width: 70%;\" border=\"1\" cellpadding=\"5\">\n<tbody>\n<tr>\n<td>\u00a0<img decoding=\"async\" class=\"alignnone size-thumbnail wp-image-19081\" src=\"https:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig41-150x150.jpg\" alt=\"Figure 4: The work flow diagram of the DT method posed method we use the dataset S1(2935 proteins) used for classification.\" width=\"150\" height=\"150\" srcset=\"https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig41-150x150.jpg 150w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig41-256x256.jpg 256w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig41.jpg 379w\" sizes=\"(max-width: 150px) 100vw, 150px\" \/><\/td>\n<td><strong>Figure 4: The work flow diagram of the DT method posed method we use the dataset S1 (2935 proteins) used for classification.<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p><a href=\"http:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig41.jpg\" target=\"_blank\">Click here to View figure<\/a><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p><strong>Methodology<\/strong><\/p>\n<p>The flow diagram for the proposed methodology is in Fig: 4 and the step by step procedures are as follows:<\/p>\n<p>Step 1: Start.<\/p>\n<p>Step 2: Input Dataset S1 (2935 membrane protein seuence)<\/p>\n<p>Step 3: Preprocessing the data from the data set S1, and create position specific scoring matrix (PSSM)<\/p>\n<p>Step 4: Extract the feature set from the dataset S1.<\/p>\n<p>Step 5: Apply the DT classifier algorithm for classifying memberane protien types.<\/p>\n<p>step 6: Evaluate the performance matrices.<\/p>\n<p>step 7: stop.<\/p>\n<p><strong>Preprocessing of Data<\/strong><\/p>\n<p>The S1 datasets of proteins are preprocessed according to their types and Protein id from the training dataset. For this create a Position Specific Scoring Matrix (PSSM) of the datasets. The PSSM is the numerical representa- tion of proteins in the dataset, which are presented in the 6 types of membrane proteins. The PSSM matrix consists of zeros and ones. If a Protein is presented in one or more membrane protein type, its entry in PSSM matrix is represented with ones, otherwise it is represented as zeros. This PSSM matrix is used for the evaluation of performance metrics.<\/p>\n<p>Feature Extraction: Features are usually extracted from the protein sequence. A sequence comprises of 20 unique amino acids namely A,C,D,E,F,G,H,I,J,K,L,M,N,P,Q,R,S, T,V, W, and Y. Even though all amino acids have a common basic chemical structure, they exhibits different chemical properties becuase of the differences in their side chains. Proteins are represented by a chain of amino acids. The difference in the amino acid string among proteins is due to their order and total number (length of the \u00a0sequence). The proposed DT classification used 968 distinct features.Extracted features are as follows:<\/p>\n<p><strong>Sequence length<\/strong><\/p>\n<p>The total number of amino acids in the given protein sequence. For example: the sequence length of \u2019acdfgyrsmeacvss\u2019 is 15<\/p>\n<p><strong>Hydrophobicity<\/strong><\/p>\n<p>The hydrophobicity of an amino acid is related to its transfer free energy from a polar medium (such as the cytoplasm) to another polar medium (like a membrane). The transfer free energy depends on the chemical nature of the two solvents, as well as on the structural context of the amino acid residue. The hydrophobicity index is a measure of the relative hydrophobicity i.e, this index is used to measure the hydrophobic affinity of a protein sequence or an amino acid sequence. In a protein, hydrophobic amino acids are likely to be found in the interior, whereas hydrophilic amino acids are likely to be in contact with the aqueous environment.<\/p>\n<p><strong>AAindex<\/strong><\/p>\n<p>It is a database representing numerical values of various physicochemical and biochemical properties of amino acids and pairs of amino acids. AAindex<sup>25<\/sup> for the amino acid index of 20 numerical values. It gives a total 544 features. Every year the updated version(9.0) of AAindex is released<\/p>\n<p><strong>Di-Amino Acid<\/strong><\/p>\n<p>Amino acids frequency is the number of combinations of amino acid residue. The count of the combination of sequence pattern \u00a0AA,AC,.., AY,CA,CC,&#8230;CY, and..,YA,YC, .., YY in the protein sequence is called the amino acid frequency. From this ,only count the combination of sequence patterns of Amino acid A,C,D, E. For example the sequence AA, AC, AD, AE,.. AY (20 numbers) and CA, CC, CD, CE&#8230;CY (20 numbers), and DA, DC, DD, &#8230;, DY (20 numbers) and EA, EC, ED,&#8230;EY (20 numbers) are counted. As a total of 400 features are generated as frequency for a particular Protein sequence.<\/p>\n<p><strong>Count of Each Amino Acid Residues<\/strong><\/p>\n<p>Amino Acid residues are the building block of proteins. Count of each amino acid residue is one of the feature used. For example, let \u2019AANDCC\u2019 be a amino acid sequence, count of amino acid residue A is 2, D is 1, C is 2 and N is 1. A total 20 features are collected as count for each amino acid.<\/p>\n<p><strong>Molecular Weight<\/strong><\/p>\n<p>Molecular weight is the mass of a molecule. The size of a protein can be represented with the number of amino acids con- tained in that protein or by using molecular weight. It is represented by unit of Daltons or in Kilo Daltons (KDa). (http:\/\/www.sciencegateway.org\/tools\/) tools used for finding the molecular weight of a protein from its protein sequence. For example, molecu- lar weight of the sequence \u2019\u2019Acdefghiklmn- Pqrstvwy\u2019 is 2.4 kilodaltons, and protein with protein id Q9P299 has the molecular \u00a0weight of 23679.0820 KDa.<\/p>\n<p><strong>Decision Tree Classification (DT)<\/strong><\/p>\n<p>A DT is a decision support tool that uses a tree structure graph or model of decisions and their possible consequences, including chance event outcomes, resource costs, and utility. Decision trees are com- monly used in operations research, specifically in decision analysis, to identify a strategy most likely to reach a goal. A DT is a flowchart-like struc- ture:internal node represents a test on an attribute, branch represents the outcome of the test and leaf\u00a0node represents a class label (decision taken after computing all attributes). The paths from root to leaf represents classification rules. The extracted features are \u00a0provided\u00a0 as \u00a0input \u00a0to \u00a0DT \u00a0classifier. The \u00a0DT classifies the protein types into six types according to the rule, with classification accuracies 69.81%, on the dataset S1.<\/p>\n<p><strong>Table 1: List of Features from Dataset S1<\/strong><\/p>\n<table style=\"width: 95%;\" border=\"1\" cellspacing=\"0\" cellpadding=\"4\">\n<tbody>\n<tr>\n<td style=\"text-align: center;\" width=\"157\"><strong>Di \u2013Amino Acid<\/strong><\/td>\n<td style=\"text-align: center;\" width=\"64\"><strong>400<\/strong><\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\" width=\"157\">Hydrophobicity<\/td>\n<td style=\"text-align: center;\" width=\"64\">1<\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\" width=\"157\">Aaindex<\/td>\n<td style=\"text-align: center;\" width=\"64\">544<\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\" width=\"157\">Count of Each Amino Acids<\/td>\n<td style=\"text-align: center;\" width=\"64\">20<\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\" width=\"157\">Sequence Molecular<\/td>\n<td style=\"text-align: center;\" width=\"64\">1<\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\" width=\"157\">Weight<\/td>\n<td style=\"text-align: center;\" width=\"64\">1<\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\" width=\"157\">Total<\/td>\n<td style=\"text-align: center;\" width=\"64\">967<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p><strong>Performance Metrics<\/strong><\/p>\n<p>The overall classification accuracy of a classification model is evaluated using Self Consistency test. It use training and testing the model with same dataset. This test is also termed as Resubstitution test, which is used to test the dataset. For multi-label classification, the concepts such as Precision, Recall, Accuracy<sup>26<\/sup> are used to measure the performance of methods. The following standard parameters are used to evaluate the performance of clasifier.<sup>27<\/sup>\u00a0 In order to find the values of Precision, Recall, Accuracy, calculate the True Positive(tp), True Negative(tn), False Positive (fp), False Nega- tive (fn). For that calculate the count of 1 values and 0 values in actual score matrix. Then generate the total count of 0 and 1 as N. Next calculate tp which is the count of 1 values in the intersection of actual score matrix and predicted score matrix. Similarly tn is the count of 0 values in the intersection of actual score matrix and predicted score matrix. fp and fn are calculated using the equation (1) and (2),<\/p>\n<p>fp = n \u2212 tn \u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0 (1)<\/p>\n<p>fn = p \u2212 tp\u00a0 \u00a0 \u00a0 \u00a0 \u00a0 \u00a0 \u00a0 \u00a0 \u00a0 \u00a0 \u00a0 \u00a0 \u00a0 \u00a0 \u00a0(2)<\/p>\n<p>Using these values, calculate Accuracy, Precision, Recall from the following equations.<\/p>\n<p><strong>Accuracy<\/strong><\/p>\n<p>It is the percentage prediction of true examples ie, True prediction divided by the total number of examples. The accuracy is defined by the equation (3), but more generalised form is shown in the equation (4)<\/p>\n<p>Accuracy = (tp + tn)\/N\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0 (3)<\/p>\n<p>Let D is a dataset with N instances. Let Yi and Zi are the set of original and predicted labels, respectively, where i D, then the accuracy becomes,<\/p>\n<p><img decoding=\"async\" class=\"alignnone size-full wp-image-19072\" src=\"https:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_f4.jpg\" alt=\"Equation 4\" width=\"368\" height=\"69\" srcset=\"https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_f4-300x56.jpg 300w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_f4.jpg 368w\" sizes=\"(max-width: 368px) 100vw, 368px\" \/><\/p>\n<p><strong>Precision<\/strong><\/p>\n<p>It is the number of correct predictions divided by the number of all returned prediction. It is calculated using the following equa- tion (5), but more generalized form is shown in the equation (6)<\/p>\n<p>P = tp + fp\u00a0 P recision = tp\/p \u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0\u00a0 (5)<\/p>\n<p><img decoding=\"async\" class=\"alignnone size-full wp-image-19073\" src=\"https:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_f6.jpg\" alt=\"Equation 6\" width=\"357\" height=\"54\" srcset=\"https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_f6-300x45.jpg 300w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_f6.jpg 357w\" sizes=\"(max-width: 357px) 100vw, 357px\" \/><\/p>\n<p><strong>Recall<\/strong><\/p>\n<p>It is the number of correct pre- dictions divided by the number of predictions. It is calculated using the following equation (7), but more generalised form is shown in the equation (8)<\/p>\n<p>p = tp + f n Recall = tp\/p\u00a0 \u00a0 \u00a0 \u00a0 (7)<\/p>\n<p><img decoding=\"async\" class=\"alignnone size-full wp-image-19074\" src=\"https:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_f8.jpg\" alt=\"Equation 8\" width=\"350\" height=\"60\" srcset=\"https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_f8-300x51.jpg 300w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_f8.jpg 350w\" sizes=\"(max-width: 350px) 100vw, 350px\" \/><\/p>\n<p><strong>Resu<\/strong><strong>l<\/strong><strong>t<\/strong><strong>s and Discussion<\/strong><\/p>\n<p>This section depicts the results of both existing Network Based Method, Shortest Distance Method and proposed DT classification. The results of pro- posed method is compared with results of existing methods. From the analysis, the Decision Tree clas- sifier is an efficient multi-label classifier for classify- ing the human membrane proteins into the following six classes, (1) Single -pass type I, (2)Single-pass type II, (3) Multi-pass, (4) Lipid-anchor, (5) GPI (Glycosylphosphatidylinisotol)-anchor, (6)\u00a0 Periph- eral membrane proteins.<\/p>\n<p>The proposed DT classification Results are shown in the Table. II. The Fig.5 illustrate the pie chart representation of decision tree classification on dataset S1. The multipass, lipid, GPI, peripheral, type1, \u00a0type2 \u00a0membrane\u00a0 proteins \u00a0are \u00a0represented by the colours, green, yellow, orange, brown, dark blue, light blue respectively, From the 2935 proteins from S1, more number of proteins are classified as multipass membrane proteins and less number of proteins as GPI anchored membrane proteins.<\/p>\n<table style=\"width: 70%;\" border=\"1\" cellpadding=\"5\">\n<tbody>\n<tr>\n<td>\u00a0<img decoding=\"async\" class=\"alignnone size-thumbnail wp-image-19082\" src=\"https:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig5-150x150.jpg\" alt=\"Figure 5: Pai-chart:Decision tree classification for dataset S1\" width=\"150\" height=\"150\" srcset=\"https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig5-150x150.jpg 150w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig5-256x256.jpg 256w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig5.jpg 467w\" sizes=\"(max-width: 150px) 100vw, 150px\" \/><\/td>\n<td><strong>Figure 5: Pai-chart:Decision tree classification for dataset S1<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p><a href=\"http:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig5.jpg\" target=\"_blank\">Click here to View figure<\/a><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p>Each membrane protein can have labeled in one or more types, Fig.6 shows the number of pro- teins having 1-6 types of the dataset S1. In dataset S1, \u00a0almost 510 \u00a0membrane proteins are \u00a0classified as Type1, 119 proteins as Type2, 1543 proteins as Multipass, 157 proteins as Lipid, 57 as GPI, and 549 proteins as Peripheral. Therefore in DT classifica- tion ,the more number of proteins are classified as Multipass and the less number of proteins as GPI in all the dataset s1. The Accuracy, Precision, Recall, \u00a0are \u00a0calculated\u00a0 and \u00a0the \u00a0results \u00a0are \u00a0shown in Table II. The Multi label Protein classification using DT gives better results with all the annotated proteins, when compared to the existing methods with few number of unannotated proteins. The clas- sification accuracy is reached 69.81% on dataset S1<\/p>\n<table style=\"width: 70%;\" border=\"1\" cellpadding=\"5\">\n<tbody>\n<tr>\n<td>\u00a0<img decoding=\"async\" class=\"alignnone size-thumbnail wp-image-19083\" src=\"https:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig6-150x150.jpg\" alt=\"Figure 6: Different Types of Membrane Proteins on Dataset S1\" width=\"150\" height=\"150\" srcset=\"https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig6-150x150.jpg 150w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig6-256x256.jpg 256w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig6.jpg 645w\" sizes=\"(max-width: 150px) 100vw, 150px\" \/><\/td>\n<td><strong>Figure 6: Different Types of Membrane Proteins on Dataset S1<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p><a href=\"http:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig6.jpg\" target=\"_blank\">Click here to View figure<\/a><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p><strong>Table 2: Performances of Decision Tree classification tested on dataset s1<\/strong><\/p>\n<table style=\"width: 95%;\" border=\"1\" cellspacing=\"0\" cellpadding=\"4\">\n<tbody>\n<tr>\n<td style=\"text-align: center;\" width=\"147\"><strong>Dataset<\/strong><\/td>\n<td style=\"text-align: center;\" width=\"64\"><strong>Accur Acy<\/strong><\/td>\n<td style=\"text-align: center;\" width=\"64\"><strong>Precisi on<\/strong><\/td>\n<td style=\"text-align: center;\" width=\"64\"><strong>Recall<\/strong><\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\" width=\"147\">S1<\/td>\n<td style=\"text-align: center;\" width=\"64\">69.81%<\/td>\n<td style=\"text-align: center;\" width=\"64\">0.1165<\/td>\n<td style=\"text-align: center;\" width=\"64\">0.1157<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p>The proposed DT classification performs clas- sification on the dataset S1. This method uses the\u00a0ONE type, TWO types,THREE types.Y axis repre- sent the total count of correct predicted membrane proteins in each type. 2482 one type membrane proteins and 34 two type membrane proteins .Very few of them are partially predicted in and some of them are not correctly predicted.<\/p>\n<p>The Fig.7 shows the correct prediction of differ- ent multi-type membrane proteins in data set S1. X axis represent the types of membrane proteins like whole number of proteins from the dataset for the classification purpose. Its classification accuracies are presented in the Table. III. It is obvious that the DT method contributed the most, annotating 2935 proteins and achieved Accuracy of 69.81%, on datasets S1, but the network-based method number of annotated protein are 467 from the dataset S1, and\u00a0 obtained Accuracy of 66.68%, and shortest-distance method with Accuracy of 54.97%, on the dataset S1.<\/p>\n<table style=\"width: 70%;\" border=\"1\" cellpadding=\"5\">\n<tbody>\n<tr>\n<td><img decoding=\"async\" class=\"alignnone size-thumbnail wp-image-19084\" src=\"https:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig7-150x150.jpg\" alt=\"Figure 7: The Distribution of Correct prediction of different multitype proteins in data set S1\" width=\"150\" height=\"150\" srcset=\"https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig7-150x150.jpg 150w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig7-256x256.jpg 256w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig7.jpg 667w\" sizes=\"(max-width: 150px) 100vw, 150px\" \/><\/td>\n<td><strong>Figure 7: The Distribution of Correct prediction of different multitype proteins in data set S1<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p><a href=\"http:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig7.jpg\" target=\"_blank\">Click here to View figure<\/a><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p><strong>Table 3: Comparison Of Classification Method<\/strong><\/p>\n<table style=\"width: 95%;\" border=\"1\" cellspacing=\"0\" cellpadding=\"4\">\n<tbody>\n<tr>\n<td style=\"text-align: center;\" width=\"65\"><strong>Dataset<\/strong><\/td>\n<td style=\"text-align: center;\" colspan=\"6\" width=\"280\"><strong>Classification\u00a0 Methods<\/strong><\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\" rowspan=\"2\" width=\"65\">S1<\/td>\n<td style=\"text-align: center;\" colspan=\"2\" width=\"98\"><strong>Network<\/strong><\/td>\n<td style=\"text-align: center;\" colspan=\"2\" width=\"91\"><strong>Shortest<\/strong><\/td>\n<td style=\"text-align: center;\" colspan=\"2\" width=\"91\"><strong>\u00a0 \u00a0DT<\/strong><\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\" width=\"57\">ACC<\/p>\n<p>66.68%<\/td>\n<td style=\"text-align: center;\" width=\"41\">NU*<\/p>\n<p>86<\/td>\n<td style=\"text-align: center;\" width=\"49\">ACC<\/p>\n<p>54.97%<\/td>\n<td style=\"text-align: center;\" width=\"41\">NU*<\/p>\n<p>0<\/td>\n<td style=\"text-align: center;\" width=\"49\">ACC<\/p>\n<p>69.81%<\/td>\n<td style=\"text-align: center;\" width=\"41\">NU*<\/p>\n<p>0<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p>The Fig 8 shows the performance of exist-ing and proposed method accuracies. This bar graph shows the classification methods like Network method(NWM), Shortest distance method(SDM), and the proposed Decision tree(DT) classification in Y axis and the corresponding accuracies in the X axis.<\/p>\n<table style=\"width: 70%;\" border=\"1\" cellpadding=\"5\">\n<tbody>\n<tr>\n<td>\u00a0<img decoding=\"async\" class=\"alignnone size-thumbnail wp-image-19085\" src=\"https:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig8-150x150.jpg\" alt=\"Figure 8: Accuracies of The Dataset in 3 Different Classification Methods 1) Network Method (NWM) 2) Shortest Distance Method(SDM) 3)Decision Tree Classification (DT)\" width=\"150\" height=\"150\" srcset=\"https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig8-150x150.jpg 150w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig8-256x256.jpg 256w, https:\/\/biomedpharmajournal.org\/staging\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig8.jpg 673w\" sizes=\"(max-width: 150px) 100vw, 150px\" \/><\/td>\n<td><strong>Figure 8: Accuracies of The Dataset in 3 Different\u00a0<\/strong><strong>Classification Methods 1) Network Method (NWM)\u00a0<\/strong><strong>2) Shortest Distance Method(SDM) 3)Decision Tree\u00a0<\/strong><strong>Classification (DT)<\/strong><\/p>\n<p>&nbsp;<\/p>\n<p><a href=\"http:\/\/biomedpharmajournal.org\/wp-content\/uploads\/2018\/02\/Vol11No1_App_Mar_fig8.jpg\" target=\"_blank\">Click here to View figure<\/a><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<p><strong>Conclusio<\/strong><strong>n<\/strong><\/p>\n<p>In multi-label classification ,each sample can be associated with a set of class labels.This paper proposed a DT classification algorithm. The 2935 membrane proteins of the datasets S1 are classified using Decision Tree based on the 967 features extracted from these proteins. As a result, the Deci- sion Tree classifier with most contributive features\u00a0achieved an acceptable accuracy of 69.81% of the dataset compared to the existing network based and shortest path method.<\/p>\n<p><strong>C<\/strong><strong>on<\/strong><strong>flict of Interest Statement<\/strong><\/p>\n<p>The authors declare that they have no competing interests.<\/p>\n<p><strong>Acknowledgment<\/strong><\/p>\n<p>The authors would like to acknowledge to Guohua Huang, Institute of system biology, China, Yuchao Zhang, Department of Maths, Shoyang University, China, Lei Chan, CIE, SM University, China, Ning Zhang, Dept. of BME, Tianjin University, China, supporting for this work.<\/p>\n<p><strong>R<\/strong><strong>eference<\/strong><strong>s<\/strong><\/p>\n<ol>\n<li>\u00a0Almen M.S, Nordstrom K.J, Fredriksson R and\u00a0 Schioth H. B. \u201cMapping the human membrane proteome: a majority of the human membrane proteins can be classified according to function and evolutionary origin,\u201d <em>BMC biology.<\/em>\u00a0 2009;7(1):1.<\/li>\n<li>\u00a0Gao Q-B,Ye X-F,\u00a0 Jin Z.-C and\u00a0 He J. \u201cImproving discrimi- nation of outer membrane proteins by fusing different forms of pseudo amino acid composition,\u201d <em>Analytical biochemistry.<\/em> 2010;398(1):52\u201359.<br \/>\n<a href=\"https:\/\/doi.org\/10.1016\/j.ab.2009.10.040\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>Krogh A, Larsson B, Heijne G.V and Sonnhammer E. L. \u201cPredicting transmembrane protein topology with a hidden markov model: application to complete genomes,\u201d <em>Journal\u00a0 of molecular biology.<\/em> 2001;305(3):567\u2013580.<br \/>\n<a href=\"https:\/\/doi.org\/10.1006\/jmbi.2000.4315\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>\u00a0Arinaminpathy Y, Khurana E, Engelman D.M and\u00a0 Gerstein\u00a0M.B. \u201cComputational analysis of membrane proteins: the largest class of drug targets,\u201d <em>Drug discovery today<\/em>.\u00a0 2009;14(23):1130\u20131135.<br \/>\n<a href=\"https:\/\/doi.org\/10.1016\/j.drudis.2009.08.006\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>Davey J. \u201cG-protein-coupled receptors: new approaches to maximise the impact of gpcrs in drug discovery,\u201d <em>Expert opinion on therapeutic targets<\/em>.\u00a0 2004;8(2):165\u2013170.<br \/>\n<a href=\"https:\/\/doi.org\/10.1517\/14728222.8.2.165\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>\u00a0Terstappen G. C and Reggiani A. \u201cIn silico research in drug discovery,\u201d <em>Trends in pharmacological sciences.\u00a0<\/em> 2001;22(1):23\u201326.<br \/>\n<a href=\"https:\/\/doi.org\/10.1016\/S0165-6147(00)01584-4\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>Wang J,Li Y,\u00a0 Wang Q, You X, Man J, Wang C and Gao X. \u201cProclusensem: predicting membrane protein types by fusing different modes of pseudo amino acid composition,\u201d <em>Computers in biology and medicine<\/em>.\u00a0 2012;42(5):564\u2013574.<br \/>\n<a href=\"https:\/\/doi.org\/10.1016\/j.compbiomed.2012.01.012\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>\u00a0Jia P, Qian Z, Feng K, Lu W, Li Y and Cai Y. \u201cPrediction of membrane protein types in a hybrid space,\u201d <em>Journal of proteome research<\/em>.\u00a02008;7(3):1131\u20131137.<br \/>\n<a href=\"https:\/\/doi.org\/10.1021\/pr700715c\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>\u00a0Lodish H, Baltimore D, Berk\u00a0 A, Zipursky S.L, Matsudaira P and Darnell J. Molecular cell biology.\u00a0 Scientific American Books New York. 1995;3.<\/li>\n<li>\u00a0Chou K.C and\u00a0 Cai Y.D.\u201cPrediction of membrane protein types by incorporating amphipathic effects,\u201d <em>Journal\u00a0 of chem- ical \u00a0information and \u00a0modeling.<\/em>\u00a0 2005;45(2):407\u2013413.<br \/>\n<a href=\"https:\/\/doi.org\/10.1021\/ci049686v\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>A. A. D. H. C. C. E. K. Murzin A. G. \u201cScop2 prototype: a new approach to protein structure mining,\u201d <em>Nucleic Acids Research.\u00a0<\/em> 2014;42:d310\u2013d314.<br \/>\n<a href=\"https:\/\/doi.org\/10.1093\/nar\/gkt1242\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>Overington J.P, Al-Lazikani B. and\u00a0 Hopkins A.L.\u00a0 How many drug targets are there? <em>Nature reviews Drug discovery<\/em>.\u00a02006;5(12):993\u2013996.<br \/>\n<a href=\"https:\/\/doi.org\/10.1038\/nrd2199\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>Chou K.C and\u00a0 Elrod D.W. \u201cPrediction of membrane protein types and subcellular locations,\u201d<em> Proteins: Structure, Function and Bioinformatics<\/em>. 1999;34(1):137\u2013153.<br \/>\n<a href=\"https:\/\/doi.org\/10.1002\/(SICI)1097-0134(19990101)34:1&lt;137::AID-PROT11&gt;3.0.CO;2-O\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>\u00a0Emanuelsson O,\u00a0 Nielsen H,Brunak S and Heijne G.V. \u201cPredicting subcellular localization of proteins based on their n-terminal amino acid sequence,\u201d<em>Journal of molecular biology<\/em>. 2000;300(4):1005\u20131016.<br \/>\n<a href=\"https:\/\/doi.org\/10.1006\/jmbi.2000.3903\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>Garg A,Bhasin M and\u00a0 Raghava\u00a0G. P. \u201cSupport vector machine-based method for subcellular localization of human proteins using amino acid compositions, their order, and simi- larity search,\u201d<em>Journal of biological Chemistry.<\/em>\u00a0 2005;280(15):14 427\u201314 432.<\/li>\n<li>\u00a0Bairoch A and\u00a0 Apweiler R. \u201cThe swiss-prot protein sequence database and its supplement trembl in 2000,\u201d<em>Nucleic acids research.<\/em>\u00a0200;28(1): 45\u201348.<br \/>\n<a href=\"https:\/\/doi.org\/10.1093\/nar\/28.1.45\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>Boeckmann B,\u00a0 \u00a0Bairoch A, Apweiler R, Blatter M.C,\u00a0 Estreicher A,\u00a0 Gasteiger E,\u00a0 \u00a0 Martin M.J, Michoud\u00a0K,Donovan C.O, Phan I. et al. The swiss-prot protein knowl- edgebase and its supplement trembl in 2003, <em>Nucleic acids research.<\/em> <strong> 2003;<\/strong>31(1):365\u2013370.<br \/>\n<a href=\"https:\/\/doi.org\/10.1093\/nar\/gkg095\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>\u00a0Wang L,\u00a0 Yuan Z,Chen X and Zhou Z. The prediction of membrane protein types with npe, IEICE Electronics Express. 2010;7(6):397\u2013402.<br \/>\n<a href=\"https:\/\/doi.org\/10.1587\/elex.7.397\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>Lin J,\u00a0 Wang Y and Xu X. \u201cA novel ensemble and composite approach for classifying proteins based on chous pseudo amino acid composition,\u201d <em>African Journal \u00a0of Biotechnology.<\/em>\u00a0 2011;10(74):16 948\u201316 952.<\/li>\n<li>\u00a0Cedano J, Aloy P, Perez-Pons J. A and\u00a0 Querol E.\u201cRela- tion between amino acid composition and cellular location of proteins,\u201d <em>Journal \u00a0of molecular\u00a0 biology<\/em>. 1997;266(3):594\u2013600.<br \/>\n<a href=\"https:\/\/doi.org\/10.1006\/jmbi.1996.0804\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>Huang G,\u00a0 Zhang Y,\u00a0 Chen L,\u00a0 Zhang N, Huang T and\u00a0 Cai Y.D. \u201cPrediction of multi-type membrane proteins in human by an integrated approach,\u201d <em>PloS one<\/em>.2014;9(3):e93553.<br \/>\n<a href=\"https:\/\/doi.org\/10.1371\/journal.pone.0093553\" target=\"_blank\">CrossRef<\/a><\/li>\n<li>\u00a0Consortium U. et al. \u201cThe universal protein resource (uniprot) in 2010,\u201d <em>Nucleic acids research<\/em>.\u00a02010;38(1):D142\u2013 D148.<\/li>\n<\/ol>\n","protected":false},"excerpt":{"rendered":"<p>Introduction Multilabel classification methods are progres- sively used in recent  [&#8230;]<\/p>\n","protected":false},"author":9,"featured_media":0,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[55],"tags":[],"class_list":["post-19054","post","type-post","status-publish","format-standard","hentry","category-vol11no1"],"_links":{"self":[{"href":"https:\/\/biomedpharmajournal.org\/staging\/wp-json\/wp\/v2\/posts\/19054","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/biomedpharmajournal.org\/staging\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/biomedpharmajournal.org\/staging\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/biomedpharmajournal.org\/staging\/wp-json\/wp\/v2\/users\/9"}],"replies":[{"embeddable":true,"href":"https:\/\/biomedpharmajournal.org\/staging\/wp-json\/wp\/v2\/comments?post=19054"}],"version-history":[{"count":7,"href":"https:\/\/biomedpharmajournal.org\/staging\/wp-json\/wp\/v2\/posts\/19054\/revisions"}],"predecessor-version":[{"id":32105,"href":"https:\/\/biomedpharmajournal.org\/staging\/wp-json\/wp\/v2\/posts\/19054\/revisions\/32105"}],"wp:attachment":[{"href":"https:\/\/biomedpharmajournal.org\/staging\/wp-json\/wp\/v2\/media?parent=19054"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/biomedpharmajournal.org\/staging\/wp-json\/wp\/v2\/categories?post=19054"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/biomedpharmajournal.org\/staging\/wp-json\/wp\/v2\/tags?post=19054"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}