
pmid: 21569819
The three dimensional structure of a protein provides major insights into its function. Protein structure comparison has implications in functional and evolutionary studies. A structural alphabet (SA) is a library of local protein structure prototypes that can abstract every part of protein main chain conformation. Protein Blocks (PBs) is a widely used SA, composed of 16 prototypes, each representing a pentapeptide backbone conformation defined in terms of dihedral angles. Through this description, the 3D structural information can be translated into a 1D sequence of PBs. In a previous study, we have used this approach to compare protein structures encoded in terms of PBs. A classical sequence alignment procedure based on dynamic programming was used, with a dedicated PB Substitution Matrix (SM). PB-based pairwise structural alignment method gave an excellent performance, when compared to other established methods for mining. In this study, we have (i) refined the SMs and (ii) improved the Protein Block Alignment methodology (named as iPBA). The SM was normalized in regards to sequence and structural similarity. Alignment of protein structures often involves similar structural regions separated by dissimilar stretches. A dynamic programming algorithm that weighs these local similar stretches has been designed. Amino acid substitutions scores were also coupled linearly with the PB substitutions. iPBA improves (i) the mining efficiency rate by 6.8% and (ii) more than 82% of the alignments have a better quality. A higher efficiency in aligning multi-domain proteins could be also demonstrated. The quality of alignment is better than DALI and MUSTANG in 81.3% of the cases. Thus our study has resulted in an impressive improvement in the quality of protein structural alignment.
Models, Molecular, Protein Folding, [SDV.BIBS] Life Sciences [q-bio]/Quantitative Methods [q-bio.QM], semi-global alignment, Protein Conformation, Proteins, structural comparison, protein structure mining, 612, Molecular Biophysics Unit, Protein Data Bank, Protein Blocks, [SDV.BBM] Life Sciences [q-bio]/Biochemistry, Molecular Biology, structural alphabet, Databases, Protein, Sequence Alignment, amino acid, anchorbased alignment, protein folds, Algorithms, [INFO.INFO-BI] Computer Science [cs]/Bioinformatics [q-bio.QM]
Models, Molecular, Protein Folding, [SDV.BIBS] Life Sciences [q-bio]/Quantitative Methods [q-bio.QM], semi-global alignment, Protein Conformation, Proteins, structural comparison, protein structure mining, 612, Molecular Biophysics Unit, Protein Data Bank, Protein Blocks, [SDV.BBM] Life Sciences [q-bio]/Biochemistry, Molecular Biology, structural alphabet, Databases, Protein, Sequence Alignment, amino acid, anchorbased alignment, protein folds, Algorithms, [INFO.INFO-BI] Computer Science [cs]/Bioinformatics [q-bio.QM]
| selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | 35 | |
| popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network. | Top 10% | |
| influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically). | Top 10% | |
| impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network. | Top 10% |
