Your browser doesn't support javascript.
loading
A deep learning framework for improving long-range residue-residue contact prediction using a hierarchical strategy.
Xiong, Dapeng; Zeng, Jianyang; Gong, Haipeng.
Afiliação
  • Xiong D; MOE Key Laboratory of Bioinformatics, School of Life Sciences.
  • Zeng J; Beijing Innovation Center of Structural Biology.
  • Gong H; Beijing Innovation Center of Structural Biology.
Bioinformatics ; 33(17): 2675-2683, 2017 Sep 01.
Article em En | MEDLINE | ID: mdl-28472263
ABSTRACT
MOTIVATION Residue-residue contacts are of great value for protein structure prediction, since contact information, especially from those long-range residue pairs, can significantly reduce the complexity of conformational sampling for protein structure prediction in practice. Despite progresses in the past decade on protein targets with abundant homologous sequences, accurate contact prediction for proteins with limited sequence information is still far from satisfaction. Methodologies for these hard targets still need further improvement.

RESULTS:

We presented a computational program DeepConPred, which includes a pipeline of two novel deep-learning-based methods (DeepCCon and DeepRCon) as well as a contact refinement step, to improve the prediction of long-range residue contacts from primary sequences. When compared with previous prediction approaches, our framework employed an effective scheme to identify optimal and important features for contact prediction, and was only trained with coevolutionary information derived from a limited number of homologous sequences to ensure robustness and usefulness for hard targets. Independent tests showed that 59.33%/49.97%, 64.39%/54.01% and 70.00%/59.81% of the top L/5, top L/10 and top 5 predictions were correct for CASP10/CASP11 proteins, respectively. In general, our algorithm ranked as one of the best methods for CASP targets. AVAILABILITY AND IMPLEMENTATION All source data and codes are available at http//166.111.152.91/Downloads.html . CONTACT hgong@tsinghua.edu.cn or zengjy321@tsinghua.edu.cn. SUPPLEMENTARY INFORMATION Supplementary data are available at Bioinformatics online.
Assuntos

Texto completo: 1 Base de dados: MEDLINE Assunto principal: Conformação Proteica / Software / Modelos Moleculares / Biologia Computacional / Aprendizado de Máquina Idioma: En Ano de publicação: 2017 Tipo de documento: Article

Texto completo: 1 Base de dados: MEDLINE Assunto principal: Conformação Proteica / Software / Modelos Moleculares / Biologia Computacional / Aprendizado de Máquina Idioma: En Ano de publicação: 2017 Tipo de documento: Article