Your browser doesn't support javascript.
loading
Mostrar: 20 | 50 | 100
Resultados 1 - 1 de 1
Filtrar
Mais filtros

Base de dados
Ano de publicação
Tipo de documento
Intervalo de ano de publicação
1.
PLoS One ; 10(5): e0128194, 2015.
Artigo em Inglês | MEDLINE | ID: mdl-26020952

RESUMO

BACKGROUND: T-cell epitopes play the important role in T-cell immune response, and they are critical components in the epitope-based vaccine design. Immunogenicity is the ability to trigger an immune response. The accurate prediction of immunogenic T-cell epitopes is significant for designing useful vaccines and understanding the immune system. METHODS: In this paper, we attempt to differentiate immunogenic epitopes from non-immunogenic epitopes based on their primary structures. First of all, we explore a variety of sequence-derived features, and analyze their relationship with epitope immunogenicity. To effectively utilize various features, a genetic algorithm (GA)-based ensemble method is proposed to determine the optimal feature subset and develop the high-accuracy ensemble model. In the GA optimization, a chromosome is to represent a feature subset in the search space. For each feature subset, the selected features are utilized to construct the base predictors, and an ensemble model is developed by taking the average of outputs from base predictors. The objective of GA is to search for the optimal feature subset, which leads to the ensemble model with the best cross validation AUC (area under ROC curve) on the training set. RESULTS: Two datasets named 'IMMA2' and 'PAAQD' are adopted as the benchmark datasets. Compared with the state-of-the-art methods POPI, POPISK, PAAQD and our previous method, the GA-based ensemble method produces much better performances, achieving the AUC score of 0.846 on IMMA2 dataset and the AUC score of 0.829 on PAAQD dataset. The statistical analysis demonstrates the performance improvements of GA-based ensemble method are statistically significant. CONCLUSIONS: The proposed method is a promising tool for predicting the immunogenic epitopes. The source codes and datasets are available in S1 File.


Assuntos
Algoritmos , Epitopos de Linfócito T/química , Modelos Genéticos , Modelos Imunológicos , Sequência de Aminoácidos , Simulação por Computador , Conjuntos de Dados como Assunto , Epitopos de Linfócito T/imunologia , Humanos , Dados de Sequência Molecular , Curva ROC , Linfócitos T/química , Linfócitos T/imunologia , Vacinas Sintéticas/biossíntese
SELEÇÃO DE REFERÊNCIAS
DETALHE DA PESQUISA