Your browser doesn't support javascript.
loading
Subcellular Localization Prediction of Human Proteins Using Multifeature Selection Methods.
Zhang, Yu-Hang; Ding, ShiJian; Chen, Lei; Huang, Tao; Cai, Yu-Dong.
Afiliação
  • Zhang YH; School of Life Sciences, Shanghai University, Shanghai 200444, China.
  • Ding S; Channing Division of Network Medicine, Brigham and Women's Hospital, Harvard Medical School, Boston, MA, USA.
  • Chen L; School of Life Sciences, Shanghai University, Shanghai 200444, China.
  • Huang T; College of Information Engineering, Shanghai Maritime University, Shanghai 201306, China.
  • Cai YD; Bio-Med Big Data Center, CAS Key Laboratory of Computational Biology, Shanghai Institute of Nutrition and Health, University of Chinese Academy of Sciences, Chinese Academy of Sciences, Shanghai 200031, China.
Biomed Res Int ; 2022: 3288527, 2022.
Article em En | MEDLINE | ID: mdl-36132086
Subcellular localization attempts to assign proteins to one of the cell compartments that performs specific biological functions. Finding the link between proteins, biological functions, and subcellular localization is an effective way to investigate the general organization of living cells in a systematic manner. However, determining the subcellular localization of proteins by traditional experimental approaches is difficult. Here, protein-protein interaction networks, functional enrichment on gene ontology and pathway, and a set of proteins having confirmed subcellular localization were applied to build prediction models for human protein subcellular localizations. To build an effective predictive model, we employed a variety of robust machine learning algorithms, including Boruta feature selection, minimum redundancy maximum relevance, Monte Carlo feature selection, and LightGBM. Then, the incremental feature selection method with random forest and support vector machine was used to discover the essential features. Furthermore, 38 key features were determined by integrating results of different feature selection methods, which may provide critical insights into the subcellular location of proteins. Their biological functions of subcellular localizations were discussed according to recent publications. In summary, our computational framework can help advance the understanding of subcellular localization prediction techniques and provide a new perspective to investigate the patterns of protein subcellular localization and their biological importance.
Assuntos

Texto completo: 1 Base de dados: MEDLINE Assunto principal: Proteínas / Biologia Computacional Idioma: En Ano de publicação: 2022 Tipo de documento: Article

Texto completo: 1 Base de dados: MEDLINE Assunto principal: Proteínas / Biologia Computacional Idioma: En Ano de publicação: 2022 Tipo de documento: Article