Your browser doesn't support javascript.
loading
Variable selection for generalized canonical correlation analysis.
Tenenhaus, Arthur; Philippe, Cathy; Guillemot, Vincent; Le Cao, Kim-Anh; Grill, Jacques; Frouin, Vincent.
Affiliation
  • Tenenhaus A; SUPELEC, Plateau de moulon, 3 rue Joliot-Curie, 91192 Gif-sur-Yvette Cedex, France arthur.tenenhaus@supelec.fr.
  • Philippe C; CNRS-IGR-Paris XI university, UMR8203, 94805 Villejuif cedex, France.
  • Guillemot V; NEUROSPIN, I2BM, CEA saclay, 91191 Gif-sur-Yvette cedex, France.
  • Le Cao KA; Queensland Facility for Advanced Bioinformatics, University of Queensland, 306 Carmody Road, St Lucia, QLD 4072, Australia.
  • Grill J; CNRS-IGR-Paris XI university, UMR8203, 94805 Villejuif cedex, France.
  • Frouin V; NEUROSPIN, I2BM, CEA saclay, 91191 Gif-sur-Yvette cedex, France.
Biostatistics ; 15(3): 569-83, 2014 Jul.
Article in En | MEDLINE | ID: mdl-24550197
ABSTRACT
Regularized generalized canonical correlation analysis (RGCCA) is a generalization of regularized canonical correlation analysis to 3 or more sets of variables. RGCCA is a component-based approach which aims to study the relationships between several sets of variables. The quality and interpretability of the RGCCA components are likely to be affected by the usefulness and relevance of the variables in each block. Therefore, it is an important issue to identify within each block which subsets of significant variables are active in the relationships between blocks. In this paper, RGCCA is extended to address the issue of variable selection. Specifically, sparse generalized canonical correlation analysis (SGCCA) is proposed to combine RGCCA with an [Formula see text]-penalty in a unified framework. Within this framework, blocks are not necessarily fully connected, which makes SGCCA a flexible method for analyzing a wide variety of practical problems. Finally, the versatility and usefulness of SGCCA are illustrated on a simulated dataset and on a 3-block dataset which combine gene expression, comparative genomic hybridization, and a qualitative phenotype measured on a set of 53 children with glioma. SGCCA is available on CRAN as part of the RGCCA package.
Subject(s)
Key words

Full text: 1 Collection: 01-internacional Database: MEDLINE Main subject: Data Interpretation, Statistical / Models, Statistical Type of study: Prognostic_studies / Qualitative_research / Risk_factors_studies Limits: Child / Humans Language: En Journal: Biostatistics Year: 2014 Type: Article Affiliation country: France

Full text: 1 Collection: 01-internacional Database: MEDLINE Main subject: Data Interpretation, Statistical / Models, Statistical Type of study: Prognostic_studies / Qualitative_research / Risk_factors_studies Limits: Child / Humans Language: En Journal: Biostatistics Year: 2014 Type: Article Affiliation country: France