Warning: Undefined array key "mm" in /www/wwwroot/www.ai-bt.com/si.php on line 10 Deprecated: trim(): Passing null to parameter #1 ($string) of type string is deprecated in /www/wwwroot/www.ai-bt.com/si.php on line 10 Finding causative genes from high-dimensional data: an appraisal of statistical and machine learning approaches.

Literature DB >> 27226102

Finding causative genes from high-dimensional data: an appraisal of statistical and machine learning approaches.

Abstract

Modern biological experiments often involve high-dimensional data with thousands or more variables. A challenging problem is to identify the key variables that are related to a specific disease. Confounding this task is the vast number of statistical methods available for variable selection. For this reason, we set out to develop a framework to investigate the variable selection capability of statistical methods that are commonly applied to analyze high-dimensional biological datasets. Specifically, we designed six simulated cancers (based on benchmark colon and prostate cancer data) where we know precisely which genes cause a dataset to be classified as cancerous or normal - we call these causative genes. We found that not one statistical method tested could identify all the causative genes for all of the simulated cancers, even though increasing the sample size does improve the variable selection capabilities in most cases. Furthermore, certain statistical tools can classify our simulated data with a low error rate, yet the variables being used for classification are not necessarily the causative genes.

Entities: Disease

Mesh：

Year: 2016 PMID： 27226102 DOI： 10.1515/sagmb-2015-0072

Source DB: PubMed Journal: Stat Appl Genet Mol Biol ISSN： 1544-6115

Keyword Cloud
Cited

1 in total

1. Biological knowledge-slanted random forest approach for the classification of calcified aortic valve stenosis.

Authors: Erika Cantor; Rodrigo Salas; Harvey Rosas; Sandra Guauque-Olarte
Journal: BioData Min Date: 2021-07-23 Impact factor: 2.522

1 in total