Warning: Undefined array key "mm" in /www/wwwroot/www.ai-bt.com/si.php on line 10 Deprecated: trim(): Passing null to parameter #1 ($string) of type string is deprecated in /www/wwwroot/www.ai-bt.com/si.php on line 10 Nonnegative principal component analysis for cancer molecular pattern discovery.

Literature DB >> 20671323

Nonnegative principal component analysis for cancer molecular pattern discovery.

Abstract

As a well-established feature selection algorithm, principal component analysis (PCA) is often combined with the state-of-the-art classification algorithms to identify cancer molecular patterns in microarray data. However, the algorithm's global feature selection mechanism prevents it from effectively capturing the latent data structures in the high-dimensional data. In this study, we investigate the benefit of adding nonnegative constraints on PCA and develop a nonnegative principal component analysis algorithm (NPCA) to overcome the global nature of PCA. A novel classification algorithm NPCA-SVM is proposed for microarray data pattern discovery. We report strong classification results from the NPCA-SVM algorithm on five benchmark microarray data sets by direct comparison with other related algorithms. We have also proved mathematically and interpreted biologically that microarray data will inevitably encounter overfitting for an SVM/PCA-SVM learning machine under a Gaussian kernel. In addition, we demonstrate that nonnegative principal component analysis can be used to capture meaningful biomarkers effectively.

Entities: Disease

Mesh：

Year: 2010 PMID： 20671323 DOI： 10.1109/TCBB.2009.36

Source DB: PubMed Journal: IEEE/ACM Trans Comput Biol Bioinform ISSN： 1545-5963 Impact factor: 3.710

Keyword Cloud
Cited

11 in total

1. A Self-Training Subspace Clustering Algorithm under Low-Rank Representation for Cancer Classification on Gene Expression Data.

Authors: Chun-Qiu Xia; Ke Han; Yong Qi; Yang Zhang; Dong-Jun Yu
Journal: IEEE/ACM Trans Comput Biol Bioinform Date: 2017-06-06 Impact factor: 3.710

8. Distribution based Fuzzy Estimate Spectral Clustering for Cancer Detection with Protein Sequence and Structural Motifs

Authors: Thenmozhi K; Karthikeyani Visalakshi N; Shanthi S
Journal: Asian Pac J Cancer Prev Date: 2018-07-27

9. Diagnostic biases in translational bioinformatics.

Authors: Henry Han
Journal: BMC Med Genomics Date: 2015-08-01 Impact factor: 3.063

10. Derivative component analysis for mass spectral serum proteomic profiles.

Authors: Henry Han
Journal: BMC Med Genomics Date: 2014-05-08 Impact factor: 3.063

Nonnegative principal component analysis for cancer molecular pattern discovery.

1. A Self-Training Subspace Clustering Algorithm under Low-Rank Representation for Cancer Classification on Gene Expression Data.

2. Transcriptome marker diagnostics using big data.

3. Nonnegative principal component analysis for mass spectral serum profiles and biomarker discovery.

4. Multi-resolution independent component analysis for high-performance tumor classification and biomarker discovery.

5. Learning a weighted meta-sample based parameter free sparse representation classification for microarray data.

Review 6. Overcome support vector machine diagnosis overfitting.

7. Disease Biomarker Query from RNA-Seq Data.

8. Distribution based Fuzzy Estimate Spectral Clustering for Cancer Detection with Protein Sequence and Structural Motifs

9. Diagnostic biases in translational bioinformatics.

10. Derivative component analysis for mass spectral serum proteomic profiles.