Download | - View accepted manuscript: Semi-supervised consensus clustering for gene expression data analysis (PDF, 1.0 MiB)
|
---|
DOI | Resolve DOI: https://doi.org/10.1186/1756-0381-7-7 |
---|
Author | Search for: Wang, Yunli1; Search for: Pan, Youlian1 |
---|
Affiliation | - National Research Council of Canada. Information and Communication Technologies
|
---|
Format | Text, Article |
---|
Subject | Semi-supervised clustering; Consensus clustering; Semi-supervised consensus clustering; Gene expression |
---|
Abstract | Background: Simple clustering methods such as hierarchical clustering and k-means are widely used for gene expression data analysis; but they are unable to deal with noise and high dimensionality associated with the microarray gene expression data. Consensus clustering appears to improve the robustness and quality of clustering results. Incorporating prior knowledge in clustering process (semi-supervised clustering) has been shown to improve the consistency between the data partitioning and domain knowledge. Methods. We proposed semi-supervised consensus clustering (SSCC) to integrate the consensus clustering with semi-supervised clustering for analyzing gene expression data. We investigated the roles of consensus clustering and prior knowledge in improving the quality of clustering. SSCC was compared with one semi-supervised clustering algorithm, one consensus clustering algorithm, and k-means. Experiments on eight gene expression datasets were performed using h-fold cross-validation. Results: Using prior knowledge improved the clustering quality by reducing the impact of noise and high dimensionality in microarray data. Integration of consensus clustering with semi-supervised clustering improved performance as compared to using consensus clustering or semi-supervised clustering separately. Our SSCC method outperformed the others tested in this paper. |
---|
Publication date | 2014-05-08 |
---|
In | |
---|
Language | English |
---|
Peer reviewed | Yes |
---|
NPARC number | 21272892 |
---|
Export citation | Export as RIS |
---|
Report a correction | Report a correction (opens in a new tab) |
---|
Record identifier | 4944bde4-ab82-453a-8eed-4442ede4f74d |
---|
Record created | 2014-12-03 |
---|
Record modified | 2020-06-04 |
---|