File Download
Links for fulltext
(May Require Subscription)
- Publisher Website: 10.2202/1544-6115.1426
- Scopus: eid_2-s2.0-62449263773
- PMID: 19222380
- WOS: WOS:000263440600016
- Find via

Supplementary
- Citations:
- Appears in Collections:
Article: Detecting outlier samples in microarray data
| Title | Detecting outlier samples in microarray data |
|---|---|
| Authors | |
| Issue Date | 2009 |
| Publisher | Berkeley Electronic Press. The Journal's web site is located at http://www.bepress.com/sagmb |
| Citation | Statistical Applications in Genetics and Molecular Biology, 2009, v. 8 n. 1, article no. 13 How to Cite? |
| Abstract | In this paper, we address the problem of detecting outlier samples with highly different expression patterns in microarray data. Although outliers are not common, they appear even in widely used benchmark data sets and can negatively affect microarray data analysis. It is important to identify outliers in order to explore underlying experimental or biological problems and remove erroneous data. We propose an outlier detection method based on principal component analysis (PCA) and robust estimation of Mahalanobis distances that is fully automatic. We demonstrate that our outlier detection method identifies biologically significant outliers with high accuracy and that outlier removal improves the prediction accuracy of classifiers. Our outlier detection method is closely related to existing robust PCA methods, so we compare our outlier detection method to a prominent robust PCA method. |
| Persistent Identifier | http://hdl.handle.net/10722/58817 |
| ISSN | 2023 Impact Factor: 0.8 2023 SCImago Journal Rankings: 0.201 |
| ISI Accession Number ID | |
| References |
| DC Field | Value | Language |
|---|---|---|
| dc.contributor.author | Shieh, AD | en_HK |
| dc.contributor.author | Hung, YS | en_HK |
| dc.date.accessioned | 2010-05-31T03:37:25Z | - |
| dc.date.available | 2010-05-31T03:37:25Z | - |
| dc.date.issued | 2009 | en_HK |
| dc.identifier.citation | Statistical Applications in Genetics and Molecular Biology, 2009, v. 8 n. 1, article no. 13 | en_HK |
| dc.identifier.issn | 1544-6115 | en_HK |
| dc.identifier.uri | http://hdl.handle.net/10722/58817 | - |
| dc.description.abstract | In this paper, we address the problem of detecting outlier samples with highly different expression patterns in microarray data. Although outliers are not common, they appear even in widely used benchmark data sets and can negatively affect microarray data analysis. It is important to identify outliers in order to explore underlying experimental or biological problems and remove erroneous data. We propose an outlier detection method based on principal component analysis (PCA) and robust estimation of Mahalanobis distances that is fully automatic. We demonstrate that our outlier detection method identifies biologically significant outliers with high accuracy and that outlier removal improves the prediction accuracy of classifiers. Our outlier detection method is closely related to existing robust PCA methods, so we compare our outlier detection method to a prominent robust PCA method. | en_HK |
| dc.language | eng | en_HK |
| dc.publisher | Berkeley Electronic Press. The Journal's web site is located at http://www.bepress.com/sagmb | en_HK |
| dc.relation.ispartof | Statistical Applications in Genetics and Molecular Biology | en_HK |
| dc.rights | Copyright © 2009 The Berkeley Electronic Press. All rights reserved. The final publication is available at www.degruyter.com | en_HK |
| dc.subject.mesh | Colonic Neoplasms - diagnosis - genetics | en_HK |
| dc.subject.mesh | Databases, Genetic | en_HK |
| dc.subject.mesh | Humans | en_HK |
| dc.subject.mesh | Oligonucleotide Array Sequence Analysis - statistics & numerical data | en_HK |
| dc.subject.mesh | Outliers, DRG - statistics & numerical data | en_HK |
| dc.subject.mesh | Principal Component Analysis | en_HK |
| dc.title | Detecting outlier samples in microarray data | en_HK |
| dc.type | Article | en_HK |
| dc.identifier.email | Hung, YS:yshung@eee.hku.hk | en_HK |
| dc.identifier.authority | Hung, YS=rp00220 | en_HK |
| dc.description.nature | published_or_final_version | - |
| dc.identifier.doi | 10.2202/1544-6115.1426 | en_HK |
| dc.identifier.pmid | 19222380 | en_HK |
| dc.identifier.scopus | eid_2-s2.0-62449263773 | en_HK |
| dc.identifier.hkuros | 163899 | en_HK |
| dc.relation.references | http://www.scopus.com/mlt/select.url?eid=2-s2.0-62449263773&selection=ref&src=s&origin=recordpage | en_HK |
| dc.identifier.volume | 8 | en_HK |
| dc.identifier.issue | 1 | en_HK |
| dc.identifier.spage | article no. 13 | - |
| dc.identifier.epage | article no. 13 | - |
| dc.identifier.eissn | 1544-6115 | - |
| dc.identifier.isi | WOS:000263440600016 | - |
| dc.publisher.place | United States | en_HK |
| dc.identifier.scopusauthorid | Shieh, AD=9237952100 | en_HK |
| dc.identifier.scopusauthorid | Hung, YS=8091656200 | en_HK |
| dc.identifier.citeulike | 6850077 | - |
| dc.identifier.issnl | 1544-6115 | - |
