Handling batch effects on cross-platform classification of microarray data

Journal article


Authors/Editors


Strategic Research Themes

No matching items found.


Publication Details

Author listEngchuan W., Meechai A., Tongsima S., Chan J.H.

PublisherInderscience

Publication year2016

JournalInternational Journal of Advanced Intelligence Paradigms (1755-0386)

Volume number8

Issue number1

Start page59

End page76

Number of pages18

ISSN1755-0386

eISSN1755-0394

URLhttps://www.scopus.com/inward/record.uri?eid=2-s2.0-84959375326&doi=10.1504%2fIJAIP.2016.074775&partnerID=40&md5=a96e0705e76814bf05dbbe0c60bf92eb

LanguagesEnglish-Great Britain (EN-GB)


View on publisher site


Abstract

Gene-set-based microarray analysis is commonly applied in the classification of complex diseases. However, the robustness of a classifier is normally limited by the small number of samples in many microarray datasets. Although a merged dataset from multiple experiments may improve classification performance, batch effects or technical/biological variations among these experiments may eventually confound the analysis. Besides the batch effects, merging multiple microarray datasets from different platforms can generate missing values, due to a different number of covered genes. In this work, we extend previous works that focused on the missing value incident by further exploring the impact of batch effects on cross-platform classification. Two quality measures of data purity are proposed and two data imputation methods are compared. The results show that by doing batch correction the quality of the merged data is improved significantly. Furthermore, the classification performance is high when the normalised purity is above a certain threshold. Copyright ฉ 2016 Inderscience Enterprises Ltd.


Keywords

Batch effects correctionBPCAComBatComplex diseasesData imputationData purityGene set-based approachkNN


Last updated on 2023-04-10 at 07:36