High-dimensional genomic data bias correction and data integration using MANCIE
Author(s) -
Chongzhi Zang,
Tao Wang,
Ke Deng,
Bo Li,
Shengen Hu,
Qian Qin,
Tengfei Xiao,
Shihua Zhang,
Clifford A. Meyer,
Housheng Hansen He,
Myles Brown,
Jun S. Liu,
Yang Xie,
X. Shirley Liu
Publication year - 2016
Publication title -
nature communications
Language(s) - English
Resource type - Journals
SCImago Journal Rank - 5.559
H-Index - 365
ISSN - 2041-1723
DOI - 10.1038/ncomms11305
Subject(s) - computational biology , computer science , biology
High-dimensional genomic data analysis is challenging due to noises and biases in high-throughput experiments. We present a computational method matrix analysis and normalization by concordant information enhancement (MANCIE) for bias correction and data integration of distinct genomic profiles on the same samples. MANCIE uses a Bayesian-supported principal component analysis-based approach to adjust the data so as to achieve better consistency between sample-wise distances in the different profiles. MANCIE can improve tissue-specific clustering in ENCODE data, prognostic prediction in Molecular Taxonomy of Breast Cancer International Consortium and The Cancer Genome Atlas data, copy number and expression agreement in Cancer Cell Line Encyclopedia data, and has broad applications in cross-platform, high-dimensional data integration.
Accelerating Research
Robert Robinson Avenue,
Oxford Science Park, Oxford
OX4 4GP, United Kingdom
Address
John Eccles HouseRobert Robinson Avenue,
Oxford Science Park, Oxford
OX4 4GP, United Kingdom