A collateral missing value estimation algorithm for DNA microarrays

Sehgal, M. S. B.; Gondal, I. and Dooley, L. (2005). A collateral missing value estimation algorithm for DNA microarrays. In: IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP '05), 18-23 Mar 2005, Philadelphia.

DOI: https://doi.org/10.1109/ICASSP.2005.1416319


Genetic microarray expression data often contains multiple missing values that can significantly affect the performance of statistical and machine learning algorithms. This paper presents an innovative missing value estimation technique, called collateral missing value estimation (CMVE) which has demonstrated superior estimation performance compared with the K-nearest neighbour (KNN) imputation algorithm, the least square impute (LSImpute) and Bayesian principal component analysis (BPCA) techniques. Experimental results confirm that CMVE provides an improvement of 89%, 12% and 10% for the BRCA1, BRCA2 and sporadic ovarian cancer mutations, respectively, compared to the average error rate of KNN, LSImpute and BPCA imputation methods, over a range of randomly selected missing values. The underlying theory behind CMVE also means that it is not restricted to bioinformatics data, but can be successfully applied to any correlated data set.

Viewing alternatives

Download history


Public Attention

Altmetrics from Altmetric

Number of Citations

Citations from Dimensions

Item Actions