[HTML][HTML] Empirical array quality weights in the analysis of microarray data

ME Ritchie, D Diyagama, J Neilson, R van Laar… - BMC …, 2006 - Springer
ME Ritchie, D Diyagama, J Neilson, R van Laar, A Dobrovic, A Holloway, GK Smyth
BMC bioinformatics, 2006Springer
Background Assessment of array quality is an essential step in the analysis of data from
microarray experiments. Once detected, less reliable arrays are typically excluded or"
filtered" from further analysis to avoid misleading results. Results In this article, a graduated
approach to array quality is considered based on empirical reproducibility of the gene
expression measures from replicate arrays. Weights are assigned to each microarray by
fitting a heteroscedastic linear model with shared array variance terms. A novel gene-by …
Background
Assessment of array quality is an essential step in the analysis of data from microarray experiments. Once detected, less reliable arrays are typically excluded or "filtered" from further analysis to avoid misleading results.
Results
In this article, a graduated approach to array quality is considered based on empirical reproducibility of the gene expression measures from replicate arrays. Weights are assigned to each microarray by fitting a heteroscedastic linear model with shared array variance terms. A novel gene-by-gene update algorithm is used to efficiently estimate the array variances. The inverse variances are used as weights in the linear model analysis to identify differentially expressed genes. The method successfully assigns lower weights to less reproducible arrays from different experiments. Down-weighting the observations from suspect arrays increases the power to detect differential expression. In smaller experiments, this approach outperforms the usual method of filtering the data. The method is available in the limma software package which is implemented in the R software environment.
Conclusion
This method complements existing normalisation and spot quality procedures, and allows poorer quality arrays, which would otherwise be discarded, to be included in an analysis. It is applicable to microarray data from experiments with some level of replication.
Springer