MethylAid: visual and interactive quality control of large Illumina 450k datasets

作者:van Iterson Maarten*; Tobi Elmar W; Slieker Roderick C; den Hollander Wouter; Luijk Rene; Slagboom P Eline; Heijmans Bastiaan T
来源:Bioinformatics, 2014, 30(23): 3435-3437.
DOI:10.1093/bioinformatics/btu566

摘要

ASummary: The Illumina 450k array is a frequently used platform for largescale genome-wide DNA methylation studies, i.e. epigenome-wide association studies. Currently, quality control of 450k data can be performed with Illumina%26apos;s GenomeStudio and is part of a limited number 450k analysis pipelines. However, GenomeStudio cannot handle largescale studies, and existing pipelines provide limited options for quality control and neither support interactive exploration by the user. To aid the detection of bad-quality samples in large-scale genome-wide DNA methylation studies as flexible and transparent as possible, we have developed MethylAid; a visual and interactiveWeb application using RStudio%26apos;s shiny package. Bad-quality samples are detected using sample-dependent and sample-independent quality control probes present on the array and user-adjustable thresholds. In-depth exploration of bad-quality samples can be performed using several interactive diagnostic plots. Furthermore, plots can be annotated with user-provided meta data, for example, to identify outlying batches. Our new tool makes quality assessment of 450k array data interactive, flexible and efficient and is, therefore, expected to be useful for both data analysts and core facilities.

  • 出版日期2014-12-1