WiggleTools: parallel processing of large collections of genome-wide datasets for visualization and statistical analysis

Bioinformatics. 2014 Apr 1;30(7):1008-9. doi: 10.1093/bioinformatics/btt737. Epub 2013 Dec 19.

Abstract

Motivation: Using high-throughput sequencing, researchers are now generating hundreds of whole-genome assays to measure various features such as transcription factor binding, histone marks, DNA methylation or RNA transcription. Displaying so much data generally leads to a confusing accumulation of plots. We describe here a multithreaded library that computes statistics on large numbers of datasets (Wiggle, BigWig, Bed, BigBed and BAM), generating statistical summaries within minutes with limited memory requirements, whether on the whole genome or on selected regions.

Availability and implementation: The code is freely available under Apache 2.0 license at www.github.com/Ensembl/Wiggletools

Publication types

  • Research Support, Non-U.S. Gov't

MeSH terms

  • Genome*
  • Genomic Library
  • Genomics / methods*
  • High-Throughput Nucleotide Sequencing / methods*
  • Internet
  • Software