Statistics, Department of

Department of Statistics: Faculty Publications

Systematic evaluation of the impact of ChIP-seq read designs on genome coverage, peak identification, and allele-specific binding detection

Qi Zhang, University of Nebraska-LincolnFollow
Xin Zeng, University of Wisconsin-Madison
Sam Younkin, University of Wisconsin Madison, Madison
Trupti Kawli, Stanford University School of Medicine, Palo Alto
Michael P. Snyder, Stanford University School of Medicine, Palo Alto
Sündüz Kele, University of Wisconsin Madison, MadisonFollow

Document Type

Article

Date of this Version

2016

Citation

Zhang et al. BMC Bioinformatics (2016) 17:96 DOI 10.1186/s12859-016-0957-1

Comments

Abstract

Background: Chromatin immunoprecipitation followed by sequencing (ChIP-seq) experiments revolutionized genome-wide profiling of transcription factors and histone modifications. Although maturing sequencing technologies allow these experiments to be carried out with short (36–50 bps), long (75–100 bps), single-end, or paired-end reads, the impact of these read parameters on the downstream data analysis are not well understood. In this paper, we evaluate the effects of different read parameters on genome sequence alignment, coverage of different classes of genomic features, peak identification, and allele-specific binding detection.

Results: We generated 101 bps paired-end ChIP-seq data for many transcription factors from human GM12878 and MCF7 cell lines. Systematic evaluations using in silico variations of these data as well as fully simulated data, revealed complex interplay between the sequencing parameters and analysis tools, and indicated clear advantages of paired-end designs in several aspects such as alignment accuracy, peak resolution, and most notably, allele-specific binding detection.

Conclusions: Our work elucidates the effect of design on the downstream analysis and provides insights to investigators in deciding sequencing parameters in ChIP-seq experiments. We present the first systematic evaluation of the impact of ChIP-seq designs on allele-specific binding detection and highlights the power of pair-end designs in such studies.

Download

Included in

Other Statistics and Probability Commons

COinS

DigitalCommons@University of Nebraska - Lincoln

Statistics, Department of

Department of Statistics: Faculty Publications

Systematic evaluation of the impact of ChIP-seq read designs on genome coverage, peak identification, and allele-specific binding detection

Document Type

Date of this Version

Citation

Comments

Abstract

Included in

Search

Browse

Author Corner

Links

DigitalCommons@University of Nebraska - Lincoln

Statistics, Department of

Department of Statistics: Faculty Publications

Systematic evaluation of the impact of ChIP-seq read designs on genome coverage, peak identification, and allele-specific binding detection

Authors

Document Type

Date of this Version

Citation

Comments

Abstract

Included in

Share

Search

Browse

Author Corner

Links