Beta approximation of ratio distribution and its application to next generation sequencing read counts

Paired sequencing data are commonly collected in genomic studies to control biological variation. However, existing data processing strategies suffer at low coverage regions, which are unavoidable due to the limitation of current sequencing technology. Furthermore, information contained in the absol...

Ausführliche Beschreibung

Bibliographische Detailangaben
Veröffentlicht in:Journal of applied statistics. - 1991. - 44(2017), 1 vom: 01., Seite 57-70
1. Verfasser: Yang, Shengping (VerfasserIn)
Weitere Verfasser: Fang, Zhide
Format: Online-Aufsatz
Sprache:English
Veröffentlicht: 2017
Zugriff auf das übergeordnete Werk:Journal of applied statistics
Schlagworte:Journal Article Beta distribution Next-generation sequencing Paired tumor and normal tissues Primary 62E17 Ratio of counts Statistical artifacts secondary 62P10
Beschreibung
Zusammenfassung:Paired sequencing data are commonly collected in genomic studies to control biological variation. However, existing data processing strategies suffer at low coverage regions, which are unavoidable due to the limitation of current sequencing technology. Furthermore, information contained in the absolute values of the read counts is commonly ignored. We propose a read count ratio processing/modification method, to not only incorporate information contained in the absolute values of paired counts into one variable, but also mitigate the discrete artifact, especially when both counts are small. Simulation shows that the processed variable fits well with a Beta distribution, thus providing an easy tool for down-stream inference analysis
Beschreibung:Date Revised 20.11.2019
published: Print-Electronic
Citation Status PubMed-not-MEDLINE
ISSN:0266-4763
DOI:10.1080/02664763.2016.1158798