Cargando…

Defining data-driven primary transcript annotations with primaryTranscriptAnnotation in R

SUMMARY: Nascent transcript measurements derived from run-on sequencing experiments are critical for the investigation of transcriptional mechanisms and regulatory networks. However, conventional mRNA gene annotations significantly differ from the boundaries of primary transcripts. New primary trans...

Descripción completa

Detalles Bibliográficos
Autores principales: Anderson, Warren D, Duarte, Fabiana M, Civelek, Mete, Guertin, Michael J
Formato: Online Artículo Texto
Lenguaje:English
Publicado: Oxford University Press 2020
Materias:
Acceso en línea:https://www.ncbi.nlm.nih.gov/pmc/articles/PMC7203734/
https://www.ncbi.nlm.nih.gov/pubmed/31917388
http://dx.doi.org/10.1093/bioinformatics/btaa011
Descripción
Sumario:SUMMARY: Nascent transcript measurements derived from run-on sequencing experiments are critical for the investigation of transcriptional mechanisms and regulatory networks. However, conventional mRNA gene annotations significantly differ from the boundaries of primary transcripts. New primary transcript annotations are needed to accurately interpret run-on data. We developed the primaryTranscriptAnnotation R package to infer the transcriptional start and termination sites of primary transcripts from genomic run-on data. We then used these inferred coordinates to annotate transcriptional units identified de novo. This package provides the novel utility to integrate data-driven primary transcript annotations with transcriptional unit coordinates identified in an unbiased manner. Highlighting the importance of using accurate primary transcript coordinates, we demonstrate that this new methodology increases the detection of differentially expressed transcripts and provides more accurate quantification of RNA polymerase pause indices. AVAILABILITY AND IMPLEMENTATION: https://github.com/WarrenDavidAnderson/genomicsRpackage/tree/master/primaryTranscriptAnnotation. SUPPLEMENTARY INFORMATION: Supplementary data are available at Bioinformatics online.