Cargando…

The over-representation of binary DNA tracts in seven sequenced chromosomes

BACKGROUND: DNA tracts composed of only two bases are possible in six combinations: A+G (purines, R), C+T (pyrimidines, Y), G+T (Keto, K), A+C (Imino, M), A+T (Weak, W) and G+C (Strong, S). It is long known that all-pyrimidine tracts, complemented by all-purines tracts ("R.Y tracts"), are...

Descripción completa

Detalles Bibliográficos
Autor principal: Yagil, Gad
Formato: Texto
Lenguaje:English
Publicado: BioMed Central 2004
Materias:
Acceso en línea:https://www.ncbi.nlm.nih.gov/pmc/articles/PMC407849/
https://www.ncbi.nlm.nih.gov/pubmed/15113401
http://dx.doi.org/10.1186/1471-2164-5-19
_version_ 1782121394790727680
author Yagil, Gad
author_facet Yagil, Gad
author_sort Yagil, Gad
collection PubMed
description BACKGROUND: DNA tracts composed of only two bases are possible in six combinations: A+G (purines, R), C+T (pyrimidines, Y), G+T (Keto, K), A+C (Imino, M), A+T (Weak, W) and G+C (Strong, S). It is long known that all-pyrimidine tracts, complemented by all-purines tracts ("R.Y tracts"), are excessively present in analyzed DNA. We have previously shown that R.Y tracts are in vast excess in yeast promoters, and brought evidence for their role in gene regulation. Here we report the systematic mapping of all six binary combinations on the level of complete sequenced chromosomes, as well as in their different subregions. RESULTS: DNA tracts composed of the above binary base combinations have been mapped in seven sequenced chromosomes: Human chromosomes 21 and 22 (the major contigs); Drosophila melanogaster chr. 2R; Caenorhabditis elegans chr. I; Arabidopsis thaliana chr. II; Saccharomyces cerevisiae chr. IV and M. jannaschii. A huge over-representation, reaching million-folds, has been found for very long tracts of all binary motifs except S, in each of the seven organisms. Long R.Y tracts are the most excessive, except in D. melanogaster, where the K.M motif predominates. S (G, C rich) tracts are in excess mainly in CpG islands; the W motif predominates in bacteria. Many excessively long W tracts are nevertheless found also in the archeon and in the eukaryotes. The survey of complete chromosomes enables us, for the first time, to map systematically the intergenic regions. In human and other chromosomes we find the highest over-representation of the binary DNA tracts in the intergenic regions. These over-representations are only partly explainable by the presence of interspersed elements. CONCLUSIONS: The over-representation of long DNA tracts composed of five of the above motifs is the largest deviation from randomness so far established for DNA, and this in a wide range of eukaryotic and archeal chromosomes. A propensity for ready DNA unwinding is proposed as the functional role, explaining the evolutionary conservation of the huge excesses observed.
format Text
id pubmed-407849
institution National Center for Biotechnology Information
language English
publishDate 2004
publisher BioMed Central
record_format MEDLINE/PubMed
spelling pubmed-4078492004-05-15 The over-representation of binary DNA tracts in seven sequenced chromosomes Yagil, Gad BMC Genomics Research Article BACKGROUND: DNA tracts composed of only two bases are possible in six combinations: A+G (purines, R), C+T (pyrimidines, Y), G+T (Keto, K), A+C (Imino, M), A+T (Weak, W) and G+C (Strong, S). It is long known that all-pyrimidine tracts, complemented by all-purines tracts ("R.Y tracts"), are excessively present in analyzed DNA. We have previously shown that R.Y tracts are in vast excess in yeast promoters, and brought evidence for their role in gene regulation. Here we report the systematic mapping of all six binary combinations on the level of complete sequenced chromosomes, as well as in their different subregions. RESULTS: DNA tracts composed of the above binary base combinations have been mapped in seven sequenced chromosomes: Human chromosomes 21 and 22 (the major contigs); Drosophila melanogaster chr. 2R; Caenorhabditis elegans chr. I; Arabidopsis thaliana chr. II; Saccharomyces cerevisiae chr. IV and M. jannaschii. A huge over-representation, reaching million-folds, has been found for very long tracts of all binary motifs except S, in each of the seven organisms. Long R.Y tracts are the most excessive, except in D. melanogaster, where the K.M motif predominates. S (G, C rich) tracts are in excess mainly in CpG islands; the W motif predominates in bacteria. Many excessively long W tracts are nevertheless found also in the archeon and in the eukaryotes. The survey of complete chromosomes enables us, for the first time, to map systematically the intergenic regions. In human and other chromosomes we find the highest over-representation of the binary DNA tracts in the intergenic regions. These over-representations are only partly explainable by the presence of interspersed elements. CONCLUSIONS: The over-representation of long DNA tracts composed of five of the above motifs is the largest deviation from randomness so far established for DNA, and this in a wide range of eukaryotic and archeal chromosomes. A propensity for ready DNA unwinding is proposed as the functional role, explaining the evolutionary conservation of the huge excesses observed. BioMed Central 2004-03-03 /pmc/articles/PMC407849/ /pubmed/15113401 http://dx.doi.org/10.1186/1471-2164-5-19 Text en Copyright © 2004 Yagil; licensee BioMed Central Ltd. This is an Open Access article: verbatim copying and redistribution of this article are permitted in all media for any purpose, provided this notice is preserved along with the article's original URL.
spellingShingle Research Article
Yagil, Gad
The over-representation of binary DNA tracts in seven sequenced chromosomes
title The over-representation of binary DNA tracts in seven sequenced chromosomes
title_full The over-representation of binary DNA tracts in seven sequenced chromosomes
title_fullStr The over-representation of binary DNA tracts in seven sequenced chromosomes
title_full_unstemmed The over-representation of binary DNA tracts in seven sequenced chromosomes
title_short The over-representation of binary DNA tracts in seven sequenced chromosomes
title_sort over-representation of binary dna tracts in seven sequenced chromosomes
topic Research Article
url https://www.ncbi.nlm.nih.gov/pmc/articles/PMC407849/
https://www.ncbi.nlm.nih.gov/pubmed/15113401
http://dx.doi.org/10.1186/1471-2164-5-19
work_keys_str_mv AT yagilgad theoverrepresentationofbinarydnatractsinsevensequencedchromosomes
AT yagilgad overrepresentationofbinarydnatractsinsevensequencedchromosomes