Cargando…
The over-representation of binary DNA tracts in seven sequenced chromosomes
BACKGROUND: DNA tracts composed of only two bases are possible in six combinations: A+G (purines, R), C+T (pyrimidines, Y), G+T (Keto, K), A+C (Imino, M), A+T (Weak, W) and G+C (Strong, S). It is long known that all-pyrimidine tracts, complemented by all-purines tracts ("R.Y tracts"), are...
Autor principal: | |
---|---|
Formato: | Texto |
Lenguaje: | English |
Publicado: |
BioMed Central
2004
|
Materias: | |
Acceso en línea: | https://www.ncbi.nlm.nih.gov/pmc/articles/PMC407849/ https://www.ncbi.nlm.nih.gov/pubmed/15113401 http://dx.doi.org/10.1186/1471-2164-5-19 |
_version_ | 1782121394790727680 |
---|---|
author | Yagil, Gad |
author_facet | Yagil, Gad |
author_sort | Yagil, Gad |
collection | PubMed |
description | BACKGROUND: DNA tracts composed of only two bases are possible in six combinations: A+G (purines, R), C+T (pyrimidines, Y), G+T (Keto, K), A+C (Imino, M), A+T (Weak, W) and G+C (Strong, S). It is long known that all-pyrimidine tracts, complemented by all-purines tracts ("R.Y tracts"), are excessively present in analyzed DNA. We have previously shown that R.Y tracts are in vast excess in yeast promoters, and brought evidence for their role in gene regulation. Here we report the systematic mapping of all six binary combinations on the level of complete sequenced chromosomes, as well as in their different subregions. RESULTS: DNA tracts composed of the above binary base combinations have been mapped in seven sequenced chromosomes: Human chromosomes 21 and 22 (the major contigs); Drosophila melanogaster chr. 2R; Caenorhabditis elegans chr. I; Arabidopsis thaliana chr. II; Saccharomyces cerevisiae chr. IV and M. jannaschii. A huge over-representation, reaching million-folds, has been found for very long tracts of all binary motifs except S, in each of the seven organisms. Long R.Y tracts are the most excessive, except in D. melanogaster, where the K.M motif predominates. S (G, C rich) tracts are in excess mainly in CpG islands; the W motif predominates in bacteria. Many excessively long W tracts are nevertheless found also in the archeon and in the eukaryotes. The survey of complete chromosomes enables us, for the first time, to map systematically the intergenic regions. In human and other chromosomes we find the highest over-representation of the binary DNA tracts in the intergenic regions. These over-representations are only partly explainable by the presence of interspersed elements. CONCLUSIONS: The over-representation of long DNA tracts composed of five of the above motifs is the largest deviation from randomness so far established for DNA, and this in a wide range of eukaryotic and archeal chromosomes. A propensity for ready DNA unwinding is proposed as the functional role, explaining the evolutionary conservation of the huge excesses observed. |
format | Text |
id | pubmed-407849 |
institution | National Center for Biotechnology Information |
language | English |
publishDate | 2004 |
publisher | BioMed Central |
record_format | MEDLINE/PubMed |
spelling | pubmed-4078492004-05-15 The over-representation of binary DNA tracts in seven sequenced chromosomes Yagil, Gad BMC Genomics Research Article BACKGROUND: DNA tracts composed of only two bases are possible in six combinations: A+G (purines, R), C+T (pyrimidines, Y), G+T (Keto, K), A+C (Imino, M), A+T (Weak, W) and G+C (Strong, S). It is long known that all-pyrimidine tracts, complemented by all-purines tracts ("R.Y tracts"), are excessively present in analyzed DNA. We have previously shown that R.Y tracts are in vast excess in yeast promoters, and brought evidence for their role in gene regulation. Here we report the systematic mapping of all six binary combinations on the level of complete sequenced chromosomes, as well as in their different subregions. RESULTS: DNA tracts composed of the above binary base combinations have been mapped in seven sequenced chromosomes: Human chromosomes 21 and 22 (the major contigs); Drosophila melanogaster chr. 2R; Caenorhabditis elegans chr. I; Arabidopsis thaliana chr. II; Saccharomyces cerevisiae chr. IV and M. jannaschii. A huge over-representation, reaching million-folds, has been found for very long tracts of all binary motifs except S, in each of the seven organisms. Long R.Y tracts are the most excessive, except in D. melanogaster, where the K.M motif predominates. S (G, C rich) tracts are in excess mainly in CpG islands; the W motif predominates in bacteria. Many excessively long W tracts are nevertheless found also in the archeon and in the eukaryotes. The survey of complete chromosomes enables us, for the first time, to map systematically the intergenic regions. In human and other chromosomes we find the highest over-representation of the binary DNA tracts in the intergenic regions. These over-representations are only partly explainable by the presence of interspersed elements. CONCLUSIONS: The over-representation of long DNA tracts composed of five of the above motifs is the largest deviation from randomness so far established for DNA, and this in a wide range of eukaryotic and archeal chromosomes. A propensity for ready DNA unwinding is proposed as the functional role, explaining the evolutionary conservation of the huge excesses observed. BioMed Central 2004-03-03 /pmc/articles/PMC407849/ /pubmed/15113401 http://dx.doi.org/10.1186/1471-2164-5-19 Text en Copyright © 2004 Yagil; licensee BioMed Central Ltd. This is an Open Access article: verbatim copying and redistribution of this article are permitted in all media for any purpose, provided this notice is preserved along with the article's original URL. |
spellingShingle | Research Article Yagil, Gad The over-representation of binary DNA tracts in seven sequenced chromosomes |
title | The over-representation of binary DNA tracts in seven sequenced chromosomes |
title_full | The over-representation of binary DNA tracts in seven sequenced chromosomes |
title_fullStr | The over-representation of binary DNA tracts in seven sequenced chromosomes |
title_full_unstemmed | The over-representation of binary DNA tracts in seven sequenced chromosomes |
title_short | The over-representation of binary DNA tracts in seven sequenced chromosomes |
title_sort | over-representation of binary dna tracts in seven sequenced chromosomes |
topic | Research Article |
url | https://www.ncbi.nlm.nih.gov/pmc/articles/PMC407849/ https://www.ncbi.nlm.nih.gov/pubmed/15113401 http://dx.doi.org/10.1186/1471-2164-5-19 |
work_keys_str_mv | AT yagilgad theoverrepresentationofbinarydnatractsinsevensequencedchromosomes AT yagilgad overrepresentationofbinarydnatractsinsevensequencedchromosomes |