Conserved elements with potential to form polymorphic G-quadruplex structures in the first intron of human genes
نویسندگان
چکیده
To understand how potential for G-quadruplex formation might influence regulation of gene expression, we examined the 2 kb spanning the transcription start sites (TSS) of the 18 217 human RefSeq genes, distinguishing contributions of template and nontemplate strands. Regions both upstream and downstream of the TSS are G-rich, but the downstream region displays a clear bias toward G-richness on the nontemplate strand. Upstream of the TSS, much of the G-richness and potential for G-quadruplex formation derives from the presence of well-defined canonical regulatory motifs in duplex DNA, including CpG dinucleotides which are sites of regulatory methylation, and motifs recognized by the transcription factor SP1. This challenges the notion that quadruplex formation upstream of the TSS contributes to regulation of gene expression. Downstream of the TSS, G-richness is concentrated in the first intron, and on the nontemplate strand, where polymorphic sequence elements with potential to form G-quadruplex structures and which cannot be accounted for by known regulatory motifs are found in almost 3000 (16%) of the human RefSeq genes, and are conserved through frogs. These elements could in principle be recognized either as DNA or as RNA, providing structural targets for regulation at the level of transcription or RNA processing.
منابع مشابه
In silico screening of G-Quadruplex Structures in Wilms tumor 1 Gene Promoter
Introduction: X-ray diffraction studies have revealed that guanines in a DNA stands may be arranged in quartet and form a structure called G-quadruplexs. Bioinformatics studies suggested the formation of G-quadruplex structure in human crucial genes, including Wilms tumor 1 (WT1). The aim of this study was to in silico analysis of the guanine-rich sequence in the promoter region of the WT1 gene...
متن کاملThe disruptive positions in human G-quadruplex motifs are less polymorphic and more conserved than their neutral counterparts
Specific guanine-rich sequence motifs in the human genome have considerable potential to form four-stranded structures known as G-quadruplexes or G4 DNA. The enrichment of these motifs in key chromosomal regions has suggested a functional role for the G-quadruplex structure in genomic regulation. In this work, we have examined the spectrum of nucleotide substitutions in G4 motifs, and related t...
متن کاملP-157: Polymorphic Core Promoter GA-repeats Alter Gene Expression of The Early Embryonic Developmental Genes
Background: We examine the GA-repeat core promoters of MECOM and GABRA3 in human embryonic kidney-293 cell line and show that those GA-repeats have promoter activity,and those different alleles of the repeats can significantly alter gene expression.We propose a novel role for GA-repeat core promoters to regulate gene expression in the genes involved in development and evolution. Materials and M...
متن کاملNovel Single Nucleotide Polymorphisms (SNPs) in Intron 2 and Exon 3 Regions of Leptin Gene in Sumba Ongole Cattle
The bovine leptin (LEP) gene was widely used as a candidate gene for molecular selection to improve productivity traits of cattle. This study was carried out to identify single nucleotide polymorphisms (SNPs) in the LEP gene of Sumba Ongole (SO, Bos indicus) cows using sequencing method. A total of 31 animals were used in this study for analyses. Research showed that total of 16 SNPs w...
متن کاملFormation of a Unique Cluster of G-Quadruplex Structures in the HIV-1 nef Coding Region: Implications for Antiviral Activity
G-quadruplexes are tetraplex structures of nucleic acids that can form in G-rich sequences. Their presence and functional role have been established in telomeres, oncogene promoters and coding regions of the human chromosome. In particular, they have been proposed to be directly involved in gene regulation at the level of transcription. Because the HIV-1 Nef protein is a fundamental factor for ...
متن کامل