945 resultados para DNA sequence representation
Resumo:
DNA sequence variation is currently a major source of data for studying human origins, evolution, and demographic history, and for detecting linkage association of complex diseases. In this dissertation, I investigated DNA variation in worldwide populations from two ∼10 kb autosomal regions on 22q11.2 (noncoding) and 1q24 (introns). A total of 75 variant sites were found among 128 human sequences in the 22q11.2 region, yielding an estimate of 0.088% for nucleotide diversity (π), and a total of 52 variant sites were found among 122 human sequences in the 1q24 region with an estimated π value of 0.057%. The data from these two regions and a 10 kb noncoding region on Xq13.3 all show a strong excess of low-frequency variants in comparison to that expected from an equilibrium population, indicating a relatively recent population expansion. The effective population sizes estimated from the three regions were 11,000, 12,700, and 8,600, respectively, which are close to the commonly used value of 10,000. In each of the two autosomal regions, the age of the most recent common ancestor (MRCA) was estimated to be older than 1 million years among all the sequences and ∼600,000 years among non-African sequences, providing first evidence from autosomal noncoding or intronic regions for a genetic history of humans much more ancient than the emergence of modern humans. The ancient genetic history of humans indicates no severe bottleneck during the evolution of humans in the last half million years; otherwise, much of the ancient genetic history would have been lost during a severe bottleneck. This study strongly suggests that both the “out of Africa” and the multiregional models are too simple for explaining the evolution of modern humans. A compilation of genome-wide data revealed that nucleotide diversity is highest in autosomal regions, intermediate in X-linked regions, and lowest in Y-linked regions. The data suggest the existence of background selection or selective sweep on Y-linked loci. In general, the nucleotide diversity in humans is low compared to that in chimpanzee and Drosophila populations. ^
Resumo:
Background: Zooplankton play an important role in our oceans, in biogeochemical cycling and providing a food source for commercially important fish larvae. However, difficulties in correctly identifying zooplankton hinder our understanding of their roles in marine ecosystem functioning, and can prevent detection of long term changes in their community structure. The advent of massively parallel Next Generation Sequencing technology allows DNA sequence data to be recovered directly from whole community samples. Here we assess the ability of such sequencing to quantify the richness and diversity of a mixed zooplankton assemblage from a productive monitoring site in the Western English Channel. Methodology/Principle Findings: Plankton WP2 replicate net hauls (200 µm) were taken at the Western Channel Observatory long-term monitoring station L4 in September 2010 and January 2011. These samples were analysed by microscopy and metagenetic analysis of the 18S nuclear small subunit ribosomal RNA gene using the 454 pyrosequencing platform. Following quality control a total of 419,042 sequences were obtained for all samples. The sequences clustered in to 205 operational taxonomic units using a 97% similarity cut-off. Allocation of taxonomy by comparison with the National Centre for Biotechnology Information database identified 138 OTUs to species level, 11 to genus level and 1 to order, <2.5% of sequences were classified as unknowns. By comparison a skilled microscopic analyst was able to routinely enumerate only 75 taxonomic groups. Conclusions: The percentage of OTUs assigned to major eukaryotic taxonomic groups broadly aligns between the metagenetic and morphological analysis and are dominated by Copepoda. However, the metagenetics reveals a previously hidden taxonomic richness, especially for Copepoda and meroplankton such as Bivalvia, Gastropoda and Polychaeta. It also reveals rare species and parasites. We conclude that Next Generation Sequencing of 18S amplicons is a powerful tool for estimating diversity and species richness of zooplankton communities.
Resumo:
In Azotobacter vinelandii, deletion of the fdxA gene that encodes a well characterized seven-iron ferredoxin (FdI) is known to lead to overexpression of the FdI redox partner, NADPH:ferredoxin reductase (FPR). Previous studies have established that this is an oxidative stress response in which the fpr gene is transcriptionally activated to the same extent in response to either addition of the superoxide propagator paraquat to the cells or to fdxA deletion. In both cases, the activation occurs through a specific DNA sequence located upstream of the fpr gene. Here, we report the identification of the A. vinelandii protein that binds specifically to the paraquat activatable fpr promoter region as the E1 subunit of the pyruvate dehydrogenase complex (PDHE1), a central enzyme in aerobic respiration. Sequence analysis shows that PDHE1, which was not previously suspected to be a DNA-binding protein, has a helix–turn–helix motif. The data presented here further show that FdI binds specifically to the DNA-bound PDHE1.
Resumo:
Although polyomavirus JC (JCV) is the proven pathogen of progressive multifocal leukoencephalopathy, the fatal demyelinating disease, this virus is ubiquitous as a usually harmless symbiote among human beings. JCV propagates in the adult kidney and excretes its progeny in urine, from which JCV DNA can readily be recovered. The main mode of transmission of JCV is from parents to children through long cohabitation. In this study, we collected a substantial number of urine samples from native inhabitants of 34 countries in Europe, Africa, and Asia. A 610-bp segment of JCV DNA was amplified from each urine sample, and its DNA sequence was determined. A worldwide phylogenetic tree subsequently constructed revealed the presence of nine subtypes including minor ones. Five subtypes (EU, Af2, B1, SC, and CY) occupied rather large territories that overlapped with each other at their boundaries. The entire Europe, northern Africa, and western Asia were the domain of EU, whereas the domain of Af2 included nearly all of Africa and southwestern Asia all the way to the northeastern edge of India. Partially overlapping domains in Asia were occupied by subtypes B1, SC, and CY. Of particular interest was the recovery of JCV subtypes in a pocket or pockets that were separated by great geographic distances from the main domains of those subtypes. Certain of these pockets can readily be explained by recent migrations of human populations carrying these subtypes. Overall, it appears that JCV genotyping promises to reveal previously unknown human migration routes: ancient as well as recent.
Resumo:
We have examined the effects on transcription initiation of promoter and enhancer strength and of the curvature of the DNA separating these entities on wild-type and mutated enhancer–promoter regions at the Escherichia coli σ54-dependent promoters glnAp2 and glnHp2 on supercoiled and linear DNA. Our results, together with previously reported observations by other investigators, show that the initiation of transcription on linear DNA requires a single intrinsic or induced bend in the DNA, as well as a promoter with high affinity for σ54-RNA polymerase, but on supercoiled DNA requires either such a bend or a high affinity promoter but not both. The examination of the DNA sequence of all nif gene activator- or nitrogen regulator I-σ54 promoters reveals that those lacking a binding site for the integration host factor have an intrinsic single bend in the DNA separating enhancer from promoter.
Molecular keys to speciation: DNA polymorphism and the control of genetic exchange in enterobacteria
Resumo:
Speciation involves the establishment of genetic barriers between closely related organisms. The extent of genetic recombination is a key determinant and a measure of genetic isolation. The results reported here reveal that genetic barriers can be established, eliminated, or modified by manipulating two systems which control genetic recombination, SOS and mismatch repair. The extent of genetic isolation between enterobacteria is a simple mathematical function of DNA sequence divergence. The function does not depend on hybrid DNA stability, but rather on the number of blocks of sequences identical in the two mating partners and sufficiently large to allow the initiation of recombination. Further, there is no obvious discontinuity in the function that could be used to define a level of divergence for distinguishing species.
Resumo:
We have used two monovalent phage display libraries containing variants of the Zif268 DNA-binding domain to obtain families of zinc fingers that bind to alterations in the last 4 bp of the DNA sequence of the Zif268 consensus operator, GCG TGGGCG. Affinity selection was performed by altering the Zif268 operator three base pairs at a time, and simultaneously selecting for sets of 16 related DNA sequences. In this way, only four experiments were required to select for all possible 64 combinations of DNA triplet sequences. The results show that (i) for high-affinity DNA binding in the range observed for the Zif268 wild-type complex (Kd = 0.5–5 nM), finger 1 specifically requires the arginine at the carboxy terminus of its recognition helix that forms a bidentate hydrogen-bond with the guanine base (G) in the crystal structure of Zif268 complexed to its DNA operator sequence GCG TGG GCG; (ii) when the guanine base (G) is replaced by A, C, or T, a lower-affinity family (Kd ⩾ 50 nM) can be detected that shows an overall tendency to bind G-rich DNA; (iii) the residues at position 2 on the finger 2 recognition helix do not appear to interact strongly with the complementary 5′ base in the finger 1 binding site; and (iv) unexpected substitutions at the amino terminus of finger 1 can occasionally result in specificity for the 3′ base in the finger 1 binding site. A DNA recognition directory was constructed for high-affinity zinc fingers that recognize all three bases in a DNA triplet for seven sequences of the type GNN. Similar approaches may be applied to other zinc fingers to broaden the scope of the directory.
Resumo:
The replication of damaged nucleotides that have escaped DNA repair leads to the formation of mutations caused by misincorporation opposite the lesion. In Escherichia coli, this process is under tight regulation of the SOS stress response and is carried out by DNA polymerase III in a process that involves also the RecA, UmuD′ and UmuC proteins. We have shown that DNA polymerase III holoenzyme is able to replicate, unassisted, through a synthetic abasic site in a gapped duplex plasmid. Here, we show that DNA polymerase III*, a subassembly of DNA polymerase III holoenzyme lacking the β subunit, is blocked very effectively by the synthetic abasic site in the same DNA substrate. Addition of the β subunit caused a dramatic increase of at least 28-fold in the ability of the polymerase to perform translesion replication, reaching 52% bypass in 5 min. When the ssDNA region in the gapped plasmid was extended from 22 nucleotides to 350 nucleotides, translesion replication still depended on the β subunit, but it was reduced by 80%. DNA sequence analysis of translesion replication products revealed mostly −1 frameshifts. This mutation type is changed to base substitution by the addition of UmuD′, UmuC, and RecA, as demonstrated in a reconstituted SOS translesion replication reaction. These results indicate that the β subunit sliding DNA clamp is the major determinant in the ability of DNA polymerase III holoenzyme to perform unassisted translesion replication and that this unassisted bypass produces primarily frameshifts.
Resumo:
The chromosomal DNA of the bacteria Streptomyces ambofaciens DSM40697 is an 8-Mb linear molecule that ends in terminal inverted repeats (TIRs) of 210 kb. The sequences of the TIRs are highly variable between the different linear replicons of Streptomyces (plasmids or chromosomes). Two spontaneous mutant strains harboring TIRs of 480 and 850 kb were isolated. The TIR polymorphism seen is a result of the deletion of one chromosomal end and its replacement by 480 or 850 kb of sequence identical to the end of the undeleted chromosomal arm. Analysis of the wild-type sequences involved in these rearrangements revealed that a recombination event took place between the two copies of a duplicated DNA sequence. Each copy was mapped to one chromosomal arm, outside of the TIR, and encoded a putative alternative sigma factor. The two ORFs, designated hasR and hasL, were found to be 99% similar at the nucleotide level. The sequence of the chimeric regions generated by the recombination showed that the chromosomal structure of the mutant strains resulted from homologous recombination events between the two copies. We suggest that this mechanism of chromosomal arm replacement contributes to the rapid evolutionary diversification of the sequences of the TIR in Streptomyces.
Resumo:
Nuclear matrix binding assays (NMBAs) define certain DNA sequences as matrix attachment regions (MARs), which often have cis-acting epigenetic regulatory functions. We used NMBAs to analyze the functionally important 15q11-q13 imprinting center (IC). We find that the IC is composed of an unusually high density of MARs, located in close proximity to the germ line elements that are proposed to direct imprint switching in this region. Moreover, we find that the organization of MARs is the same at the homologous mouse locus, despite extensive divergence of DNA sequence. MARs of this size are not usually associated with genes but rather with heterochromatin-forming areas of the genome. In contrast, the 15q11-q13 region contains multiple transcribed genes and is unusual for being subject to genomic imprinting, causing the maternal chromosome to be more transcriptionally silent, methylated, and late replicating than the paternal chromosome. We suggest that the extensive MAR sequences at the IC are organized as heterochromatin during oogenesis, an organization disrupted during spermatogenesis. Consistent with this model, multicolor fluorescence in situ hybridization to halo nuclei demonstrates a strong matrix association of the maternal IC, whereas the paternal IC is more decondensed, extending into the nuclear halo. This model also provides a mechanism for spreading of the imprinting signal, because heterochromatin at the IC on the maternal chromosome may exert a suppressive position effect in cis. We propose that the germ line elements at the 15q11-q13 IC mediate their effects through the candidate heterochromatin-forming DNA identified in this study.
Resumo:
The end of a telomeric DNA sequence isolated from a polytene chromosome of a hypotrichous ciliate folds back and hybridizes with downstream telomeric sequence to form a t loop that is stable in the absence of protein and DNA cross-linking. The single-stranded, telomeric DNA sequence at the end of a macronuclear molecule does not form a t loop but, instead, is complexed with a heterodimeric, telomere-binding protein. Thus, two mechanisms for capping the ends of DNA molecules are used in the same cell.
Resumo:
We have designed a p53 DNA binding domain that has virtually the same binding affinity for the gadd45 promoter as does wild-type protein but is considerably more stable. The design strategy was based on molecular evolution of the protein domain. Naturally occurring amino acid substitutions were identified by comparing the sequences of p53 homologues from 23 species, introducing them into wild-type human p53, and measuring the changes in stability. The most stable substitutions were combined in a multiple mutant. The advantage of this strategy is that, by substituting with naturally occurring residues, the function is likely to be unimpaired. All point mutants bind the consensus DNA sequence. The changes in stability ranged from +1.27 (less stable Q165K) to −1.49 (more stable N239Y) kcal mol−1, respectively. The changes in free energy of unfolding on mutation are additive. Of interest, the two most stable mutants (N239Y and N268D) have been known to act as suppressors and restored the activity of two of the most common tumorigenic mutants. Of the 20 single mutants, 10 are cancer-associated, though their frequency of occurrence is extremely low: A129D, Q165K, Q167E, and D148E are less stable and M133L, V203A and N239Y are more stable whereas the rest are neutral. The quadruple mutant (M133LV203AN239YN268D), which is stabilized by 2.65 kcal mol−1 and Tm raised by 5.6°C is of potential interest for trials in vivo.
Resumo:
Leishmania are parasites that survive within macrophages by mechanism(s) not entirely known. Depression of cellular immunity and diminished production of interleukin 1β (IL-1β) and tumor necrosis factor α are potential ways by which the parasite survives within macrophages. We examined the mechanism(s) by which lipophosphoglycan (LPG), a major glycolipid of Leishmania, perturbs cytokine gene expression. LPG treatment of THP-1 monocytes suppressed endotoxin induction of IL-1β steady-state mRNA by greater than 90%, while having no effect on the expression of a control gene. The addition of LPG 2 h before or 2 h after endotoxin challenge significantly suppressed steady-state IL-1β mRNA by 90% and 70%, respectively. LPG also inhibited tumor necrosis factor α and Staphylococcus induction of IL-1β gene expression. The inhibitory effect of LPG is agonist-specific because LPG did not suppress the induction of IL-1β mRNA by phorbol 12-myristate 13-acetate. A unique DNA sequence located within the −310 to −57 nucleotide region of the IL-1β promoter was found to mediate LPG’s inhibitory activity. The requirement for the −310 to −57 promoter gene sequence for LPG’s effect is demonstrated by the abrogation of LPG’s inhibitory activity by truncation or deletion of the −310 to −57 promoter gene sequence. Furthermore, the minimal IL-1β promoter (positions −310 to +15) mediated LPG’s inhibitory activity with dose and kinetic profiles that were similar to LPG’s suppression of steady-state IL-1β mRNA. These findings delineated a promoter gene sequence that responds to LPG to act as a “gene silencer,” a function, to our knowledge, not previously described. LPG’s inhibitory activity for several mediators of inflammation and the persistence of significant inhibitory activity 2 h after endotoxin challenge suggest that LPG has therapeutic potential and may be exploited for therapy of sepsis, acute respiratory distress syndrome, and autoimmune diseases.
Resumo:
We describe an adaptation of the rolling circle amplification (RCA) reporter system for the detection of protein Ags, termed “immunoRCA.” In immunoRCA, an oligonucleotide primer is covalently attached to an Ab; thus, in the presence of circular DNA, DNA polymerase, and nucleotides, amplification results in a long DNA molecule containing hundreds of copies of the circular DNA sequence that remain attached to the Ab and that can be detected in a variety of ways. Using immunoRCA, analytes were detected at sensitivities exceeding those of conventional enzyme immunoassays in ELISA and microparticle formats. The signal amplification afforded by immunoRCA also enabled immunoassays to be carried out in microspot and microarray formats with exquisite sensitivity. When Ags are present at concentrations down to fM levels, specifically bound Abs can be scored by counting discrete fluorescent signals arising from individual Ag–Ab complexes. Multiplex immunoRCA also was demonstrated by accurately quantifying Ags mixed in different ratios in a two-color, single-molecule-counting assay on a glass slide. ImmunoRCA thus combines high sensitivity and a very wide dynamic range with an unprecedented capability for single molecule detection. This Ag-detection method is of general applicability and is extendable to multiplexed immunoassays that employ a battery of different Abs, each labeled with a unique oligonucleotide primer, that can be discriminated by a color-coded visualization system. ImmunoRCA-profiling based on the simultaneous quantitation of multiple Ags should expand the power of immunoassays by exploiting the increased information content of ratio-based expression analysis.
Resumo:
FokI is a member an unusual class of restriction enzymes that recognize a specific DNA sequence and cleave nonspecifically a short distance away from that sequence. FokI consists of an N-terminal DNA recognition domain and a C-terminal cleavage domain. The bipartite nature of FokI has led to the development of artificial enzymes with novel specificities. We have solved the structure of FokI to 2.3 Å resolution. The structure reveals a dimer, in which the dimerization interface is mediated by the cleavage domain. Each monomer has an overall conformation similar to that found in the FokI–DNA complex, with the cleavage domain packing alongside the DNA recognition domain. In corroboration with the cleavage data presented in the accompanying paper in this issue of Proceedings, we propose a model for FokI DNA cleavage that requires the dimerization of FokI on DNA to cleave both DNA strands.