5 resultados para Enhanced genetic algorithms
em DigitalCommons@The Texas Medical Center
Resumo:
Essential biological processes are governed by organized, dynamic interactions between multiple biomolecular systems. Complexes are thus formed to enable the biological function and get dissembled as the process is completed. Examples of such processes include the translation of the messenger RNA into protein by the ribosome, the folding of proteins by chaperonins or the entry of viruses in host cells. Understanding these fundamental processes by characterizing the molecular mechanisms that enable then, would allow the (better) design of therapies and drugs. Such molecular mechanisms may be revealed trough the structural elucidation of the biomolecular assemblies at the core of these processes. Various experimental techniques may be applied to investigate the molecular architecture of biomolecular assemblies. High-resolution techniques, such as X-ray crystallography, may solve the atomic structure of the system, but are typically constrained to biomolecules of reduced flexibility and dimensions. In particular, X-ray crystallography requires the sample to form a three dimensional (3D) crystal lattice which is technically di‑cult, if not impossible, to obtain, especially for large, dynamic systems. Often these techniques solve the structure of the different constituent components within the assembly, but encounter difficulties when investigating the entire system. On the other hand, imaging techniques, such as cryo-electron microscopy (cryo-EM), are able to depict large systems in near-native environment, without requiring the formation of crystals. The structures solved by cryo-EM cover a wide range of resolutions, from very low level of detail where only the overall shape of the system is visible, to high-resolution that approach, but not yet reach, atomic level of detail. In this dissertation, several modeling methods are introduced to either integrate cryo-EM datasets with structural data from X-ray crystallography, or to directly interpret the cryo-EM reconstruction. Such computational techniques were developed with the goal of creating an atomic model for the cryo-EM data. The low-resolution reconstructions lack the level of detail to permit a direct atomic interpretation, i.e. one cannot reliably locate the atoms or amino-acid residues within the structure obtained by cryo-EM. Thereby one needs to consider additional information, for example, structural data from other sources such as X-ray crystallography, in order to enable such a high-resolution interpretation. Modeling techniques are thus developed to integrate the structural data from the different biophysical sources, examples including the work described in the manuscript I and II of this dissertation. At intermediate and high-resolution, cryo-EM reconstructions depict consistent 3D folds such as tubular features which in general correspond to alpha-helices. Such features can be annotated and later on used to build the atomic model of the system, see manuscript III as alternative. Three manuscripts are presented as part of the PhD dissertation, each introducing a computational technique that facilitates the interpretation of cryo-EM reconstructions. The first manuscript is an application paper that describes a heuristics to generate the atomic model for the protein envelope of the Rift Valley fever virus. The second manuscript introduces the evolutionary tabu search strategies to enable the integration of multiple component atomic structures with the cryo-EM map of their assembly. Finally, the third manuscript develops further the latter technique and apply it to annotate consistent 3D patterns in intermediate-resolution cryo-EM reconstructions. The first manuscript, titled An assembly model for Rift Valley fever virus, was submitted for publication in the Journal of Molecular Biology. The cryo-EM structure of the Rift Valley fever virus was previously solved at 27Å-resolution by Dr. Freiberg and collaborators. Such reconstruction shows the overall shape of the virus envelope, yet the reduced level of detail prevents the direct atomic interpretation. High-resolution structures are not yet available for the entire virus nor for the two different component glycoproteins that form its envelope. However, homology models may be generated for these glycoproteins based on similar structures that are available at atomic resolutions. The manuscript presents the steps required to identify an atomic model of the entire virus envelope, based on the low-resolution cryo-EM map of the envelope and the homology models of the two glycoproteins. Starting with the results of the exhaustive search to place the two glycoproteins, the model is built iterative by running multiple multi-body refinements to hierarchically generate models for the different regions of the envelope. The generated atomic model is supported by prior knowledge regarding virus biology and contains valuable information about the molecular architecture of the system. It provides the basis for further investigations seeking to reveal different processes in which the virus is involved such as assembly or fusion. The second manuscript was recently published in the of Journal of Structural Biology (doi:10.1016/j.jsb.2009.12.028) under the title Evolutionary tabu search strategies for the simultaneous registration of multiple atomic structures in cryo-EM reconstructions. This manuscript introduces the evolutionary tabu search strategies applied to enable a multi-body registration. This technique is a hybrid approach that combines a genetic algorithm with a tabu search strategy to promote the proper exploration of the high-dimensional search space. Similar to the Rift Valley fever virus, it is common that the structure of a large multi-component assembly is available at low-resolution from cryo-EM, while high-resolution structures are solved for the different components but lack for the entire system. Evolutionary tabu search strategies enable the building of an atomic model for the entire system by considering simultaneously the different components. Such registration indirectly introduces spatial constrains as all components need to be placed within the assembly, enabling the proper docked in the low-resolution map of the entire assembly. Along with the method description, the manuscript covers the validation, presenting the benefit of the technique in both synthetic and experimental test cases. Such approach successfully docked multiple components up to resolutions of 40Å. The third manuscript is entitled Evolutionary Bidirectional Expansion for the Annotation of Alpha Helices in Electron Cryo-Microscopy Reconstructions and was submitted for publication in the Journal of Structural Biology. The modeling approach described in this manuscript applies the evolutionary tabu search strategies in combination with the bidirectional expansion to annotate secondary structure elements in intermediate resolution cryo-EM reconstructions. In particular, secondary structure elements such as alpha helices show consistent patterns in cryo-EM data, and are visible as rod-like patterns of high density. The evolutionary tabu search strategy is applied to identify the placement of the different alpha helices, while the bidirectional expansion characterizes their length and curvature. The manuscript presents the validation of the approach at resolutions ranging between 6 and 14Å, a level of detail where alpha helices are visible. Up to resolution of 12 Å, the method measures sensitivities between 70-100% as estimated in experimental test cases, i.e. 70-100% of the alpha-helices were correctly predicted in an automatic manner in the experimental data. The three manuscripts presented in this PhD dissertation cover different computation methods for the integration and interpretation of cryo-EM reconstructions. The methods were developed in the molecular modeling software Sculptor (http://sculptor.biomachina.org) and are available for the scientific community interested in the multi-resolution modeling of cryo-EM data. The work spans a wide range of resolution covering multi-body refinement and registration at low-resolution along with annotation of consistent patterns at high-resolution. Such methods are essential for the modeling of cryo-EM data, and may be applied in other fields where similar spatial problems are encountered, such as medical imaging.
Resumo:
Familial hemiplegic migraine type 1 (FHM1) is an autosomal dominant subtype of migraine with aura that is associated with hemiparesis. As with other types of migraine, it affects women more frequently than men. FHM1 is caused by mutations in the CACNA1A gene, which encodes the alpha1A subunit of Cav2.1 channels; the R192Q mutation in CACNA1A causes a mild form of FHM1, whereas the S218L mutation causes a severe, often lethal phenotype. Spreading depression (SD), a slowly propagating neuronal and glial cell depolarization that leads to depression of neuronal activity, is the most likely cause of migraine aura. Here, we have shown that transgenic mice expressing R192Q or S218L FHM1 mutations have increased SD frequency and propagation speed; enhanced corticostriatal propagation; and, similar to the human FHM1 phenotype, more severe and prolonged post-SD neurological deficits. The susceptibility to SD and neurological deficits is affected by allele dosage and is higher in S218L than R192Q mutants. Further, female S218L and R192Q mutant mice were more susceptible to SD and neurological deficits than males. This sex difference was abrogated by ovariectomy and senescence and was partially restored by estrogen replacement, implicating ovarian hormones in the observed sex differences in humans with FHM1. These findings demonstrate that genetic and hormonal factors modulate susceptibility to SD and neurological deficits in FHM1 mutant mice, providing a potential mechanism for the phenotypic diversity of human migraine and aura.
Resumo:
Genetic anticipation is defined as a decrease in age of onset or increase in severity as the disorder is transmitted through subsequent generations. Anticipation has been noted in the literature for over a century. Recently, anticipation in several diseases including Huntington's Disease, Myotonic Dystrophy and Fragile X Syndrome were shown to be caused by expansion of triplet repeats. Anticipation effects have also been observed in numerous mental disorders (e.g. Schizophrenia, Bipolar Disorder), cancers (Li-Fraumeni Syndrome, Leukemia) and other complex diseases. ^ Several statistical methods have been applied to determine whether anticipation is a true phenomenon in a particular disorder, including standard statistical tests and newly developed affected parent/affected child pair methods. These methods have been shown to be inappropriate for assessing anticipation for a variety of reasons, including familial correlation and low power. Therefore, we have developed family-based likelihood modeling approaches to model the underlying transmission of the disease gene and penetrance function and hence detect anticipation. These methods can be applied in extended families, thus improving the power to detect anticipation compared with existing methods based only upon parents and children. The first method we have proposed is based on the regressive logistic hazard model. This approach models anticipation by a generational covariate. The second method allows alleles to mutate as they are transmitted from parents to offspring and is appropriate for modeling the known triplet repeat diseases in which the disease alleles can become more deleterious as they are transmitted across generations. ^ To evaluate the new methods, we performed extensive simulation studies for data simulated under different conditions to evaluate the effectiveness of the algorithms to detect genetic anticipation. Results from analysis by the first method yielded empirical power greater than 87% based on the 5% type I error critical value identified in each simulation depending on the method of data generation and current age criteria. Analysis by the second method was not possible due to the current formulation of the software. The application of this method to Huntington's Disease and Li-Fraumeni Syndrome data sets revealed evidence for a generation effect in both cases. ^
Resumo:
The molecular mechanisms responsible for the expansion and deletion of trinucleotide repeat sequences (TRS) are the focus of our studies. Several hereditary neurological diseases including Huntington's disease, myotonic dystrophy, and fragile X syndrome are associated with the instability of TRS. Using the well defined and controllable model system of Escherichia coli, the influences of three types of DNA incisions on genetic instability of CTG•CAG repeats were studied: DNA double-strand breaks (DSB), single-strand nicks, and single-strand gaps. The DNA incisions were generated in pUC19 derivatives by in vitro cleavage with restriction endonucleases. The cleaved DNA was then transformed into E. coli parental and mutant strains. Double-strand breaks induced deletions throughout the TRS region in an orientation dependent manner relative to the origin of replication. The extent of instability was enhanced by the repeat length and sequence (CTG•CAG vs. CGG•CCG). Mutations in recA and recBC increased deletions, mutations in recF stabilized the TRS, whereas mutations in ruvA had no effect. DSB were repaired by intramolecular recombination, versus an intermolecular gene conversion or crossover mechanism. 30 nt gaps formed a distinct 30 nt deletion product, whereas single strand nicks and gaps of 15 nts did not induce expansions or deletions. Formation of this deletion product required the CTG•CAG repeats to be present in the single-stranded region and was stimulated by E. coli DNA ligase, but was not dependent upon the RecFOR pathway. Models are presented to explain the DSB induced instabilities and formation of the 30 nucleotide deletion product. In addition to the in vitro creation of DSBs, several attempts to generate this incision in vivo with the use of EcoR I restriction modification systems were conducted. ^
Resumo:
Two molecular epidemiological studies were conducted to examine associations between genetic variation and risk of squamous cell carcinoma of the head and neck (SCCHN). In the first study, we hypothesized that genetic variation in p53 response elements (REs) may play roles in the etiology of SCCHN. We selected and genotyped five polymorphic p53 REs as well as a most frequently studied p53 codon 72 (Arg72Pro, rs1042522) polymorphism in 1,100 non-Hispanic White SCCHN patients and 1,122 age-and sex-matched cancer-free controls recruited at The University of Texas M. D. Anderson Cancer Center. In multivariate logistic regression analysis with adjustment for age, sex, smoking and drinking status, marital status and education level, we observed that the EOMES rs3806624 CC genotype had a significant effect of protection against SCCHN risk (adjusted odds ratio= 0.79, 95% confidence interval =0.64–0.98), compared with the -838TT+CT genotypes. Moreover, a significantly increased risk associated with the combined genotypes of p53 codon 72CC and EOMES -838TT+CT was observed, especially in the subgroup of non-oropharyneal cancer patients. The values of false-positive report probability were also calculated for significant findings. In the second study, we assessed the association between SCCHN risk and four potential regulatory single nucleotide polymorphisms (SNPs) of DEC1 (deleted in esophageal cancer 1) gene, a candidate tumor suppressor gene for esophageal cancer. After adjustment for age, sex, and smoking and drinking status, the variant -606CC (i.e., -249CC) homozygotes had a significantly reduced SCCHN risk (adjusted odds ratio = 0.71, 95% confidence interval = 0.52–0.99), compared with the -606TT homozygotes. Stratification analyses showed that a reduced risk associated with the -606CC genotype was more pronounced in subgroups of non-smokers, non-drinkers, younger subjects (defined as ≤ 57 years), carriers of TP53 Arg/Arg (rs1042522) genotype, patients with oropharyngeal cancer or late-stage SCCHN. Further in silico analysis revealed that the -249 T-to-C change led to a gain of a transcription factor binding site. Additional functional analysis showed that the -249T-to-C change significantly enhanced transcriptional activity of the DEC1 promoter and the DNA-protein binding activity. We conclude that the DEC1 promoter -249 T>C (rs2012775) polymorphism is functional, modulating susceptibility to SCCHN among non-Hispanic Whites. Additional large-scale, preferably population-based studies are needed to validate our findings.^