994 resultados para Functional Annotation


Relevância:

60.00% 60.00%

Publicador:

Resumo:

During the last few years, next-generation sequencing (NGS) technologies have accelerated the detection of genetic variants resulting in the rapid discovery of new disease-associated genes. However, the wealth of variation data made available by NGS alone is not sufficient to understand the mechanisms underlying disease pathogenesis and manifestation. Multidisciplinary approaches combining sequence and clinical data with prior biological knowledge are needed to unravel the role of genetic variants in human health and disease. In this context, it is crucial that these data are linked, organized, and made readily available through reliable online resources. The Swiss-Prot section of the Universal Protein Knowledgebase (UniProtKB/Swiss-Prot) provides the scientific community with a collection of information on protein functions, interactions, biological pathways, as well as human genetic diseases and variants, all manually reviewed by experts. In this article, we present an overview of the information content of UniProtKB/Swiss-Prot to show how this knowledgebase can support researchers in the elucidation of the mechanisms leading from a molecular defect to a disease phenotype.

Relevância:

60.00% 60.00%

Publicador:

Resumo:

INTRODUCTION: Diverse microarray and sequencing technologies have been widely used to characterise the molecular changes in malignant epithelial cells in breast cancers. Such gene expression studies to identify markers and targets in tumour cells are, however, compromised by the cellular heterogeneity of solid breast tumours and by the lack of appropriate counterparts representing normal breast epithelial cells. METHODS: Malignant neoplastic epithelial cells from primary breast cancers and luminal and myoepithelial cells isolated from normal human breast tissue were isolated by immunomagnetic separation methods. Pools of RNA from highly enriched preparations of these cell types were subjected to expression profiling using massively parallel signature sequencing (MPSS) and four different genome wide microarray platforms. Functional related transcripts of the differential tumour epithelial transcriptome were used for gene set enrichment analysis to identify enrichment of luminal and myoepithelial type genes. Clinical pathological validation of a small number of genes was performed on tissue microarrays. RESULTS: MPSS identified 6,553 differentially expressed genes between the pool of normal luminal cells and that of primary tumours substantially enriched for epithelial cells, of which 98% were represented and 60% were confirmed by microarray profiling. Significant expression level changes between these two samples detected only by microarray technology were shown by 4,149 transcripts, resulting in a combined differential tumour epithelial transcriptome of 8,051 genes. Microarray gene signatures identified a comprehensive list of 907 and 955 transcripts whose expression differed between luminal epithelial cells and myoepithelial cells, respectively. Functional annotation and gene set enrichment analysis highlighted a group of genes related to skeletal development that were associated with the myoepithelial/basal cells and upregulated in the tumour sample. One of the most highly overexpressed genes in this category, that encoding periostin, was analysed immunohistochemically on breast cancer tissue microarrays and its expression in neoplastic cells correlated with poor outcome in a cohort of poor prognosis estrogen receptor-positive tumours. CONCLUSION: Using highly enriched cell populations in combination with multiplatform gene expression profiling studies, a comprehensive analysis of molecular changes between the normal and malignant breast tissue was established. This study provides a basis for the identification of novel and potentially important targets for diagnosis, prognosis and therapy in breast cancer.

Relevância:

60.00% 60.00%

Publicador:

Resumo:

Background: Global analyses of human disease genes by computational methods have yielded important advances in the understanding of human diseases. Generally these studies have treated the group of disease genes uniformly, thus ignoring the type of disease-causing mutations (dominant or recessive). In this report we present a comprehensive study of the evolutionary history of autosomal disease genes separated by mode of inheritance.Results: We examine differences in protein and coding sequence conservation between dominant and recessive human disease genes. Our analysis shows that disease genes affected by dominant mutations are more conserved than those affected by recessive mutations. This could be a consequence of the fact that recessive mutations remain hidden from selection while heterozygous. Furthermore, we employ functional annotation analysis and investigations into disease severity to support this hypothesis. Conclusion: This study elucidates important differences between dominantly- and recessively-acting disease genes in terms of protein and DNA sequence conservation, paralogy and essentiality. We propose that the division of disease genes by mode of inheritance will enhance both understanding of the disease process and prediction of candidate disease genes in the future.

Relevância:

60.00% 60.00%

Publicador:

Resumo:

En este trabajo se describe una base de conocimiento de las ALU humanas. La ontología incorpora términos SO y GO y está orientada a describir el contexto genómico del conjunto de ALU. Para cada elemento ALU se almacenan el gen y transcrito más cercanos, así como su anotación funcional de acuerdo a GO, el estado de la cromatina circundante y los factores de transcripción presentes en la ALU. Se han incorporado reglas semánticas para facilitar el almacenamiento, consulta e integración de la información. La ontología de ALU es plenamente analizable mediante razonadores como Pellet y está parcialmente transferida a una wiki semántica.

Relevância:

60.00% 60.00%

Publicador:

Resumo:

With the increasing availability of various 'omics data, high-quality orthology assignment is crucial for evolutionary and functional genomics studies. We here present the fourth version of the eggNOG database (available at http://eggnog.embl.de) that derives nonsupervised orthologous groups (NOGs) from complete genomes, and then applies a comprehensive characterization and analysis pipeline to the resulting gene families. Compared with the previous version, we have more than tripled the underlying species set to cover 3686 organisms, keeping track with genome project completions while prioritizing the inclusion of high-quality genomes to minimize error propagation from incomplete proteome sets. Major technological advances include (i) a robust and scalable procedure for the identification and inclusion of high-quality genomes, (ii) provision of orthologous groups for 107 different taxonomic levels compared with 41 in eggNOGv3, (iii) identification and annotation of particularly closely related orthologous groups, facilitating analysis of related gene families, (iv) improvements of the clustering and functional annotation approach, (v) adoption of a revised tree building procedure based on the multiple alignments generated during the process and (vi) implementation of quality control procedures throughout the entire pipeline. As in previous versions, eggNOGv4 provides multiple sequence alignments and maximum-likelihood trees, as well as broad functional annotation. Users can access the complete database of orthologous groups via a web interface, as well as through bulk download.

Relevância:

60.00% 60.00%

Publicador:

Resumo:

BackgroundBipolar disorder is a highly heritable polygenic disorder. Recent enrichment analyses suggest that there may be true risk variants for bipolar disorder in the expression quantitative trait loci (eQTL) in the brain.AimsWe sought to assess the impact of eQTL variants on bipolar disorder risk by combining data from both bipolar disorder genome-wide association studies (GWAS) and brain eQTL.MethodTo detect single nucleotide polymorphisms (SNPs) that influence expression levels of genes associated with bipolar disorder, we jointly analysed data from a bipolar disorder GWAS (7481 cases and 9250 controls) and a genome-wide brain (cortical) eQTL (193 healthy controls) using a Bayesian statistical method, with independent follow-up replications. The identified risk SNP was then further tested for association with hippocampal volume (n = 5775) and cognitive performance (n = 342) among healthy individuals.ResultsIntegrative analysis revealed a significant association between a brain eQTL rs6088662 on chromosome 20q11.22 and bipolar disorder (log Bayes factor = 5.48; bipolar disorder P = 5.85×10(-5)). Follow-up studies across multiple independent samples confirmed the association of the risk SNP (rs6088662) with gene expression and bipolar disorder susceptibility (P = 3.54×10(-8)). Further exploratory analysis revealed that rs6088662 is also associated with hippocampal volume and cognitive performance in healthy individuals.ConclusionsOur findings suggest that 20q11.22 is likely a risk region for bipolar disorder; they also highlight the informative value of integrating functional annotation of genetic variants for gene expression in advancing our understanding of the biological basis underlying complex disorders, such as bipolar disorder.

Relevância:

60.00% 60.00%

Publicador:

Resumo:

The HERC gene family encodes proteins with two characteristic domains: HECT and RCC1-like. Proteins with HECT domain shave been described to function as ubiquitin ligases, and those that contain RCC1-like domains have been reported to function as GTPases regulators. These two activities are essential in a number of important cellular processes such as cell cycle, cell signaling, and membrane trafficking. Mutations affecting these domains have been found associated with retinitis pigmentosa, amyotrophic lateral sclerosis, and cancer. In humans, six HERC genes have been reported which encode two subgroups of HERC proteins: large (HERC1-2) and small (HERC3-6). The giant HERC1 protein was the first to be identified. It has been involved in membrane trafficking and cell proliferation/growth through its interactions with clathrin, M2-pyruvate kinase, and TSC2 proteins. Mutations affecting other members of the HERC family have been found to be associated with sterility and growth retardation. Here, we report the characterization of a recessive mutation named tambaleante, which causes progressive Purkinje cell degeneration leading to severe ataxia with reduced growth and lifespan in homozygous mice aged over two months. We mapped this mutation in mouse chromosome 9 and then performed positional cloning. We found a GuA transition at position 1448, causing a Gly to Glu substitution (Gly483Glu) in the highly conserved N- terminal RCC1-like domain of the HERC1 protein. Successful transgenic rescue, with either a mouse BAC containing the normal copy of Herc1 or with the human HERC1 cDNA, validated our findings. Histological and biochemical studies revealed extensive autophagy associated with an increase of the mutant protein level and a decrease of mTOR activity. Our observations concerning this first mutation in the Herc1 gene contribute to the functional annotation of the encoded E3 ubiquitin ligase and underline the crucial and unexpected role of this protein in Purkinje cell physiology.

Relevância:

60.00% 60.00%

Publicador:

Resumo:

Chez les végétaux supérieurs, l’embryogenèse est une phase clé du développement au cours de laquelle l’embryon établit les principales structures qui formeront la future plante et synthétise et accumule des réserves définissant le rendement et la qualité nutritionnelle des graines. Ainsi, la compréhension des évènements moléculaires et physiologiques menant à la formation de la graine représente un intérêt agronomique majeur. Toutefois, l'analyse des premiers stades de développement est souvent difficile parce que l'embryon est petit et intégré à l'intérieur du tissu maternel. Solanum chacoense qui présente des fleurs relativement grande facilitant l’isolation des ovules, a été utilisée pour l’étude de la biologie de la reproduction plus précisément la formation des gamètes femelles, la pollinisation, la fécondation et le développement des embryons. Afin d'analyser le programme transcriptionnel induit au cours de la structuration de ces étapes de la reproduction sexuée, nous avons mis à profit un projet de séquençage de 7741 ESTs (6700 unigènes) exprimés dans l’ovule à différents stades du développement embryonnaire. L’ADN de ces ESTs a été utilisé pour la fabrication de biopuces d’ADN. Dans un premier temps, ces biopuces ont été utilisé pour comparer des ADNc issus des ovules de chaque stade de développement embryonnaire (depuis le zygote jusqu’au embryon mature) versus un ovule non fécondé. Trois profils d’expression correspondant au stade précoce, intermédiaire et tardive ont été trouvés. Une analyse plus approfondie entre chaque point étudié (de 0 à 22 jours après pollinisation), a permis d'identifier des gènes spécifiques caractérisant des phases de transition spécifiques. Les annotations Fonctionnelles des gènes differentiellement exprimés nous ont permis d'identifier les principales fonctions cellulaires impliquées à chaque stade de développement, révélant que les embryons sont engagés dans des actifs processus de différenciation. Ces biopuces d’ADN ont été par la suite utilisé pour comparer différent types de pollinisation (compatible, incompatible, semi-compatible et inter-espèce) afin d’identifier les gènes répondants à plusieurs stimuli avant l'arrivé du tube pollinique aux ovules (activation à distance). Nous avons pu démontrer que le signal perçu par l’ovaire était différent et dépend de plusieurs facteurs, incluant le type de pollen et la distance parcourue par le pollen dans le style. Une autre analyse permettant la comparaison des différentes pollinisations et la blessure du style nous a permis d’identifier que les programmes génétiques de la pollinisation chevauchent en partie avec ceux du stress. Cela était confirmé en traitant les fleurs par une hormone de stress, méthyle jasmonate. Dans le dernier chapitre, nous avons utilisé ces biopuces pour étudier le changement transcriptionnel d’un mutant sur exprimant une protéine kinase FRK2 impliqué dans l’identité des ovules. Nous avons pu sélectionner plusieurs gènes candidat touchés par la surexpression de cette kinase pour mieux comprendre la voie se signalisation. Ces biopuces ont ainsi servi à déterminer la variation au niveau transcriptionnelle des gènes impliqués lors de différents stades de la reproduction sexuée chez les plantes et nous a permis de mieux comprendre ces étapes.

Relevância:

60.00% 60.00%

Publicador:

Resumo:

Les habitudes de consommation de substances psychoactives, le stress, l’obésité et les traits cardiovasculaires associés seraient en partie reliés aux mêmes facteurs génétiques. Afin d’explorer cette hypothèse, nous avons effectué, chez 119 familles multi-générationnelles québécoises de la région du Saguenay-Lac-St-Jean, des études d’association et de liaison pangénomiques pour les composantes génétiques : de la consommation usuelle d’alcool, de tabac et de café, de la réponse au stress physique et psychologique, des traits anthropométriques reliés à l’obésité, ainsi que des mesures du rythme cardiaque (RC) et de la pression artérielle (PA). 58000 SNPs et 437 marqueurs microsatellites ont été utilisés et l’annotation fonctionnelle des gènes candidats identifiés a ensuite été réalisée. Nous avons détecté des corrélations phénotypiques significatives entre les substances psychoactives, le stress, l’obésité et les traits hémodynamiques. Par exemple, les consommateurs d’alcool et de tabac ont montré un RC significativement diminué en réponse au stress psychologique. De plus, les consommateurs de tabac avaient des PA plus basses que les non-consommateurs. Aussi, les hypertendus présentaient des RC et PA systoliques accrus en réponse au stress psychologique et un indice de masse corporelle (IMC) élevé, comparativement aux normotendus. D’autre part, l’utilisation de tabac augmenterait les taux corporels d’épinéphrine, et des niveaux élevés d’épinéphrine ont été associés à des IMC diminués. Ainsi, en accord avec les corrélations inter-phénotypiques, nous avons identifié plusieurs gènes associés/liés à la consommation de substances psychoactives, à la réponse au stress physique et psychologique, aux traits reliés à l’obésité et aux traits hémodynamiques incluant CAMK4, CNTN4, DLG2, DAG1, FHIT, GRID2, ITPR2, NOVA1, NRG3 et PRKCE. Ces gènes codent pour des protéines constituant un réseau d’interactions, impliquées dans la plasticité synaptique, et hautement exprimées dans le cerveau et ses tissus associés. De plus, l’analyse des sentiers de signalisation pour les gènes identifiés (P = 0,03) a révélé une induction de mécanismes de Potentialisation à Long Terme. Les variations des traits étudiés seraient en grande partie liées au sexe et au statut d’hypertension. Pour la consommation de tabac, nous avons noté que le degré et le sens des corrélations avec l’obésité, les traits hémodynamiques et le stress sont spécifiques au sexe et à la pression artérielle. Par exemple, si des variations ont été détectées entre les hommes fumeurs et non-fumeurs (anciens et jamais), aucune différence n’a été observée chez les femmes. Nous avons aussi identifié de nombreux traits reliés à l’obésité dont la corrélation avec la consommation de tabac apparaît essentiellement plus liée à des facteurs génétiques qu’au fait de fumer en lui-même. Pour le sexe et l’hypertension, des différences dans l’héritabilité de nombreux traits ont également été observées. En effet, des analyses génétiques sur des sous-groupes spécifiques ont révélé des gènes additionnels partageant des fonctions synaptiques : CAMK4, CNTN5, DNM3, KCNAB1 (spécifique à l’hypertension), CNTN4, DNM3, FHIT, ITPR1 and NRXN3 (spécifique au sexe). Ces gènes codent pour des protéines interagissant avec les protéines de gènes détectés dans l’analyse générale. De plus, pour les gènes des sous-groupes, les résultats des analyses des sentiers de signalisation et des profils d’expression des gènes ont montré des caractéristiques similaires à celles de l’analyse générale. La convergence substantielle entre les déterminants génétiques des substances psychoactives, du stress, de l’obésité et des traits hémodynamiques soutiennent la notion selon laquelle les variations génétiques des voies de plasticité synaptique constitueraient une interface commune avec les différences génétiques liées au sexe et à l’hypertension. Nous pensons, également, que la plasticité synaptique interviendrait dans de nombreux phénotypes complexes influencés par le mode de vie. En définitive, ces résultats indiquent que des approches basées sur des sous-groupes et des réseaux amélioreraient la compréhension de la nature polygénique des phénotypes complexes, et des processus moléculaires communs qui les définissent.

Relevância:

60.00% 60.00%

Publicador:

Resumo:

Laryngeal squamous cell carcinoma is very common in head and neck cancer, with high mortality rates and poor prognosis. In this study, we compared expression profiles of clinical samples from 13 larynx tumors and 10 non-neoplastic larynx tissues using a custom-built cDNA microarray containing 331 probes for 284 genes previously identified by informatics analysis of EST databases as markers of head and neck tumors. Thirty-five genes showed statistically significant differences (SNR >= 11.01, p <= 0.001) in the expression between tumor and non-tumor larynx tissue samples. Functional annotation indicated that these genes are involved in cellular processes relevant to the cancer phenotype, such as apoptosis, cell cycle, DNA repair, proteolysis, protease inhibition, signal transduction and transcriptional regulation. Six of the identified transcripts map to intronic regions of protein-coding genes and may comprise non-annotated exons or as yet uncharacterized long ncRNAs with a regulatory role in the gene expression program of larynx tissue. The differential expression of 10 of these genes (ADCY6, AES, AL2SCR3, CRR9, CSTB, DUSP1, MAP3K5, PLAT, UBL1 and ZNF706) was independently confirmed by quantitative real-time RT-PCR. Among these, the CSTB gene product has cysteine protease inhibitor activity that has been associated with an antimetastatic function. Interestingly, CSTB showed a low expression in the tumor samples analyzed (p<0.0001). The set of genes identified here contribute to a better understanding of the molecular basis of larynx cancer, and provide candidate markers for improving diagnosis, prognosis and treatment of this carcinoma.

Relevância:

60.00% 60.00%

Publicador:

Resumo:

Fundação de Amparo à Pesquisa do Estado de São Paulo (FAPESP)

Relevância:

60.00% 60.00%

Publicador:

Resumo:

Fundação de Amparo à Pesquisa do Estado de São Paulo (FAPESP)

Relevância:

60.00% 60.00%

Publicador:

Resumo:

Conselho Nacional de Desenvolvimento Científico e Tecnológico (CNPq)

Relevância:

60.00% 60.00%

Publicador:

Resumo:

Pós-graduação em Genética - IBILCE

Relevância:

60.00% 60.00%

Publicador:

Resumo:

Pós-graduação em Genética e Melhoramento Animal - FCAV