985 resultados para PROTEIN NETWORKS


Relevância:

30.00% 30.00%

Publicador:

Resumo:

In the last decade, two tools, one drawn from information theory and the other from artificial neural networks, have proven particularly useful in many different areas of sequence analysis. The work presented herein indicates that these two approaches can be joined in a general fashion to produce a very powerful search engine that is capable of locating members of a given nucleic acid sequence family in either local or global sequence searches. This program can, in turn, be queried for its definition of the motif under investigation, ranking each base in context for its contribution to membership in the motif family. In principle, the method used can be applied to any binding motif, including both DNA and RNA sequence families, given sufficient family size.

Relevância:

30.00% 30.00%

Publicador:

Resumo:

The small GTP-binding protein Cdc42 is thought to induce filopodium formation by regulating actin polymerization at the cell cortex. Although several Cdc42-binding proteins have been identified and some of them have been implicated in filopodium formation, the precise role of Cdc42 in modulating actin polymerization has not been defined. To understand the biochemical pathways that link Cdc42 to the actin cytoskeleton, we have reconstituted Cdc42-induced actin polymerization in Xenopus egg extracts. Using this cell-free system, we have developed a rapid and specific assay that has allowed us to fractionate the extract and isolate factors involved in this activity. We report here that at least two biochemically distinct components are required, based on their chromatographic behavior and affinity for Cdc42. One component is purified to homogeneity and is identified as the Arp2/3 complex, a protein complex that has been shown to nucleate actin polymerization. However, the purified complex alone is not sufficient to mediate the activity; a second component that binds Cdc42 directly and mediates the interaction between Cdc42 and the complex also is required. These results establish an important link between a signaling molecule, Cdc42, and a complex that can directly modulate actin networks in vitro. We propose that activation of the Arp2/3 complex by Cdc42 and other signaling molecules plays a central role in stimulating actin polymerization at the cell surface.

Relevância:

30.00% 30.00%

Publicador:

Resumo:

SBASE 8.0 is the eighth release of the SBASE library of protein domain sequences that contains 294 898 annotated structural, functional, ligand-binding and topogenic segments of proteins, cross-referenced to most major sequence databases and sequence pattern collections. The entries are clustered into over 2005 statistically validated domain groups (SBASE-A) and 595 non-validated groups (SBASE-B), provided with several WWW-based search and browsing facilities for online use. A domain-search facility was developed, based on non-parametric pattern recognition methods, including artificial neural networks. SBASE 8.0 is freely available by anonymous ‘ftp’ file transfer from ftp.icgeb.trieste.it. Automated searching of SBASE can be carried out with the WWW servers http://www.icgeb.trieste.it/sbase/ and http://sbase.abc.hu/sbase/.

Relevância:

30.00% 30.00%

Publicador:

Resumo:

Functional annotation of novel genes can be achieved by detection of interactions of their encoded proteins with known proteins followed by assays to validate that the gene participates in a specific cellular function. We report an experimental strategy that allows for detection of protein interactions and functional assays with a single reporter system. Interactions among biochemical network component proteins are detected and probed with stimulators and inhibitors of the network. In addition, the cellular location of the interacting proteins is determined. We used this strategy to map a signal transduction network that controls initiation of translation in eukaryotes. We analyzed 35 different pairs of full-length proteins and identified 14 interactions, of which five have not been observed previously, suggesting that the organization of the pathway is more ramified and integrated than previously shown. Our results demonstrate the feasibility of using this strategy in efforts of genomewide functional annotation.

Relevância:

30.00% 30.00%

Publicador:

Resumo:

Cells are intrinsically noisy biochemical reactors: low reactant numbers can lead to significant statistical fluctuations in molecule numbers and reaction rates. Here we use an analytic model to investigate the emergent noise properties of genetic systems. We find for a single gene that noise is essentially determined at the translational level, and that the mean and variance of protein concentration can be independently controlled. The noise strength immediately following single gene induction is almost twice the final steady-state value. We find that fluctuations in the concentrations of a regulatory protein can propagate through a genetic cascade; translational noise control could explain the inefficient translation rates observed for genes encoding such regulatory proteins. For an autoregulatory protein, we demonstrate that negative feedback efficiently decreases system noise. The model can be used to predict the noise characteristics of networks of arbitrary connectivity. The general procedure is further illustrated for an autocatalytic protein and a bistable genetic switch. The analysis of intrinsic noise reveals biological roles of gene network structures and can lead to a deeper understanding of their evolutionary origin.

Relevância:

30.00% 30.00%

Publicador:

Resumo:

Nerve cells contain abundant subpopulations of cold-stable microtubules. We have previously isolated a calmodulin-regulated brain protein, STOP (stable tubule-only polypeptide), which reconstitutes microtubule cold stability when added to cold-labile microtubules in vitro. We have now cloned cDNA encoding STOP. We find that STOP is a 100.5-kDa protein with no homology to known proteins. The primary structure of STOP includes two distinct domains of repeated motifs. The central region of STOP contains 5 tandem repeats of 46 amino acids, 4 with 98% homology to the consensus sequence. The STOP C terminus contains 28 imperfect repeats of an 11-amino acid motif. STOP also contains a putative SH3-binding motif close to its N terminus. In vitro translated STOP binds to both microtubules and Ca2+-calmodulin. When STOP cDNA is expressed in cells that lack cold-stable microtubules, STOP associates with microtubules at 37 degrees C, and stabilizes microtubule networks, inducing cold stability, nocodazole resistance, and tubulin detyrosination on microtubules in transfected cells. We conclude that STOP must play an important role in the generation of microtubule cold stability and in the control of microtubule dynamics in brain.

Relevância:

30.00% 30.00%

Publicador:

Resumo:

We describe a new method for using neural networks to predict residue contact pairs in a protein. The main inputs to the neural network are a set of 25 measures of correlated mutation between all pairs of residues in two windows of size 5 centered on the residues of interest. While the individual pair-wise correlations are a relatively weak predictor of contact, by training the network on windows of correlation the accuracy of prediction is significantly improved. The neural network is trained on a set of 100 proteins and then tested on a disjoint set of 1033 proteins of known structure. An average predictive accuracy of 21.7% is obtained taking the best L/2 predictions for each protein, where L is the sequence length. Taking the best L/10 predictions gives an average accuracy of 30.7%. The predictor is also tested on a set of 59 proteins from the CASP5 experiment. The accuracy is found to be relatively consistent across different sequence lengths, but to vary widely according to the secondary structure. Predictive accuracy is also found to improve by using multiple sequence alignments containing many sequences to calculate the correlations. (C) 2004 Wiley-Liss, Inc.

Relevância:

30.00% 30.00%

Publicador:

Resumo:

Networks of interactions evolve in many different domains. They tend to have topological characteristics in common, possibly due to common factors in the way the networks grow and develop. It has been recently suggested that one such common characteristic is the presence of a hierarchically modular organization. In this paper, we describe a new algorithm for the detection and quantification of hierarchical modularity, and demonstrate that the yeast protein-protein interaction network does have a hierarchically modular organization. We further show that such organization is evident in artificial networks produced by computational evolution using a gene duplication operator, but not in those developing via preferential attachment of new nodes to highly connected existing nodes. (C) 2004 Elsevier Ireland Ltd. All rights reserved.

Relevância:

30.00% 30.00%

Publicador:

Resumo:

Background: The multitude of motif detection algorithms developed to date have largely focused on the detection of patterns in primary sequence. Since sequence-dependent DNA structure and flexibility may also play a role in protein-DNA interactions, the simultaneous exploration of sequence-and structure-based hypotheses about the composition of binding sites and the ordering of features in a regulatory region should be considered as well. The consideration of structural features requires the development of new detection tools that can deal with data types other than primary sequence. Results: GANN ( available at http://bioinformatics.org.au/gann) is a machine learning tool for the detection of conserved features in DNA. The software suite contains programs to extract different regions of genomic DNA from flat files and convert these sequences to indices that reflect sequence and structural composition or the presence of specific protein binding sites. The machine learning component allows the classification of different types of sequences based on subsamples of these indices, and can identify the best combinations of indices and machine learning architecture for sequence discrimination. Another key feature of GANN is the replicated splitting of data into training and test sets, and the implementation of negative controls. In validation experiments, GANN successfully merged important sequence and structural features to yield good predictive models for synthetic and real regulatory regions. Conclusion: GANN is a flexible tool that can search through large sets of sequence and structural feature combinations to identify those that best characterize a set of sequences.

Relevância:

30.00% 30.00%

Publicador:

Resumo:

Motivation: Targeting peptides direct nascent proteins to their specific subcellular compartment. Knowledge of targeting signals enables informed drug design and reliable annotation of gene products. However, due to the low similarity of such sequences and the dynamical nature of the sorting process, the computational prediction of subcellular localization of proteins is challenging. Results: We contrast the use of feed forward models as employed by the popular TargetP/SignalP predictors with a sequence-biased recurrent network model. The models are evaluated in terms of performance at the residue level and at the sequence level, and demonstrate that recurrent networks improve the overall prediction performance. Compared to the original results reported for TargetP, an ensemble of the tested models increases the accuracy by 6 and 5% on non-plant and plant data, respectively.

Relevância:

30.00% 30.00%

Publicador:

Resumo:

Selection of machine learning techniques requires a certain sensitivity to the requirements of the problem. In particular, the problem can be made more tractable by deliberately using algorithms that are biased toward solutions of the requisite kind. In this paper, we argue that recurrent neural networks have a natural bias toward a problem domain of which biological sequence analysis tasks are a subset. We use experiments with synthetic data to illustrate this bias. We then demonstrate that this bias can be exploitable using a data set of protein sequences containing several classes of subcellular localization targeting peptides. The results show that, compared with feed forward, recurrent neural networks will generally perform better on sequence analysis tasks. Furthermore, as the patterns within the sequence become more ambiguous, the choice of specific recurrent architecture becomes more critical.

Relevância:

30.00% 30.00%

Publicador:

Resumo:

The regulation of osteoclast differentiation in the bone microenvironment is critical for normal bone remodeling, as well as for various human bone diseases. Over the last decade, our knowledge of how osteoclast differentiation occurs has progressed rapidly. We highlight some of the major advances in understanding how cell signaling and transcription are integrated to direct the differentiation of this cell type. These studies used genetic, molecular, and biochemical approaches. Additionally, we summarize data obtained from studies of osteoclast differentiation that used the functional genomic approach of global gene profiling applied to osteoclast differentiation. This genomic data confirms results from studies using the classical experimental approaches and also may suggest new modes by which osteoclast differentiation and function can be modulated. Two conclusions that emerge are that osteoclast differentiation depends on a combination of fairly ubiquitously expressed transcription factors rather than unique osteoclast factors, and that the overlay of cell signaling pathways on this set of transcription factors provides a powerful mechanism to fine tune the differentiation program in response to the local bone microenvironment.

Relevância:

30.00% 30.00%

Publicador:

Resumo:

Background: The structure of proteins may change as a result of the inherent flexibility of some protein regions. We develop and explore probabilistic machine learning methods for predicting a continuum secondary structure, i.e. assigning probabilities to the conformational states of a residue. We train our methods using data derived from high-quality NMR models. Results: Several probabilistic models not only successfully estimate the continuum secondary structure, but also provide a categorical output on par with models directly trained on categorical data. Importantly, models trained on the continuum secondary structure are also better than their categorical counterparts at identifying the conformational state for structurally ambivalent residues. Conclusion: Cascaded probabilistic neural networks trained on the continuum secondary structure exhibit better accuracy in structurally ambivalent regions of proteins, while sustaining an overall classification accuracy on par with standard, categorical prediction methods.

Relevância:

30.00% 30.00%

Publicador:

Resumo:

This paper presents a composite multi-layer classifier system for predicting the subcellular localization of proteins based on their amino acid sequence. The work is an extension of our previous predictor PProwler v1.1 which is itself built upon the series of predictors SignalP and TargetP. In this study we outline experiments conducted to improve the classifier design. The major improvement came from using Support Vector machines as a "smart gate" sorting the outputs of several different targeting peptide detection networks. Our final model (PProwler v1.2) gives MCC values of 0.873 for non-plant and 0.849 for plant proteins. The model improves upon the accuracy of our previous subcellular localization predictor (PProwler v1.1) by 2% for plant data (which represents 7.5% improvement upon TargetP).

Relevância:

30.00% 30.00%

Publicador:

Resumo:

Time delay is an important aspect in the modelling of genetic regulation due to slow biochemical reactions such as gene transcription and translation, and protein diffusion between the cytosol and nucleus. In this paper we introduce a general mathematical formalism via stochastic delay differential equations for describing time delays in genetic regulatory networks. Based on recent developments with the delay stochastic simulation algorithm, the delay chemical masterequation and the delay reaction rate equation are developed for describing biological reactions with time delay, which leads to stochastic delay differential equations derived from the Langevin approach. Two simple genetic regulatory networks are used to study the impact of' intrinsic noise on the system dynamics where there are delays. (c) 2006 Elsevier B.V. All rights reserved.