949 resultados para Document databases
Resumo:
Report for the scientific sojourn at the Swiss Federal Institute of Technology Zurich, Switzerland, between September and December 2007. In order to make robots useful assistants for our everyday life, the ability to learn and recognize objects is of essential importance. However, object recognition in real scenes is one of the most challenging problems in computer vision, as it is necessary to deal with difficulties. Furthermore, in mobile robotics a new challenge is added to the list: computational complexity. In a dynamic world, information about the objects in the scene can become obsolete before it is ready to be used if the detection algorithm is not fast enough. Two recent object recognition techniques have achieved notable results: the constellation approach proposed by Lowe and the bag of words approach proposed by Nistér and Stewénius. The Lowe constellation approach is the one currently being used in the robot localization project of the COGNIRON project. This report is divided in two main sections. The first section is devoted to briefly review the currently used object recognition system, the Lowe approach, and bring to light the drawbacks found for object recognition in the context of indoor mobile robot navigation. Additionally the proposed improvements for the algorithm are described. In the second section the alternative bag of words method is reviewed, as well as several experiments conducted to evaluate its performance with our own object databases. Furthermore, some modifications to the original algorithm to make it suitable for object detection in unsegmented images are proposed.
Resumo:
Este trabajo presenta un sistema para detectar y clasificar objetos binarios según la forma de éstos. En el primer paso del procedimiento, se aplica un filtrado para extraer el contorno del objeto. Con la información de los puntos de forma se obtiene un descriptor BSM con características altamente descriptivas, universales e invariantes. En la segunda fase del sistema se aprende y se clasifica la información del descriptor mediante Adaboost y Códigos Correctores de Errores. Se han usado bases de datos públicas, tanto en escala de grises como en color, para validar la implementación del sistema diseñado. Además, el sistema emplea una interfaz interactiva en la que diferentes métodos de procesamiento de imágenes pueden ser aplicados.
Resumo:
En la presente memoria se detallan con exactitud los pasos y procesos realizados para construir una aplicación que posibilite el cruce de datos genéticos a partir de información contenida en bases de datos remotas. Desarrolla un estudio en profundidad del contenido y estructura de las bases de datos remotas del NCBI y del KEGG, documentando una minería de datos con el objetivo de extraer de ellas la información necesaria para desarrollar la aplicación de cruce de datos genéticos. Finalmente se establecen los programas, scripts y entornos gráficos que han sido implementados para la construcción y posterior puesta en marcha de la aplicación que proporciona la funcionalidad de cruce de la que es objeto este proyecto fin de carrera.
Resumo:
Des de que començà aquest projecte, el grup de recerca ha intentat aprofundir el coneixement de la Catalunya de la Guerra del Francès (1808 -1814) a partir d’una òptica britànica. El grup pretenia així desenvolupar la relació que es va establir entre els catalans i els britànics al llarg de tota la guerra, des dels primers contactes permesos per la presència de la flota britànica en la costa catalana fins a la intervenció de forces britàniques en territori català. D’aquesta manera, i primerament, el grup inicià la consulta de les bases de dades i catàlegs catalans i britànics per a completar el nostre llistat de referències arxivístiques i bibliogràfiques. El segon pas van ésser les tres estades d’investigació que entre el 2006 i el 2007 es van fer a Anglaterra, principalment a Londres. La investigació es realitzà a la British Library, al Institute of Historical Research of the School of Advanced Studies de la University of London, al National Maritime Museum i als National Archives of the United Kingdom. A continuació, el grup analitzà la informació recollida de la lectura de fonts primàries i bibliogràfiques en aquests centres de recerca. Finalment, el grup creu que la intensa relació que es va establir entre les dues parts, reflecteix la importància que les autoritats britàniques van donar a Catalunya, i que el seu aïllament amb el centre polític del bàndol patriota va permetre que desenvolupés les seves pròpies dinàmiques i cronologies, encara que s’integraven en el desenvolupament general de la guerra.
Resumo:
DNA-based techniques are important tools for species assignment, in particular when identification with morphological criteria is difficult. The aim of this study was to genetically determine the species identity of tree frogs (Hyla spp.) populations from western and northern Switzerland (Swiss Plateau), this area being frequently subjected to introductions of species or sub-species from south of the Alps. We sequenced 261 base pairs of the mitochondrial DNA cytochrome b gene from 24 samples of tree frogs from the Swiss Plateau, Ticino (southern Switzerland) and the Dombes region (Ain, France), and compared them with homologous sequences retrieved from DNA databases. The phylogenetic analyses revealed two distinct clades. The first one is represented by samples of Green tree frog (Hyla arborea) from the Swiss Plateau, France, Germany and Greece, confirming the current knowledge about the species' distribution. The second clade includes samples belonging to the Italian tree frog (Hyla intermedia) from south of the Alps (Ticino and Italy), and unexpectedly from the Grangettes site in western Switzerland. These results suggest the introduction of the Italian tree frog H. intermedia north of the Alps, and raise questions about the management of the Grangettes protected area.
Resumo:
This study presents a first attempt to extend the “Multi-scale integrated analysis of societal and ecosystem metabolism (MuSIASEM)” approach to a spatial dimension using GIS techniques in the Metropolitan area of Barcelona. We use a combination of census and commercial databases along with a detailed land cover map to create a layer of Common Geographic Units that we populate with the local values of human time spent in different activities according to MuSIASEM hierarchical typology. In this way, we mapped the hours of available human time, in regards to the working hours spent in different locations, putting in evidence the gradients in spatial density between the residential location of workers (generating the work supply) and the places where the working hours are actually taking place. We found a strong three-modal pattern of clumps of areas with different combinations of values of time spent on household activities and on paid work. We also measured and mapped spatial segregation between these two activities and put forward the conjecture that this segregation increases with higher energy throughput, as the size of the functional units must be able to cope with the flow of exosomatic energy. Finally, we discuss the effectiveness of the approach by comparing our geographic representation of exosomatic throughput to the one issued from conventional methods.
Resumo:
En termes de temps d'execució i ús de dades, les aplicacions paral·leles/distribuïdes poden tenir execucions variables, fins i tot quan s'empra el mateix conjunt de dades d'entrada. Existeixen certs aspectes de rendiment relacionats amb l'entorn que poden afectar dinàmicament el comportament de l'aplicació, tals com: la capacitat de la memòria, latència de la xarxa, el nombre de nodes, l'heterogeneïtat dels nodes, entre d'altres. És important considerar que l'aplicació pot executar-se en diferents configuracions de maquinari i el desenvolupador d'aplicacions no port garantir que els ajustaments de rendiment per a un sistema en particular continuïn essent vàlids per a d'altres configuracions. L'anàlisi dinàmica de les aplicacions ha demostrat ser el millor enfocament per a l'anàlisi del rendiment per dues raons principals. En primer lloc, ofereix una solució molt còmoda des del punt de vista dels desenvolupadors mentre que aquests dissenyen i evaluen les seves aplicacions paral·leles. En segon lloc, perquè s'adapta millor a l'aplicació durant l'execució. Aquest enfocament no requereix la intervenció de desenvolupadors o fins i tot l'accés al codi font de l'aplicació. S'analitza l'aplicació en temps real d'execució i es considra i analitza la recerca dels possibles colls d'ampolla i optimitzacions. Per a optimitzar l'execució de l'aplicació bioinformàtica mpiBLAST, vam analitzar el seu comportament per a identificar els paràmetres que intervenen en el rendiment d'ella, com ara: l'ús de la memòria, l'ús de la xarxa, patrons d'E/S, el sistema de fitxers emprat, l'arquitectura del processador, la grandària de la base de dades biològica, la grandària de la seqüència de consulta, la distribució de les seqüències dintre d'elles, el nombre de fragments de la base de dades i/o la granularitat dels treballs assignats a cada procés. El nostre objectiu és determinar quins d'aquests paràmetres tenen major impacte en el rendiment de les aplicacions i com ajustar-los dinàmicament per a millorar el rendiment de l'aplicació. Analitzant el rendiment de l'aplicació mpiBLAST hem trobat un conjunt de dades que identifiquen cert nivell de serial·lització dintre l'execució. Reconeixent l'impacte de la caracterització de les seqüències dintre de les diferents bases de dades i una relació entre la capacitat dels workers i la granularitat de la càrrega de treball actual, aquestes podrien ser sintonitzades dinàmicament. Altres millores també inclouen optimitzacions relacionades amb el sistema de fitxers paral·lel i la possibilitat d'execució en múltiples multinucli. La grandària de gra de treball està influenciat per factors com el tipus de base de dades, la grandària de la base de dades, i la relació entre grandària de la càrrega de treball i la capacitat dels treballadors.
Resumo:
Type 2 diabetes mellitus (T2DM) is a major disease affecting nearly 280 million people worldwide. Whilst the pathophysiological mechanisms leading to disease are poorly understood, dysfunction of the insulin-producing pancreatic beta-cells is key event for disease development. Monitoring the gene expression profiles of pancreatic beta-cells under several genetic or chemical perturbations has shed light on genes and pathways involved in T2DM. The EuroDia database has been established to build a unique collection of gene expression measurements performed on beta-cells of three organisms, namely human, mouse and rat. The Gene Expression Data Analysis Interface (GEDAI) has been developed to support this database. The quality of each dataset is assessed by a series of quality control procedures to detect putative hybridization outliers. The system integrates a web interface to several standard analysis functions from R/Bioconductor to identify differentially expressed genes and pathways. It also allows the combination of multiple experiments performed on different array platforms of the same technology. The design of this system enables each user to rapidly design a custom analysis pipeline and thus produce their own list of genes and pathways. Raw and normalized data can be downloaded for each experiment. The flexible engine of this database (GEDAI) is currently used to handle gene expression data from several laboratory-run projects dealing with different organisms and platforms. Database URL: http://eurodia.vital-it.ch.
Resumo:
MOTIVATION: Microarray results accumulated in public repositories are widely reused in meta-analytical studies and secondary databases. The quality of the data obtained with this technology varies from experiment to experiment, and an efficient method for quality assessment is necessary to ensure their reliability. RESULTS: The lack of a good benchmark has hampered evaluation of existing methods for quality control. In this study, we propose a new independent quality metric that is based on evolutionary conservation of expression profiles. We show, using 11 large organ-specific datasets, that IQRray, a new quality metrics developed by us, exhibits the highest correlation with this reference metric, among 14 metrics tested. IQRray outperforms other methods in identification of poor quality arrays in datasets composed of arrays from many independent experiments. In contrast, the performance of methods designed for detecting outliers in a single experiment like Normalized Unscaled Standard Error and Relative Log Expression was low because of the inability of these methods to detect datasets containing only low-quality arrays and because the scores cannot be directly compared between experiments. AVAILABILITY AND IMPLEMENTATION: The R implementation of IQRray is available at: ftp://lausanne.isb-sib.ch/pub/databases/Bgee/general/IQRray.R. CONTACT: Marta.Rosikiewicz@unil.ch SUPPLEMENTARY INFORMATION: Supplementary data are available at Bioinformatics online.
Resumo:
El Consorci de Biblioteques Universitàries de Catalunya (CBUC) va ser creat el 1996 amb l'objectiu de fer i mantenir el catàleg col·lectiu de les universitats de Catalunya (CCUC) però aviat va ampliar les seves activitats amb el préstec interbibliotecari i les compres conjuntes d'informació electrònica. Aquesta darrera activitat va iniciar-se a finals de 1997 quan el CBUC va presentar als vicerectors de recerca de les universitats públiques de Catalunya el projecte de comprar bases de dades de manera consorciada. Aquests van estar-hi d'acord i van manifestar el seu interès de que en les compres conjuntes també s'incloguessin revistes electròniques. El CBUC va decidir englobar aquestes activitats sota el nom Biblioteca Digital de Catalunya (BDC) la qual naixia amb la "finalitat de proporcionar un conjunt nuclear comú d'informació electrònica per a la totalitat dels usuaris de les biblioteques del CBUC". A finals de 1998 el projecte de la BDC es va presentar a la Generalitat de Catalunya i es va obtenir un finançament per al projecte que cobria el període 1999-2001. Des de llavors la BDC ha passat per almenys tres fases: Una de formació, 1999-2001, que es va iniciar amb un ajut del llavors Departament d'Universitats, Recerca i Societat de la Informació (DURSI) de la Generalitat de Catalunya, ajut que es va traduir en una inversió de 180.000€/any i que va permetre l'inici de subscripcions conjuntes, principalment bases de dades. Una de creixement, 2002-2004, realitzada a partir d'un increment de l'ajut del DURSI, ajut que s'usa com a "capital llavor" per subscriure de forma especial revistes. En aquest moment la BDC s'amplia a universitats no membres del CBUC. Una d'estabilització, 2005-2009, en la que s'han fet algunes compres per a una part de les universitats (i no per a totes com fins llavors) i s'han iniciat alguns intents d'estendre la BDC a altres institucions de recerca. L'article caracteritza les diferents fases i mostra les causes de la seva evolució. Finalment, s'exposen els principals assoliments de la BDC així com els reptes de futur més immediats.
Resumo:
BACKGROUND: Superinfection with drug resistant HIV strains could potentially contribute to compromised therapy in patients initially infected with drug-sensitive virus and receiving antiretroviral therapy. To investigate the importance of this potential route to drug resistance, we developed a bioinformatics pipeline to detect superinfection from routinely collected genotyping data, and assessed whether superinfection contributed to increased drug resistance in a large European cohort of viremic, drug treated patients. METHODS: We used sequence data from routine genotypic tests spanning the protease and partial reverse transcriptase regions in the Virolab and EuResist databases that collated data from five European countries. Superinfection was indicated when sequences of a patient failed to cluster together in phylogenetic trees constructed with selected sets of control sequences. A subset of the indicated cases was validated by re-sequencing pol and env regions from the original samples. RESULTS: 4425 patients had at least two sequences in the database, with a total of 13816 distinct sequence entries (of which 86% belonged to subtype B). We identified 107 patients with phylogenetic evidence for superinfection. In 14 of these cases, we analyzed newly amplified sequences from the original samples for validation purposes: only 2 cases were verified as superinfections in the repeated analyses, the other 12 cases turned out to involve sample or sequence misidentification. Resistance to drugs used at the time of strain replacement did not change in these two patients. A third case could not be validated by re-sequencing, but was supported as superinfection by an intermediate sequence with high degenerate base pair count within the time frame of strain switching. Drug resistance increased in this single patient. CONCLUSIONS: Routine genotyping data are informative for the detection of HIV superinfection; however, most cases of non-monophyletic clustering in patient phylogenies arise from sample or sequence mix-up rather than from superinfection, which emphasizes the importance of validation. Non-transient superinfection was rare in our mainly treatment experienced cohort, and we found a single case of possible transmitted drug resistance by this route. We therefore conclude that in our large cohort, superinfection with drug resistant HIV did not compromise the efficiency of antiretroviral treatment.
Obtenció de nous anàlegs amb activitat brassinoesteroide mitjançant modelització molecular i síntesi
Resumo:
Els brassinoesteroides són productes naturals que actuen com a potents reguladors del creixement vegetal. Presenten aplicacions prometedores en l’agricultura degut a que, aplicats exògenament, augmenten la qualitat i la quantitat de les collites. Ara bé, el seu ús s’ha vist restringit degut a la seva costosa obtenció. Aquest fet ha motivat la recerca de nous compostos actius més assequibles. En aquest projecte es planteja el disseny i obtenció de nous anàlegs seguint diferents estratègies que impliquen tant l’ús de mètodes de modelització molecular com de síntesi orgànica. La primera d’aquestes estratègies consisteix en buscar compostos actius en bases de dades de compostos comercials a través de processos de Virtual Screening desenvolupats amb mètodes computacionals basats en Camps d’Interacció Molecular. Així, es van establir i interpretar models de Relacions Quantitatives Estructura-Activitat (QSAR) emprant descriptors independents de l’alineament (GRIND) i, amb col•laboració amb la Universitat de Perugia, aquest criteri de cerca es va ampliar amb l’aplicació de descriptors FLAP de nova generació. Una altra estratègia es va basar en intentar substituir l’esquelet esteroide dels brassinoesteroides per una estructura equivalent, fixant com a cadena lateral el grup (R)-hexahidromandelil. S’han aplicat dos criteris: mètodes computacionals basats en models QSAR establerts amb descriptors GRIND i també en la metodologia SHOP (scaffold hopping), i, per altra banda, anàlegs proposats racionalment a partir d’un estudi efectuat sobre disruptors endocrins no esteroïdals. Sobre les estructures trobades s’hi va unir la cadena lateral comercial esmentada per via sintètica, en la qual s’ha hagut de fer un èmfasi especial en grups protectors. En total, 49 estructures es proposen per a ser obtingudes sintèticament. També s’ha treballat en l’obtenció un agonista derivat de l’hipotètic antagonista KM-01. Totes les molècules candidates, ja siguin comercials o obtingudes sintèticament, estant sent avaluades en el test d’inclinació de la làmina d’arròs (RLIT).