Biblioteca Digital

985 resultados para Field Programmable Gate Array (FPGA)

Analysis, Design, and Hardware-in-the-Loop Validation of Multi-Active Bridge Converters

Relevância:

100.00% 100.00%

Publicador:

Resumo:

This master's thesis investigates different aspects of Dual-Active-Bridge (DAB) Converter and extends aspects further to Multi-Active-Bridges (MAB). The thesis starts with an overview of the applications of the DAB and MAB and their importance. The analytical part of the thesis includes the derivation of the peak and RMS currents, which is required for finding the losses present in the system. The power converters, considered in this thesis are DAB, Triple-Active Bridge (TAB) and Quad-Active Bridge (QAB). All the theoretical calculations are compared with the simulation results from PLECS software for identifying the correctness of the reviewed and developed theory. The Hardware-in-the-Loop (HIL) simulation is conducted for checking the control operation in real-time with the help of the RT box from the Plexim. Additionally, as in real systems digital signal processor (DSP), system-on-chip or field programmable gate array is employed for the control of the power electronic systems, and the execution of the control in the real-time simulation (RTS) conducted is performed by DSP.

A many-core co-processor for embedded parallel computing on FPGA

Relevância:

100.00% 100.00%

Publicador:

Resumo:

Single processor architectures are unable to provide the required performance of high performance embedded systems. Parallel processing based on general-purpose processors can achieve these performances with a considerable increase of required resources. However, in many cases, simplified optimized parallel cores can be used instead of general-purpose processors achieving better performance at lower resource utilization. In this paper, we propose a configurable many-core architecture to serve as a co-processor for high-performance embedded computing on Field-Programmable Gate Arrays. The architecture consists of an array of configurable simple cores with support for floating-point operations interconnected with a configurable interconnection network. For each core it is possible to configure the size of the internal memory, the supported operations and number of interfacing ports. The architecture was tested in a ZYNQ-7020 FPGA in the execution of several parallel algorithms. The results show that the proposed many-core architecture achieves better performance than that achieved with a parallel generalpurpose processor and that up to 32 floating-point cores can be implemented in a ZYNQ-7020 SoC FPGA.

An FPGA architecture for improved arithmetic performance

Relevância:

100.00% 100.00%

Publicador:

Gera��o autom��tica de controladores em FPGA integrando anima��o gr��fica

Relevância:

100.00% 100.00%

Publicador:

Resumo:

Disserta��o apresentada na Faculdade de Ci��ncias e Tecnologia da Universidade Nova de Lisboa para obten��o do grau de Mestre em Engenharia Electrot��cnica e de Computadores

FPGA-based architecture for hyperspectral unmixing

Relevância:

100.00% 100.00%

Publicador:

Resumo:

This paper proposes an FPGA-based architecture for onboard hyperspectral unmixing. This method based on the Vertex Component Analysis (VCA) has several advantages, namely it is unsupervised, fully automatic, and it works without dimensionality reduction (DR) pre-processing step. The architecture has been designed for a low cost Xilinx Zynq board with a Zynq-7020 SoC FPGA based on the Artix-7 FPGA programmable logic and tested using real hyperspectral datasets. Experimental results indicate that the proposed implementation can achieve real-time processing, while maintaining the methods accuracy, which indicate the potential of the proposed platform to implement high-performance, low cost embedded systems.

Efficient FPGA-based regular expression pattern matching

Relevância:

100.00% 100.00%

Publicador:

Resumo:

An approach to the automatic generation of efficient Field Programmable Gate Arrays (FPGAs) circuits for the Regular Expression-based (RegEx) Pattern Matching problems is presented. Using a novel design strategy, as proposed, circuits that are highly area-and-time-efficient can be automatically generated for arbitrary sets of regular expressions. This makes the technique suitable for applications that must handle very large sets of patterns at high speed, such as in the network security and intrusion detection application domains. We have combined several existing techniques to optimise our solution for such domains and proposed the way the whole process of dynamic generation of FPGAs for RegEX pattern matching could be automated efficiently.

The route from VHDL (Very high speed integrated circuits Hardware Description Language) to FPGA using synthesis

Relevância:

100.00% 100.00%

Publicador:

LALP: a language to program custom FPGA-based acceleration engines

Relevância:

100.00% 100.00%

Publicador:

Resumo:

Field-Programmable Gate Arrays (FPGAs) are becoming increasingly important in embedded and high-performance computing systems. They allow performance levels close to the ones obtained with Application-Specific Integrated Circuits, while still keeping design and implementation flexibility. However, to efficiently program FPGAs, one needs the expertise of hardware developers in order to master hardware description languages (HDLs) such as VHDL or Verilog. Attempts to furnish a high-level compilation flow (e.g., from C programs) still have to address open issues before broader efficient results can be obtained. Bearing in mind an FPGA available resources, it has been developed LALP (Language for Aggressive Loop Pipelining), a novel language to program FPGA-based accelerators, and its compilation framework, including mapping capabilities. The main ideas behind LALP are to provide a higher abstraction level than HDLs, to exploit the intrinsic parallelism of hardware resources, and to allow the programmer to control execution stages whenever the compiler techniques are unable to generate efficient implementations. Those features are particularly useful to implement loop pipelining, a well regarded technique used to accelerate computations in several application domains. This paper describes LALP, and shows how it can be used to achieve high-performance computing solutions.

Run-Time Scalable Hardware for Reconfigurable Systems

Relevância:

100.00% 100.00%

Publicador:

Resumo:

La optimizaci��n de par��metros tales como el consumo de potencia, la cantidad de recursos l��gicos empleados o la ocupaci��n de memoria ha sido siempre una de las preocupaciones principales a la hora de dise��ar sistemas embebidos. Esto es debido a que se trata de sistemas dotados de una cantidad de recursos limitados, y que han sido tradicionalmente empleados para un prop��sito espec��fico, que permanece invariable a lo largo de toda la vida ��til del sistema. Sin embargo, el uso de sistemas embebidos se ha extendido a ��reas de aplicaci��n fuera de su ��mbito tradicional, caracterizadas por una mayor demanda computacional. As��, por ejemplo, algunos de estos sistemas deben llevar a cabo un intenso procesado de se��ales multimedia o la transmisi��n de datos mediante sistemas de comunicaciones de alta capacidad. Por otra parte, las condiciones de operaci��n del sistema pueden variar en tiempo real. Esto sucede, por ejemplo, si su funcionamiento depende de datos medidos por el propio sistema o recibidos a trav��s de la red, de las demandas del usuario en cada momento, o de condiciones internas del propio dispositivo, tales como la duraci��n de la bater��a. Como consecuencia de la existencia de requisitos de operaci��n din��micos es necesario ir hacia una gesti��n din��mica de los recursos del sistema. Si bien el software es inherentemente flexible, no ofrece una potencia computacional tan alta como el hardware. Por lo tanto, el hardware reconfigurable aparece como una soluci��n adecuada para tratar con mayor flexibilidad los requisitos variables din��micamente en sistemas con alta demanda computacional. La flexibilidad y adaptabilidad del hardware requieren de dispositivos reconfigurables que permitan la modificaci��n de su funcionalidad bajo demanda. En esta tesis se han seleccionado las FPGAs (Field Programmable Gate Arrays) como los dispositivos m��s apropiados, hoy en d��a, para implementar sistemas basados en hardware reconfigurable De entre todas las posibilidades existentes para explotar la capacidad de reconfiguraci��n de las FPGAs comerciales, se ha seleccionado la reconfiguraci��n din��mica y parcial. Esta t��cnica consiste en substituir una parte de la l��gica del dispositivo, mientras el resto contin��a en funcionamiento. La capacidad de reconfiguraci��n din��mica y parcial de las FPGAs es empleada en esta tesis para tratar con los requisitos de flexibilidad y de capacidad computacional que demandan los dispositivos embebidos. La propuesta principal de esta tesis doctoral es el uso de arquitecturas de procesamiento escalables espacialmente, que son capaces de adaptar su funcionalidad y rendimiento en tiempo real, estableciendo un compromiso entre dichos par��metros y la cantidad de l��gica que ocupan en el dispositivo. A esto nos referimos con arquitecturas con huellas escalables. En particular, se propone el uso de arquitecturas altamente paralelas, modulares, regulares y con una alta localidad en sus comunicaciones, para este prop��sito. El tama��o de dichas arquitecturas puede ser modificado mediante la adici��n o eliminaci��n de algunos de los m��dulos que las componen, tanto en una dimensi��n como en dos. Esta estrategia permite implementar soluciones escalables, sin tener que contar con una versi��n de las mismas para cada uno de los tama��os posibles de la arquitectura. De esta manera se reduce significativamente el tiempo necesario para modificar su tama��o, as�� como la cantidad de memoria necesaria para almacenar todos los archivos de configuraci��n. En lugar de proponer arquitecturas para aplicaciones espec��ficas, se ha optado por patrones de procesamiento gen��ricos, que pueden ser ajustados para solucionar distintos problemas en el estado del arte. A este respecto, se proponen patrones basados en esquemas sist��licos, as�� como de tipo wavefront. Con el objeto de poder ofrecer una soluci��n integral, se han tratado otros aspectos relacionados con el dise��o y el funcionamiento de las arquitecturas, tales como el control del proceso de reconfiguraci��n de la FPGA, la integraci��n de las arquitecturas en el resto del sistema, as�� como las t��cnicas necesarias para su implementaci��n. Por lo que respecta a la implementaci��n, se han tratado distintos aspectos de bajo nivel dependientes del dispositivo. Algunas de las propuestas realizadas a este respecto en la presente tesis doctoral son un router que es capaz de garantizar el correcto rutado de los m��dulos reconfigurables dentro del ��rea destinada para ellos, as�� como una estrategia para la comunicaci��n entre m��dulos que no introduce ning��n retardo ni necesita emplear recursos configurables del dispositivo. El flujo de dise��o propuesto se ha automatizado mediante una herramienta denominada DREAMS. La herramienta se encarga de la modificaci��n de las netlists correspondientes a cada uno de los m��dulos reconfigurables del sistema, y que han sido generadas previamente mediante herramientas comerciales. Por lo tanto, el flujo propuesto se entiende como una etapa de post-procesamiento, que adapta esas netlists a los requisitos de la reconfiguraci��n din��mica y parcial. Dicha modificaci��n la lleva a cabo la herramienta de una forma completamente autom��tica, por lo que la productividad del proceso de dise��o aumenta de forma evidente. Para facilitar dicho proceso, se ha dotado a la herramienta de una interfaz gr��fica. El flujo de dise��o propuesto, y la herramienta que lo soporta, tienen caracter��sticas espec��ficas para abordar el dise��o de las arquitecturas din��micamente escalables propuestas en esta tesis. Entre ellas est�� el soporte para el realojamiento de m��dulos reconfigurables en posiciones del dispositivo distintas a donde el m��dulo es originalmente implementado, as�� como la generaci��n de estructuras de comunicaci��n compatibles con la simetr��a de la arquitectura. El router has sido empleado tambi��n en esta tesis para obtener un rutado sim��trico entre nets equivalentes. Dicha posibilidad ha sido explotada para aumentar la protecci��n de circuitos con altos requisitos de seguridad, frente a ataques de canal lateral, mediante la implantaci��n de l��gica complementaria con rutado id��ntico. Para controlar el proceso de reconfiguraci��n de la FPGA, se propone en esta tesis un motor de reconfiguraci��n especialmente adaptado a los requisitos de las arquitecturas din��micamente escalables. Adem��s de controlar el puerto de reconfiguraci��n, el motor de reconfiguraci��n ha sido dotado de la capacidad de realojar m��dulos reconfigurables en posiciones arbitrarias del dispositivo, en tiempo real. De esta forma, basta con generar un ��nico bitstream por cada m��dulo reconfigurable del sistema, independientemente de la posici��n donde va a ser finalmente reconfigurado. La estrategia seguida para implementar el proceso de realojamiento de m��dulos es diferente de las propuestas existentes en el estado del arte, pues consiste en la composici��n de los archivos de configuraci��n en tiempo real. De esta forma se consigue aumentar la velocidad del proceso, mientras que se reduce la longitud de los archivos de configuraci��n parciales a almacenar en el sistema. El motor de reconfiguraci��n soporta m��dulos reconfigurables con una altura menor que la altura de una regi��n de reloj del dispositivo. Internamente, el motor se encarga de la combinaci��n de los frames que describen el nuevo m��dulo, con la configuraci��n existente en el dispositivo previamente. El escalado de las arquitecturas de procesamiento propuestas en esta tesis tambi��n se puede beneficiar de este mecanismo. Se ha incorporado tambi��n un acceso directo a una memoria externa donde se pueden almacenar bitstreams parciales. Para acelerar el proceso de reconfiguraci��n se ha hecho funcionar el ICAP por encima de la m��xima frecuencia de reloj aconsejada por el fabricante. As��, en el caso de Virtex-5, aunque la m��xima frecuencia del reloj deber��an ser 100 MHz, se ha conseguido hacer funcionar el puerto de reconfiguraci��n a frecuencias de operaci��n de hasta 250 MHz, incluyendo el proceso de realojamiento en tiempo real. Se ha previsto la posibilidad de portar el motor de reconfiguraci��n a futuras familias de FPGAs. Por otro lado, el motor de reconfiguraci��n se puede emplear para inyectar fallos en el propio dispositivo hardware, y as�� ser capaces de evaluar la tolerancia ante los mismos que ofrecen las arquitecturas reconfigurables. Los fallos son emulados mediante la generaci��n de archivos de configuraci��n a los que intencionadamente se les ha introducido un error, de forma que se modifica su funcionalidad. Con el objetivo de comprobar la validez y los beneficios de las arquitecturas propuestas en esta tesis, se han seguido dos l��neas principales de aplicaci��n. En primer lugar, se propone su uso como parte de una plataforma adaptativa basada en hardware evolutivo, con capacidad de escalabilidad, adaptabilidad y recuperaci��n ante fallos. En segundo lugar, se ha desarrollado un deblocking filter escalable, adaptado a la codificaci��n de v��deo escalable, como ejemplo de aplicaci��n de las arquitecturas de tipo wavefront propuestas. El hardware evolutivo consiste en el uso de algoritmos evolutivos para dise��ar hardware de forma aut��noma, explotando la flexibilidad que ofrecen los dispositivos reconfigurables. En este caso, los elementos de procesamiento que componen la arquitectura son seleccionados de una biblioteca de elementos presintetizados, de acuerdo con las decisiones tomadas por el algoritmo evolutivo, en lugar de definir la configuraci��n de las mismas en tiempo de dise��o. De esta manera, la configuraci��n del core puede cambiar cuando lo hacen las condiciones del entorno, en tiempo real, por lo que se consigue un control aut��nomo del proceso de reconfiguraci��n din��mico. As��, el sistema es capaz de optimizar, de forma aut��noma, su propia configuraci��n. El hardware evolutivo tiene una capacidad inherente de auto-reparaci��n. Se ha probado que las arquitecturas evolutivas propuestas en esta tesis son tolerantes ante fallos, tanto transitorios, como permanentes y acumulativos. La plataforma evolutiva se ha empleado para implementar filtros de eliminaci��n de ruido. La escalabilidad tambi��n ha sido aprovechada en esta aplicaci��n. Las arquitecturas evolutivas escalables permiten la adaptaci��n aut��noma de los cores de procesamiento ante fluctuaciones en la cantidad de recursos disponibles en el sistema. Por lo tanto, constituyen un ejemplo de escalabilidad din��mica para conseguir un determinado nivel de calidad, que puede variar en tiempo real. Se han propuesto dos variantes de sistemas escalables evolutivos. El primero consiste en un ��nico core de procesamiento evolutivo, mientras que el segundo est�� formado por un n��mero variable de arrays de procesamiento. La codificaci��n de v��deo escalable, a diferencia de los codecs no escalables, permite la decodificaci��n de secuencias de v��deo con diferentes niveles de calidad, de resoluci��n temporal o de resoluci��n espacial, descartando la informaci��n no deseada. Existen distintos algoritmos que soportan esta caracter��stica. En particular, se va a emplear el est��ndar Scalable Video Coding (SVC), que ha sido propuesto como una extensi��n de H.264/AVC, ya que este ��ltimo es ampliamente utilizado tanto en la industria, como a nivel de investigaci��n. Para poder explotar toda la flexibilidad que ofrece el est��ndar, hay que permitir la adaptaci��n de las caracter��sticas del decodificador en tiempo real. El uso de las arquitecturas din��micamente escalables es propuesto en esta tesis con este objetivo. El deblocking filter es un algoritmo que tiene como objetivo la mejora de la percepci��n visual de la imagen reconstruida, mediante el suavizado de los "artefactos" de bloque generados en el lazo del codificador. Se trata de una de las tareas m��s intensivas en procesamiento de datos de H.264/AVC y de SVC, y adem��s, su carga computacional es altamente dependiente del nivel de escalabilidad seleccionado en el decodificador. Por lo tanto, el deblocking filter ha sido seleccionado como prueba de concepto de la aplicaci��n de las arquitecturas din��micamente escalables para la compresi��n de video. La arquitectura propuesta permite a��adir o eliminar unidades de computaci��n, siguiendo un esquema de tipo wavefront. La arquitectura ha sido propuesta conjuntamente con un esquema de procesamiento en paralelo del deblocking filter a nivel de macrobloque, de tal forma que cuando se var��a del tama��o de la arquitectura, el orden de filtrado de los macrobloques varia de la misma manera. El patr��n propuesto se basa en la divisi��n del procesamiento de cada macrobloque en dos etapas independientes, que se corresponden con el filtrado horizontal y vertical de los bloques dentro del macrobloque. Las principales contribuciones originales de esta tesis son las siguientes: - El uso de arquitecturas altamente regulares, modulares, paralelas y con una intensa localidad en sus comunicaciones, para implementar cores de procesamiento din��micamente reconfigurables. - El uso de arquitecturas bidimensionales, en forma de malla, para construir arquitecturas din��micamente escalables, con una huella escalable. De esta forma, las arquitecturas permiten establecer un compromiso entre el ��rea que ocupan en el dispositivo, y las prestaciones que ofrecen en cada momento. Se proponen plantillas de procesamiento gen��ricas, de tipo sist��lico o wavefront, que pueden ser adaptadas a distintos problemas de procesamiento. - Un flujo de dise��o y una herramienta que lo soporta, para el dise��o de sistemas reconfigurables din��micamente, centradas en el dise��o de las arquitecturas altamente paralelas, modulares y regulares propuestas en esta tesis. - Un esquema de comunicaciones entre m��dulos reconfigurables que no introduce ning��n retardo ni requiere el uso de recursos l��gicos propios. - Un router flexible, capaz de resolver los conflictos de rutado asociados con el dise��o de sistemas reconfigurables din��micamente. - Un algoritmo de optimizaci��n para sistemas formados por m��ltiples cores escalables que optimice, mediante un algoritmo gen��tico, los par��metros de dicho sistema. Se basa en un modelo conocido como el problema de la mochila. - Un motor de reconfiguraci��n adaptado a los requisitos de las arquitecturas altamente regulares y modulares. Combina una alta velocidad de reconfiguraci��n, con la capacidad de realojar m��dulos en tiempo real, incluyendo el soporte para la reconfiguraci��n de regiones que ocupan menos que una regi��n de reloj, as�� como la r��plica de un m��dulo reconfigurable en m��ltiples posiciones del dispositivo. - Un mecanismo de inyecci��n de fallos que, empleando el motor de reconfiguraci��n del sistema, permite evaluar los efectos de fallos permanentes y transitorios en arquitecturas reconfigurables. - La demostraci��n de las posibilidades de las arquitecturas propuestas en esta tesis para la implementaci��n de sistemas de hardware evolutivos, con una alta capacidad de procesamiento de datos. - La implementaci��n de sistemas de hardware evolutivo escalables, que son capaces de tratar con la fluctuaci��n de la cantidad de recursos disponibles en el sistema, de una forma aut��noma. - Una estrategia de procesamiento en paralelo para el deblocking filter compatible con los est��ndares H.264/AVC y SVC que reduce el n��mero de ciclos de macrobloque necesarios para procesar un frame de video. - Una arquitectura din��micamente escalable que permite la implementaci��n de un nuevo deblocking filter, totalmente compatible con los est��ndares H.264/AVC y SVC, que explota el paralelismo a nivel de macrobloque. El presente documento se organiza en siete cap��tulos. En el primero se ofrece una introducci��n al marco tecnol��gico de esta tesis, especialmente centrado en la reconfiguraci��n din��mica y parcial de FPGAs. Tambi��n se motiva la necesidad de las arquitecturas din��micamente escalables propuestas en esta tesis. En el cap��tulo 2 se describen las arquitecturas din��micamente escalables. Dicha descripci��n incluye la mayor parte de las aportaciones a nivel arquitectural realizadas en esta tesis. Por su parte, el flujo de dise��o adaptado a dichas arquitecturas se propone en el cap��tulo 3. El motor de reconfiguraci��n se propone en el 4, mientras que el uso de dichas arquitecturas para implementar sistemas de hardware evolutivo se aborda en el 5. El deblocking filter escalable se describe en el 6, mientras que las conclusiones finales de esta tesis, as�� como la descripci��n del trabajo futuro, son abordadas en el cap��tulo 7. ABSTRACT The optimization of system parameters, such as power dissipation, the amount of hardware resources and the memory footprint, has been always a main concern when dealing with the design of resource-constrained embedded systems. This situation is even more demanding nowadays. Embedded systems cannot anymore be considered only as specific-purpose computers, designed for a particular functionality that remains unchanged during their lifetime. Differently, embedded systems are now required to deal with more demanding and complex functions, such as multimedia data processing and high-throughput connectivity. In addition, system operation may depend on external data, the user requirements or internal variables of the system, such as the battery life-time. All these conditions may vary at run-time, leading to adaptive scenarios. As a consequence of both the growing computational complexity and the existence of dynamic requirements, dynamic resource management techniques for embedded systems are needed. Software is inherently flexible, but it cannot meet the computing power offered by hardware solutions. Therefore, reconfigurable hardware emerges as a suitable technology to deal with the run-time variable requirements of complex embedded systems. Adaptive hardware requires the use of reconfigurable devices, where its functionality can be modified on demand. In this thesis, Field Programmable Gate Arrays (FPGAs) have been selected as the most appropriate commercial technology existing nowadays to implement adaptive hardware systems. There are different ways of exploiting reconfigurability in reconfigurable devices. Among them is dynamic and partial reconfiguration. This is a technique which consists in substituting part of the FPGA logic on demand, while the rest of the device continues working. The strategy followed in this thesis is to exploit the dynamic and partial reconfiguration of commercial FPGAs to deal with the flexibility and complexity demands of state-of-the-art embedded systems. The proposal of this thesis to deal with run-time variable system conditions is the use of spatially scalable processing hardware IP cores, which are able to adapt their functionality or performance at run-time, trading them off with the amount of logic resources they occupy in the device. This is referred to as a scalable footprint in the context of this thesis. The distinguishing characteristic of the proposed cores is that they rely on highly parallel, modular and regular architectures, arranged in one or two dimensions. These architectures can be scaled by means of the addition or removal of the composing blocks. This strategy avoids implementing a full version of the core for each possible size, with the corresponding benefits in terms of scaling and adaptation time, as well as bitstream storage memory requirements. Instead of providing specific-purpose architectures, generic architectural templates, which can be tuned to solve different problems, are proposed in this thesis. Architectures following both systolic and wavefront templates have been selected. Together with the proposed scalable architectural templates, other issues needed to ensure the proper design and operation of the scalable cores, such as the device reconfiguration control, the run-time management of the architecture and the implementation techniques have been also addressed in this thesis. With regard to the implementation of dynamically reconfigurable architectures, device dependent low-level details are addressed. Some of the aspects covered in this thesis are the area constrained routing for reconfigurable modules, or an inter-module communication strategy which does not introduce either extra delay or logic overhead. The system implementation, from the hardware description to the device configuration bitstream, has been fully automated by modifying the netlists corresponding to each of the system modules, which are previously generated using the vendor tools. This modification is therefore envisaged as a post-processing step. Based on these implementation proposals, a design tool called DREAMS (Dynamically Reconfigurable Embedded and Modular Systems) has been created, including a graphic user interface. The tool has specific features to cope with modular and regular architectures, including the support for module relocation and the inter-module communications scheme based on the symmetry of the architecture. The core of the tool is a custom router, which has been also exploited in this thesis to obtain symmetric routed nets, with the aim of enhancing the protection of critical reconfigurable circuits against side channel attacks. This is achieved by duplicating the logic with an exactly equal routing. In order to control the reconfiguration process of the FPGA, a Reconfiguration Engine suited to the specific requirements set by the proposed architectures was also proposed. Therefore, in addition to controlling the reconfiguration port, the Reconfiguration Engine has been enhanced with the online relocation ability, which allows employing a unique configuration bitstream for all the positions where the module may be placed in the device. Differently to the existing relocating solutions, which are based on bitstream parsers, the proposed approach is based on the online composition of bitstreams. This strategy allows increasing the speed of the process, while the length of partial bitstreams is also reduced. The height of the reconfigurable modules can be lower than the height of a clock region. The Reconfiguration Engine manages the merging process of the new and the existing configuration frames within each clock region. The process of scaling up and down the hardware cores also benefits from this technique. A direct link to an external memory where partial bitstreams can be stored has been also implemented. In order to accelerate the reconfiguration process, the ICAP has been overclocked over the speed reported by the manufacturer. In the case of Virtex-5, even though the maximum frequency of the ICAP is reported to be 100 MHz, valid operations at 250 MHz have been achieved, including the online relocation process. Portability of the reconfiguration solution to today's and probably, future FPGAs, has been also considered. The reconfiguration engine can be also used to inject faults in real hardware devices, and this way being able to evaluate the fault tolerance offered by the reconfigurable architectures. Faults are emulated by introducing partial bitstreams intentionally modified to provide erroneous functionality. To prove the validity and the benefits offered by the proposed architectures, two demonstration application lines have been envisaged. First, scalable architectures have been employed to develop an evolvable hardware platform with adaptability, fault tolerance and scalability properties. Second, they have been used to implement a scalable deblocking filter suited to scalable video coding. Evolvable Hardware is the use of evolutionary algorithms to design hardware in an autonomous way, exploiting the flexibility offered by reconfigurable devices. In this case, processing elements composing the architecture are selected from a presynthesized library of processing elements, according to the decisions taken by the algorithm, instead of being decided at design time. This way, the configuration of the array may change as run-time environmental conditions do, achieving autonomous control of the dynamic reconfiguration process. Thus, the self-optimization property is added to the native self-configurability of the dynamically scalable architectures. In addition, evolvable hardware adaptability inherently offers self-healing features. The proposal has proved to be self-tolerant, since it is able to self-recover from both transient and cumulative permanent faults. The proposed evolvable architecture has been used to implement noise removal image filters. Scalability has been also exploited in this application. Scalable evolvable hardware architectures allow the autonomous adaptation of the processing cores to a fluctuating amount of resources available in the system. Thus, it constitutes an example of the dynamic quality scalability tackled in this thesis. Two variants have been proposed. The first one consists in a single dynamically scalable evolvable core, and the second one contains a variable number of processing cores. Scalable video is a flexible approach for video compression, which offers scalability at different levels. Differently to non-scalable codecs, a scalable video bitstream can be decoded with different levels of quality, spatial or temporal resolutions, by discarding the undesired information. The interest in this technology has been fostered by the development of the Scalable Video Coding (SVC) standard, as an extension of H.264/AVC. In order to exploit all the flexibility offered by the standard, it is necessary to adapt the characteristics of the decoder to the requirements of each client during run-time. The use of dynamically scalable architectures is proposed in this thesis with this aim. The deblocking filter algorithm is the responsible of improving the visual perception of a reconstructed image, by smoothing blocking artifacts generated in the encoding loop. This is one of the most computationally intensive tasks of the standard, and furthermore, it is highly dependent on the selected scalability level in the decoder. Therefore, the deblocking filter has been selected as a proof of concept of the implementation of dynamically scalable architectures for video compression. The proposed architecture allows the run-time addition or removal of computational units working in parallel to change its level of parallelism, following a wavefront computational pattern. Scalable architecture is offered together with a scalable parallelization strategy at the macroblock level, such that when the size of the architecture changes, the macroblock filtering order is modified accordingly. The proposed pattern is based on the division of the macroblock processing into two independent stages, corresponding to the horizontal and vertical filtering of the blocks within the macroblock. The main contributions of this thesis are: - The use of highly parallel, modular, regular and local architectures to implement dynamically reconfigurable processing IP cores, for data intensive applications with flexibility requirements. - The use of two-dimensional mesh-type arrays as architectural templates to build dynamically reconfigurable IP cores, with a scalable footprint. The proposal consists in generic architectural templates, which can be tuned to solve different computational problems. ��A design flow and a tool targeting the design of DPR systems, focused on highly parallel, modular and local architectures. - An inter-module communication strategy, which does not introduce delay or area overhead, named Virtual Borders. - A custom and flexible router to solve the routing conflicts as well as the inter-module communication problems, appearing during the design of DPR systems. - An algorithm addressing the optimization of systems composed of multiple scalable cores, which size can be decided individually, to optimize the system parameters. It is based on a model known as the multi-dimensional multi-choice Knapsack problem. - A reconfiguration engine tailored to the requirements of highly regular and modular architectures. It combines a high reconfiguration throughput with run-time module relocation capabilities, including the support for sub-clock reconfigurable regions and the replication in multiple positions. - A fault injection mechanism which takes advantage of the system reconfiguration engine, as well as the modularity of the proposed reconfigurable architectures, to evaluate the effects of transient and permanent faults in these architectures. - The demonstration of the possibilities of the architectures proposed in this thesis to implement evolvable hardware systems, while keeping a high processing throughput. - The implementation of scalable evolvable hardware systems, which are able to adapt to the fluctuation of the amount of resources available in the system, in an autonomous way. - A parallelization strategy for the H.264/AVC and SVC deblocking filter, which reduces the number of macroblock cycles needed to process the whole frame. - A dynamically scalable architecture that permits the implementation of a novel deblocking filter module, fully compliant with the H.264/AVC and SVC standards, which exploits the macroblock level parallelism of the algorithm. This document is organized in seven chapters. In the first one, an introduction to the technology framework of this thesis, specially focused on dynamic and partial reconfiguration, is provided. The need for the dynamically scalable processing architectures proposed in this work is also motivated in this chapter. In chapter 2, dynamically scalable architectures are described. Description includes most of the architectural contributions of this work. The design flow tailored to the scalable architectures, together with the DREAMs tool provided to implement them, are described in chapter 3. The reconfiguration engine is described in chapter 4. The use of the proposed scalable archtieectures to implement evolvable hardware systems is described in chapter 5, while the scalable deblocking filter is described in chapter 6. Final conclusions of this thesis, and the description of future work, are addressed in chapter 7.

Explointing FPGA block memories for protected cryptographic implementations

Relevância:

100.00% 100.00%

Publicador:

Resumo:

Modern Field Programmable Gate Arrays (FPGAs) are power packed with features to facilitate designers. Availability of features like huge block memory (BRAM), Digital Signal Processing (DSP) cores, embedded CPU makes the design strategy of FPGAs quite different from ASICs. FPGA are also widely used in security-critical application where protection against known attacks is of prime importance. We focus ourselves on physical attacks which target physical implementations. To design countermeasures against such attacks, the strategy for FPGA designers should also be different from that in ASIC. The available features should be exploited to design compact and strong countermeasures. In this paper, we propose methods to exploit the BRAMs in FPGAs for designing compact countermeasures. BRAM can be used to optimize intrinsic countermeasures like masking and dual-rail logic, which otherwise have significant overhead (at least 2X). The optimizations are applied on a real AES-128 co-processor and tested for area overhead and resistance on Xilinx Virtex-5 chips. The presented masking countermeasure has an overhead of only 16% when applied on AES. Moreover Dual-rail Precharge Logic (DPL) countermeasure has been optimized to pack the whole sequential part in the BRAM, hence enhancing the security. Proper robustness evaluations are conducted to analyze the optimization for area and security.

A Survey on FPGA-Based Sensor Systems: Towards Intelligent and Reconfigurable Low-Power Sensors for Computer Vision, Control and Signal Processing

Relevância:

100.00% 100.00%

Publicador:

Resumo:

The current trend in the evolution of sensor systems seeks ways to provide more accuracy and resolution, while at the same time decreasing the size and power consumption. The use of Field Programmable Gate Arrays (FPGAs) provides specific reprogrammable hardware technology that can be properly exploited to obtain a reconfigurable sensor system. This adaptation capability enables the implementation of complex applications using the partial reconfigurability at a very low-power consumption. For highly demanding tasks FPGAs have been favored due to the high efficiency provided by their architectural flexibility (parallelism, on-chip memory, etc.), reconfigurability and superb performance in the development of algorithms. FPGAs have improved the performance of sensor systems and have triggered a clear increase in their use in new fields of application. A new generation of smarter, reconfigurable and lower power consumption sensors is being developed in Spain based on FPGAs. In this paper, a review of these developments is presented, describing as well the FPGA technologies employed by the different research groups and providing an overview of future research within this field.

Constructive Synthesis of Memory-Intensive Accelerators for FPGA From Nested Loop Kernels

Relevância:

100.00% 100.00%

Publicador:

Resumo:

Field-programmable gate arrays are ideal hosts to custom accelerators for signal, image, and data processing but de- mand manual register transfer level design if high performance and low cost are desired. High-level synthesis reduces this design burden but requires manual design of complex on-chip and off-chip memory architectures, a major limitation in applications such as video processing. This paper presents an approach to resolve this shortcoming. A constructive process is described that can derive such accelerators, including on- and off-chip memory storage from a C description such that a user-defined throughput constraint is met. By employing a novel statement-oriented approach, dataflow intermediate models are derived and used to support simple ap- proaches for on-/off-chip buffer partitioning, derivation of custom on-chip memory hierarchies and architecture transformation to ensure user-defined throughput constraints are met with minimum cost. When applied to accelerators for full search motion estima- tion, matrix multiplication, Sobel edge detection, and fast Fourier transform, it is shown how real-time performance up to an order of magnitude in advance of existing commercial HLS tools is enabled whilst including all requisite memory infrastructure. Further, op- timizations are presented that reduce the on-chip buffer capacity and physical resource cost by up to 96% and 75%, respectively, whilst maintaining real-time performance.

Implementa��o em FPGA de um Modem QPSK

Relevância:

100.00% 100.00%

Publicador:

Resumo:

Esta disserta��o insere-se num conjunto de trabalhos a decorrer no Instituto de Telecomunica��es de Aveiro que tem como objetivo o desenvolvimento de um sistema de comunica��o para um UAV. Neste sentido, apresenta a implementa��o e valida��o de um modem em banda base aberto e flex��vel implementado em FPGA, baseado em abordagem SDR, com possibilidade de integra��oo no sistema de comunica��o com o UAV. Ao longo desta disserta��o implementou-se, utilizando o MATLAB, um modem de modula��o adapt��vel, ao qual foram integrados algoritmos de sincronismo e de corre��o de fase. Desta forma, foi poss��vel realizar uma an��lise ao modelo comportamental dos v��rios constituintes do modem abstraindose dos tempos de atraso do processamento ou da precis��o da representa��o dos dados, e assim simplificar a sua implementa��o em hardware. Analisado o modelo comportamental do modem desenvolvido em MATLAB realizou-se a sua implementa��o em hardware para a modula��o QPSK. A sua prototipagem foi realizada, com recurso �� ferramenta computacional Vivado Design Suite 2014.2, utilizando o kit de desenvolvimento ZedBoard e o frontend AD-FMCOMMS1-EBZ. O correto funcionamento dos m��dulos implementados em hardware foi posteriormente avaliado atrav��s de uma interface entre o MATLAB e a Zed- Board, sendo que, os resultados obtidos no modelo em MATLAB serviram como termo de compara��o. Atrav��s da utiliza��o desta interface �� poss��vel validar parte do modem implementado em FPGA, mantendo o restante processamento a ser realizado em MATLAB, validando assim os m��dulos em FPGA de uma forma isolada.

Arquitetura h��brida com DSP e FPGA para implementa��o de controladores de filtros ativos de pot��ncia

Relevância:

100.00% 100.00%

Publicador:

Resumo:

The presence of non-linear loads at a point in the distribution system may deform voltage waveform due to the consumption of non-sinusoidal currents. The use of active power filters allows significant reduction of the harmonic content in the supply current. However, the processing of digital control structures for these filters may require high performance hardware, particularly for reference currents calculation. This work describes the development of hardware structures with high processing capability for application in active power filters. In this sense, it considers an architecture that allows parallel processing using programmable logic devices. The developed structure uses a hybrid model using a DSP and an FPGA. The DSP is used for the acquisition of current and voltage signals, calculation of fundamental current related controllers and PWM generation. The FPGA is used for intensive signal processing, such as the harmonic compensators. In this way, from the experimental analysis, significant reductions of the processing time are achieved when compared to traditional approaches using only DSP. The experimental results validate the designed structure and these results are compared with other ones from architectures reported in the literature.

Implementa��o em FPGA de compensadores de desvios para conversor anal��gico digital intercalado

Relevância:

100.00% 100.00%

Publicador:

Resumo:

This work presents the modeling and FPGA implementation of digital TIADC mismatches compensation systems. The development of the whole work follows a top-down methodology. Following this methodology was developed a two channel TIADC behavior modeling and their respective offset, gain and clock skew mismatches on Simulink. In addition was developed digital mismatch compensation system behavior modeling. For clock skew mismatch compensation fractional delay filters were used, more specifically, the efficient Farrow struct. The definition of wich filter design methodology would be used, and wich Farrow structure, required the study of various design methods presented in literature. The digital compensation systems models were converted to VHDL, for FPGA implementation and validation. These system validation was carried out using the test methodology FPGA In Loop . The results obtained with TIADC mismatch compensators show the high performance gain provided by these structures. Beyond this result, these work illustrates the potential of design, implementation and FPGA test methodologies.

«
1
2
...
5
6
7
8
9
10
11
...
65
66
»