Javascript must be enabled to continue!
RC-SIMD: Reconfigurable communication SIMD architecture for image processing applications
View through CrossRef
During the last two decades, Single Instruction Multiple Data (SIMD) processors have become important architectures in embedded systems for image processing applications. The main reasons are their area and energy efficiency. Often the processing elements (PEs) of an SIMD processor are only locally connected. This may result in a communication bottleneck (only access to direct neighbors). One way to solve this is to use a fully connected communication network (FC-SIMD) between PEs. However, this solution leads to an excessive communication area cost, low communication network utilization, and scalability problems. E.g., the area overhead of an FC-SIMD is more than 100% when the number of PEs gets bigger than 64. In this paper, we introduce a new type of SIMD architecture, called RC-SIMD, with a reconfigurable communication network. It uses a delay-line in the instruction bus, causing the accesses to the communication network to be distributed over time. This architecture requires only a very cheap communication network while performing almost the same as expensive FC-SIMD architectures. However, the new architecture causes irregular resource conflicts. We therefore introduce a conflict model that existing schedulers are able to cope with. Experimental results show that, on average (compared to locally connected SIMDs), RC-SIMD require 21% fewer cycles than architecture without the delay-line, while the area overhead is at most 10%.
Title: RC-SIMD: Reconfigurable communication SIMD architecture for image processing applications
Description:
During the last two decades, Single Instruction Multiple Data (SIMD) processors have become important architectures in embedded systems for image processing applications.
The main reasons are their area and energy efficiency.
Often the processing elements (PEs) of an SIMD processor are only locally connected.
This may result in a communication bottleneck (only access to direct neighbors).
One way to solve this is to use a fully connected communication network (FC-SIMD) between PEs.
However, this solution leads to an excessive communication area cost, low communication network utilization, and scalability problems.
E.
g.
, the area overhead of an FC-SIMD is more than 100% when the number of PEs gets bigger than 64.
In this paper, we introduce a new type of SIMD architecture, called RC-SIMD, with a reconfigurable communication network.
It uses a delay-line in the instruction bus, causing the accesses to the communication network to be distributed over time.
This architecture requires only a very cheap communication network while performing almost the same as expensive FC-SIMD architectures.
However, the new architecture causes irregular resource conflicts.
We therefore introduce a conflict model that existing schedulers are able to cope with.
Experimental results show that, on average (compared to locally connected SIMDs), RC-SIMD require 21% fewer cycles than architecture without the delay-line, while the area overhead is at most 10%.
Related Results
The architecture of differences
The architecture of differences
Following in the footsteps of the protagonists of the Italian architectural debate is a mark of culture and proactivity. The synthesis deriving from the artistic-humanistic factors...
Latest advancement in image processing techniques
Latest advancement in image processing techniques
Image processing is method of performing some operations on an image, for enhancing the image or for getting some information from that image, or for some other applications is not...
Virtualizable hardware/software design infrastructure for dynamically partially reconfigurable systems
Virtualizable hardware/software design infrastructure for dynamically partially reconfigurable systems
In most existing works, reconfigurable hardware modules are still managed as conventional hardware devices. Further, the software reconfiguration overhead incurred by loading corre...
Transforming TLP into DLP with the dynamic inter-thread vectorization architecture
Transforming TLP into DLP with the dynamic inter-thread vectorization architecture
Transformer le TLP en DLP avec l'architecture de vectorisation dynamique inter-thread
De nombreux microprocesseurs modernes mettent en œuvre le multi-threading simu...
SIMDOM: A framework for SIMD instruction translation and offloading in heterogeneous mobile architectures
SIMDOM: A framework for SIMD instruction translation and offloading in heterogeneous mobile architectures
AbstractFog and mobile edge computing is a paradigm that augments resource‐scarce mobile devices with resource‐rich network servers to enable ubiquitous computing. Smartphone appli...
Recent development in reconfigurable dielectric resonator antenna and microwave filter: design and application
Recent development in reconfigurable dielectric resonator antenna and microwave filter: design and application
SummaryDeveloping wireless communication systems depends on reconfigurable microwave filters (MF) and dielectric resonator antenna (DRA) because the functionality of the various fi...
Loop unrolling optimization for dual SIMD extension
Loop unrolling optimization for dual SIMD extension
Abstract
SIMD extensions are playing an increasingly important role in high-performance computing and artificial intelligence fields. To fully utilize these components, var...
Analysis of Types in Business Communication using the TOPSIS Method
Analysis of Types in Business Communication using the TOPSIS Method
Information interchange between employees and others outside the corporation is referred to as business communication. To accomplish organizational objectives, managers and staff i...

