Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Reconfigurable Binary Neural Network Accelerator with Adaptive Parallelism Scheme

View through CrossRef
Binary neural networks (BNNs) have attracted significant interest for the implementation of deep neural networks (DNNs) on resource-constrained edge devices, and various BNN accelerator architectures have been proposed to achieve higher efficiency. BNN accelerators can be divided into two categories: streaming and layer accelerators. Although streaming accelerators designed for a specific BNN network topology provide high throughput, they are infeasible for various sensor applications in edge AI because of their complexity and inflexibility. In contrast, layer accelerators with reasonable resources can support various network topologies, but they operate with the same parallelism for all the layers of the BNN, which degrades throughput performance at certain layers. To overcome this problem, we propose a BNN accelerator with adaptive parallelism that offers high throughput performance in all layers. The proposed accelerator analyzes target layer parameters and operates with optimal parallelism using reasonable resources. In addition, this architecture is able to fully compute all types of BNN layers thanks to its reconfigurability, and it can achieve a higher area–speed efficiency than existing accelerators. In performance evaluation using state-of-the-art BNN topologies, the designed BNN accelerator achieved an area–speed efficiency 9.69 times higher than previous FPGA implementations and 24% higher than existing VLSI implementations for BNNs.
Title: Reconfigurable Binary Neural Network Accelerator with Adaptive Parallelism Scheme
Description:
Binary neural networks (BNNs) have attracted significant interest for the implementation of deep neural networks (DNNs) on resource-constrained edge devices, and various BNN accelerator architectures have been proposed to achieve higher efficiency.
BNN accelerators can be divided into two categories: streaming and layer accelerators.
Although streaming accelerators designed for a specific BNN network topology provide high throughput, they are infeasible for various sensor applications in edge AI because of their complexity and inflexibility.
In contrast, layer accelerators with reasonable resources can support various network topologies, but they operate with the same parallelism for all the layers of the BNN, which degrades throughput performance at certain layers.
To overcome this problem, we propose a BNN accelerator with adaptive parallelism that offers high throughput performance in all layers.
The proposed accelerator analyzes target layer parameters and operates with optimal parallelism using reasonable resources.
In addition, this architecture is able to fully compute all types of BNN layers thanks to its reconfigurability, and it can achieve a higher area–speed efficiency than existing accelerators.
In performance evaluation using state-of-the-art BNN topologies, the designed BNN accelerator achieved an area–speed efficiency 9.
69 times higher than previous FPGA implementations and 24% higher than existing VLSI implementations for BNNs.

Related Results

ПОЛІТИЧНИЙ ПАРАЛЕЛІЗМ: УКРАЇНСЬКИЙ КОНТЕКСТ
ПОЛІТИЧНИЙ ПАРАЛЕЛІЗМ: УКРАЇНСЬКИЙ КОНТЕКСТ
<p><em>The study uses an empirical method which involves free finding of right material to study the origin and genesis of political parallelism as an integral characte...
Recent Patents on Measurement of Parallelism of Plates
Recent Patents on Measurement of Parallelism of Plates
Background: Parallel plate structures are widely used in micro-electromechanical systems, and the measuring technology of parallelism of parallel plates becomes more and more inevi...
Virtualizable hardware/software design infrastructure for dynamically partially reconfigurable systems
Virtualizable hardware/software design infrastructure for dynamically partially reconfigurable systems
In most existing works, reconfigurable hardware modules are still managed as conventional hardware devices. Further, the software reconfiguration overhead incurred by loading corre...
LINGUOPOETIC CLASSIFICATION OF PARALLELISM IN AZERBAIJAN AND ENGLISH
LINGUOPOETIC CLASSIFICATION OF PARALLELISM IN AZERBAIJAN AND ENGLISH
The article deals with the ways of classification of linguopoetic features of parallelism in the Azerbaijan and English languages. We present research on the classification of diff...
BINARY TOPOLOGY BASED ON SOME NEW SETS
BINARY TOPOLOGY BASED ON SOME NEW SETS
In this chapter, we introduce and some new sets called binary -open sets, binary -sets, binary -sets, binary -closed sets, binary -sets and binary -sets , which are simple forms of...
Reconfigurable antennas for wireless network security
Reconfigurable antennas for wireless network security
Large scale proliferation of wireless technology coupled with the increasingly hostile information security landscape is of serious concern as organizations continue to widely adop...
Electrostatic Accelerators
Electrostatic Accelerators
Abstract The article contains sections titled: Introduction Types of Electrostatic Accelerators ...
Analisis Legalitas Transaksional Binary Option di Indonesia
Analisis Legalitas Transaksional Binary Option di Indonesia
Abstract The development of financial technology has given birth to a new financial transaction, namely Binary Options. Binary Options market their products as an investment that ...

Back to Top