Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Investigation into designing VLSI of a flexible architecture for a deep neural network accelerator

View through CrossRef
Deep learning, a branch of AI, makes use of specialised neural networks. Computational acceleration of high-performance deep neural network algorithms continues to necessitate efficient architecture, high memory bandwidth, parallel processing, and resources, despite decades of study on such algorithms. Excessive space requirements are a problem for DNN implementations caused by resource-intensive components like activation functions (AF) and multiply-and-accumulate (MAC) units. In addition, edge-AI applications necessitate a densely packed, power-hungry, high-throughput DNN accelerator. If we want to build a DNN accelerator that uses little power and occupies little space, we need to optimise the MAC architecture, the AF, and the network complexity so that data flows efficiently. In addition, providing functional configurability while working with restricted chip surface is a difficulty for DNN hardware designs based on ASICs. This dissertation explores the efficient and low-power VLSI architecture of DNN accelerators, addressing the hardware implementation of DNN and targeting applications with limited resources. In order to assess MAC and non-linear AF operations, we investigate and enhance the CORDIC architecture. The poor throughput is a major downside of CORDIC-based architectures, even though they are area and power efficient. Consequently, we suggest a pipelined design for CORDIC-based MAC and AF that focuses on performance. Due to the increased hardware resource consumption that comes with pipeline stages, this study investigates the mutual exclusivity of CORDIC stages and investigates in depth the accuracy variation related to the number of stages needed to achieve high throughput.
Title: Investigation into designing VLSI of a flexible architecture for a deep neural network accelerator
Description:
Deep learning, a branch of AI, makes use of specialised neural networks.
Computational acceleration of high-performance deep neural network algorithms continues to necessitate efficient architecture, high memory bandwidth, parallel processing, and resources, despite decades of study on such algorithms.
Excessive space requirements are a problem for DNN implementations caused by resource-intensive components like activation functions (AF) and multiply-and-accumulate (MAC) units.
In addition, edge-AI applications necessitate a densely packed, power-hungry, high-throughput DNN accelerator.
If we want to build a DNN accelerator that uses little power and occupies little space, we need to optimise the MAC architecture, the AF, and the network complexity so that data flows efficiently.
In addition, providing functional configurability while working with restricted chip surface is a difficulty for DNN hardware designs based on ASICs.
This dissertation explores the efficient and low-power VLSI architecture of DNN accelerators, addressing the hardware implementation of DNN and targeting applications with limited resources.
In order to assess MAC and non-linear AF operations, we investigate and enhance the CORDIC architecture.
The poor throughput is a major downside of CORDIC-based architectures, even though they are area and power efficient.
Consequently, we suggest a pipelined design for CORDIC-based MAC and AF that focuses on performance.
Due to the increased hardware resource consumption that comes with pipeline stages, this study investigates the mutual exclusivity of CORDIC stages and investigates in depth the accuracy variation related to the number of stages needed to achieve high throughput.

Related Results

The architecture of differences
The architecture of differences
Following in the footsteps of the protagonists of the Italian architectural debate is a mark of culture and proactivity. The synthesis deriving from the artistic-humanistic factors...
POWER-EFFICIENT VLSI DESIGN: STRATEGIES FOR LOW-POWER APPLICATIONS
POWER-EFFICIENT VLSI DESIGN: STRATEGIES FOR LOW-POWER APPLICATIONS
“Power-Efficient VLSI Design: Strategies for Low-Power Applications” is a comprehensive guide that explores the intricacies of designing energy-efficient integrated circuits, addre...
Electrostatic Accelerators
Electrostatic Accelerators
Abstract The article contains sections titled: Introduction Types of Electrostatic Accelerators ...
AI-DRIVEN VLSI DESIGN AND AUTOMATION
AI-DRIVEN VLSI DESIGN AND AUTOMATION
The rapid increase in the complexity of Very Large-Scale Integration (VLSI) systems, along with continuous semiconductor scaling, has made conventional design and automation techni...
NEURAL NETWORKS AND DEEP LEARNING: THEORITICAL INSIGHTS AND FRAMEWORKS
NEURAL NETWORKS AND DEEP LEARNING: THEORITICAL INSIGHTS AND FRAMEWORKS
“NEURAL NETWORKS AND DEEP LEARNING: THEORITICAL INSIGHTS AND FRAMEWORKS” is a comprehensive guide that dives deep into the world of neural networks and their applications in modern...
CHALLENGES AND CONVERGENCE OF QUANTUM COMPUTING AND VLSI
CHALLENGES AND CONVERGENCE OF QUANTUM COMPUTING AND VLSI
The convergence of quantum computing and VLSI technology represents a promising yet challenging direction for next-generation computing systems. This chapter examines the fundament...
Modified neural networks for rapid recovery of tokamak plasma parameters for real time control
Modified neural networks for rapid recovery of tokamak plasma parameters for real time control
Two modified neural network techniques are used for the identification of the equilibrium plasma parameters of the Superconducting Steady State Tokamak I from external magnetic mea...
FinFET Devices and Integration
FinFET Devices and Integration
Through more than a decade of industry wide R&D effort, 3D-FinFET has found its way into manufacturing. In this abstract, we review the key progress in process and integration ...

Back to Top