FEECA: Design Space Exploration for Low-Latency and Energy-Efficient Capsule Network Accelerators

Alberto Marchisio,Muhammad Abdullah Hanif,Vojtech Mrazek,Muhammad Shafique

doi:10.1109/tvlsi.2021.3059518

Abstract

In the past few years, Capsule Networks (CapsNets) have taken the spotlight compared to traditional convolutional neural networks (CNNs) for image classification. Unlike CNNs, CapsNets have the ability to learn the spatial relationship between features of the images. However, their complexity grows because of their heterogeneous capsule structure and the dynamic routing , which is an iterative algorithm to dynamically learn the coupling coefficients of two consecutive capsule layers. This necessitates specialized hardware accelerators for CapsNets. Moreover, a high-performance and energy-efficient design of CapsNet accelerators requires exploration of different design decisions (such as the size and configuration of the processing array and the structure of the processing elements). Toward this, we make the following key contributions: 1) FEECA , a novel methodology to explore the design space of the (micro)architectural parameters of a CapsNet hardware accelerator and 2) CapsAcc , the first specialized RTL-level hardware architecture to perform CapsNets inference with high performance and high energy efficiency. Our CapsAcc achieves significant performance improvement, compared to an optimized GPU implementation, due to its efficient implementation of key activation functions, such as squash and softmax , and an efficient data reuse for the dynamic routing. The FEECA methodology employs the Non-dominated Sorting Genetic Algorithm (NSGA-II) to explore the Pareto-optimal points with respect to area, performance, and energy consumption. This requires analytical modeling of the number of clock cycles required to perform each operation of the CapsNet inference and the memory accesses to enable a fast yet accurate design space exploration. We synthesized the complete accelerator architecture in a 45-nm CMOS technology using Synopsys design tools and evaluated it for the MNIST benchmark (as done by the original CapsNet paper from Google Brain’s team) and for a more complex data set, the German Traffic Sign Recognition Benchmark (GTSRB).

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

FEECA: Design Space Exploration for Low-Latency and Energy-Efficient Capsule Network Accelerators

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Very Large Scale Integration (VLSI) Systems

Lead the way for us

Journal: IEEE Transactions on Very Large Scale Integration (VLSI) Systems	Publication Date: Apr 1, 2021
Citations: 11

Similar Papers

An Evasion Attack against Stacked Capsule Autoencoder
Jiazhu Dai ... Siwei Xiong
Algorithms | VOL. 15
Jiazhu Dai, et. al.Jiazhu Dai ... Siwei Xiong
19 Jan 2022
Algorithms | VOL. 15

Automatic Speed-Limit Sign Detection and Recognition for Advanced Driver Assistance Systems
-
International Journal of Innovative Technology and Exploring Engineering | VOL. 8
--
31 Aug 2019
International Journal of Innovative Technology and Exploring Engineering | VOL. 8

SeVuc: A study on the Security Vulnerabilities of Capsule Networks against adversarial attacks
Alberto Marchisio ... Muhammad Shafique
Microprocessors and Microsystems | VOL. 96
Alberto Marchisio, et. al.Alberto Marchisio ... Muhammad Shafique
30 Nov 2022
Microprocessors and Microsystems | VOL. 96

POLAR: Performance-aware On-device Learning Capable Programmable Processing-in-Memory Architecture for Low-Power ML Applications
Sathwika Bavikadi ... Mark A Indovina
-
Sathwika Bavikadi, et. al.Sathwika Bavikadi ... Mark A Indovina
01 Aug 2022
01 Aug 2022

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

FEECA: Design Space Exploration for Low-Latency and Energy-Efficient Capsule Network Accelerators

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Very Large Scale Integration (VLSI) Systems