Parallel GEMM-based convolutions for deep learning on multicore ARM and RISC-V architectures

Héctor Martínez,Sandra Catalán,Adrián Castelló,Enrique S Quintana-Ortí

doi:10.1016/j.sysarc.2024.103186

Abstract

We present high performance, multi-threaded implementations of three GEMM-based convolution algorithms for multicore processors with ARM and RISC-V architectures. The codes are integrated into CONVLIB, a library that has the following unique features: (1) scripts to automatically generate a key component of GEMM, known as the micro-kernel, which is typically written in assembly language; (2) a modified analytical model to automatically tune the algorithms to the underlying cache architecture; (3) the ability to select four hyper-parameters: micro-kernel, cache parameters, parallel loop, and GEMM algorithm dynamically between calls to the library, without recompiling it; and (4) a driver to identify the best hyper-parameters. In addition, we provide a detailed performance evaluation of the convolution algorithms, on five ARM and RISC-V processors, and we publicly release the codes.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Parallel GEMM-based convolutions for deep learning on multicore ARM and RISC-V architectures

Abstract

Talk to us

Similar Papers

More From: Journal of Systems Architecture

Lead the way for us

Similar Papers

On Problems in OpenBLAS Library Usage in Productized Code on RISC-V
Ksenia Alexeyevna Zaytseva ... Andrey Dmitrievich Sokolov
Proceedings of the Institute for System Programming of the RAS | VOL. 35
Ksenia Alexeyevna Zaytseva, et. al.Ksenia Alexeyevna Zaytseva ... Andrey Dmitrievich Sokolov
01 Jan 2023
Proceedings of the Institute for System Programming of the RAS | VOL. 35

Development an efficient AXI-interconnect unit between set of customized peripheral devices and an implemented dual-core RISC-V processor
Demyana Emil ... Gihan Nagib
The Journal of Supercomputing | VOL. 79
Demyana Emil, et. al.Demyana Emil ... Gihan Nagib
05 May 2023
The Journal of Supercomputing | VOL. 79

A Survey and Analysis on SoC Platform Security in ARM, Intel and RISC-V Architecture
Geraldine Shirley Nicholas ... Fareena Saqib
-
Geraldine Shirley Nicholas, et. al.Geraldine Shirley Nicholas ... Fareena Saqib
01 Aug 2020
01 Aug 2020

Evaluating the CCSDS 123 Compressor Running on RISC-V and ARM Architectures
Carolina Imianosky ... Felipe Viel
-
Carolina Imianosky, et. al.Carolina Imianosky ... Felipe Viel
24 Nov 2020
24 Nov 2020

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Parallel GEMM-based convolutions for deep learning on multicore ARM and RISC-V architectures

Abstract

Talk to us

Similar Papers

More From: Journal of Systems Architecture