A Comprehensive Methodology to Optimize FPGA Designs via the Roofline Model

Marco Siracusa,Donatella Sciuto,Samuel Williams,Lorenzo Di Tucci,Marco Rabozzi,Marco Domenico Santambrogio,Emanuele Del Sozzo

doi:10.1109/tc.2021.3111761

Marco Siracusa, Donatella Sciuto + Show 5 more

Open Access

https://doi.org/10.1109/tc.2021.3111761

Copy DOI

Abstract

With reconfigurable fabrics delivering increasing performance over the years, Field-Programmable Gate Arrays (FPGAs) are becoming an appealing solution for next-generation High-Performance Computing (HPC) systems. However, in order to gain traction among traditional von Neumann architectures, the optimization process of Field-Programmable Gate Array (FPGA) designs should be further abstracted to a higher level. In fact, while High-Level Synthesis (HLS) already provides a handy way to write FPGA code with common high-level languages, substantial effort and expertise are still required to optimize the resulting FPGA design for the underlying hardware. To overcome this problem, we propose a semi-automated performance optimization methodology based on a Hierarchical Roofline model for FPGAs. System-wide and applications-specific optimizations such as off-chip memory transfer and data locality optimizations are guided by the FPGA Roofline model whereas FPGA-specific optimizations are automatically searched by a Design Space Exploration (DSE) engine. We demonstrate the way this methodology allows to easily analyze and optimize to peak system performance a wide set of applications ranging from particle methods, wavefront algorithms, and sparse arithmetic computations. In addition, we prove that the integrated Design Space Exploration (DSE) engine achieves a 14.36x maximum speedup if compared to previous automated solutions in the literature.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: IEEE Transactions on Computers	Publication Date: Aug 1, 2022
Citations: 18	License type: other-oa

R Discovery Prime

R Discovery Prime

A Comprehensive Methodology to Optimize FPGA Designs via the Roofline Model

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Computers

Lead the way for us

Similar Papers

Programming Reconfigurable Heterogeneous Computing Clusters Using MPI With Transpilation
Burkhard Ringlein ... Francois Abel
-
Burkhard Ringlein, et. al.Burkhard Ringlein ... Francois Abel
01 Nov 2020
01 Nov 2020

FPGA design for constrained energy minimization
Jianwei Wang ... Chein-I Chang
-
Jianwei Wang, et. al.Jianwei Wang ... Chein-I Chang
27 Feb 2004
27 Feb 2004

Chapter 3 - Optimizing the Development Cycle
R.C Cofer ... Benjamin F Harding
Rapid System Prototyping with FPGAs | VOL. -
R.C Cofer, et. al.R.C Cofer ... Benjamin F Harding
01 Jan 2006
Rapid System Prototyping with FPGAs | VOL. -

Design automation tools for FPGA design (panel)
Kella Knack ... Gordan Hyland
-
Kella Knack, et. al.Kella Knack ... Gordan Hyland
01 Jan 1993
01 Jan 1993

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

A Comprehensive Methodology to Optimize FPGA Designs via the Roofline Model

Abstract

Talk to us

Similar Papers

More From: IEEE Transactions on Computers