Divergence Reduction in Monte Carlo Neutron Transport with On-GPU Asynchronous Scheduling

Braxton Cuneo,Mike Bailey

doi:10.1145/3626957

Abstract

While Monte Carlo Neutron Transport (MCNT) is near-embarrasingly parallel, the effectively unpredictable lifetime of neutrons can lead to divergence when MCNT is evaluated on GPUs. Divergence is the phenomenon of adjacent threads in a warp executing different control flow paths; on GPUS, it reduces performance because each work group may only execute one path at a time. The process of Thread Data Remapping (TDR) resolves these discrepancies by moving data across hardware such that data in the same warp will be processed through similar paths. A common issue among prior implementations of TDR is the synchronous nature of its remapping and processing cycles, which exhaustively sort data produced by prior processing passes and exhaustively evaluate the sorted data. In another work, we defined a method of remapping data through an asynchronous scheduler which allows for work to be stored in shared memory and deferred arbitrarily until that work is a viable option for low-divergence evaluation. This article surveys a wider set of cases, with the goal of characterizing performance trends across a more comprehensive set of parameters. These parameters include cross sections of scattering/capturing/fission, use of implicit capture, source neutron counts, simulation time spans, and tuned memory allocations. Across these cases, we have recorded minimum and average execution times, as well as a heuristically tuned near-optimal memory allocation size for both synchronous and asynchronous scheduling. Across the collected data, it is shown that the asynchronous method is faster and more memory efficient in the majority of cases, and that it requires less tuning to achieve competitive performance.

Full Text

Published version (

Free)

Open DOI Link

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: ACM transactions on modeling and computer simulation : a publication of the Association for Computing Machinery	Publication Date: Jan 14, 2024
Citations: 1	License type: other-oa

R Discovery Prime

R Discovery Prime

Divergence Reduction in Monte Carlo Neutron Transport with On-GPU Asynchronous Scheduling

Abstract

Talk to us

Similar Papers

More From: ACM transactions on modeling and computer simulation : a publication of the Association for Computing Machinery

Lead the way for us

Similar Papers

Determining average program execution times and their variance
V Sarkar
-
V SarkarV Sarkar
21 Jun 1989
21 Jun 1989

Determining average program execution times and their variance
V Sarkar
ACM SIGPLAN Notices | VOL. 24
V SarkarV Sarkar
21 Jun 1989
ACM SIGPLAN Notices | VOL. 24

ASSESSMENT OF TECHNICAL ABILITIES IN ENDOSCOPIC SURGERY SIMULATED BY MEDICINE STUDENTS
G P Ibero Casadiego ... M Merayo Alvarez
The European Journal of Surgery | VOL. 108
G P Ibero Casadiego, et. al.G P Ibero Casadiego ... M Merayo Alvarez
16 May 2021
The European Journal of Surgery | VOL. 108

Analysis of checkpointing schemes for multiprocessor systems
A Ziv ... J Bruck
-
A Ziv, et. al.A Ziv ... J Bruck
25 Oct 1994
25 Oct 1994

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Divergence Reduction in Monte Carlo Neutron Transport with On-GPU Asynchronous Scheduling

Abstract

Talk to us

Similar Papers

More From: ACM transactions on modeling and computer simulation : a publication of the Association for Computing Machinery