Application Performance on the Tri-Lab Linux Capacity Cluster - TLCC

Mahesh Rajan,Jeff Ogden,Courtenay T Vaughan,Douglas Doerfler,Marcus Epperson

doi:10.4018/jdst.2010040102

Mahesh Rajan, Jeff Ogden + Show 3 more

Open Access

https://doi.org/10.4018/jdst.2010040102

Copy DOI

Abstract

In a recent acquisition by DOE/NNSA several large capacity computing clusters called TLCC have been installed at the DOE labs: SNL, LANL and LLNL. TLCC architecture with ccNUMA, multi-socket, multi-core nodes, and InfiniBand interconnect, is representative of the trend in HPC architectures. This paper examines application performance on TLCC contrasting them with Red Storm/Cray XT4. TLCC and Red Storm share similar AMD processors and memory DIMMs. Red Storm however has single socket nodes and custom interconnect. Micro-benchmarks and performance analysis tools help understand the causes for the observed performance differences. Control of processor and memory affinity on TLCC with the numactl utility is shown to result in significant performance gains and is essential to attenuate the detrimental impact of OS interference and cache-coherency overhead. While previous studies have investigated impact of affinity control mostly in the context of small SMP systems, the focus of this paper is on highly parallel MPI applications.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

Journal: International Journal of Distributed Systems and Technologies	Publication Date: Apr 1, 2010
Citations: 6	License type: other-oa

R Discovery Prime

R Discovery Prime

Application Performance on the Tri-Lab Linux Capacity Cluster - TLCC

Abstract

Talk to us

Similar Papers

More From: International Journal of Distributed Systems and Technologies

Lead the way for us

Similar Papers

Performance of an MPI-only semiconductor device simulator on a quad socket/quad core InfiniBand platform.
John Shadid ... Paul Lin
-
John Shadid, et. al.John Shadid ... Paul Lin
01 Jan 2009
01 Jan 2009

The Architecture of Heterogeneous Petascale HPC RIVR
Miran Ulbin ... Zoran Ren
-
Miran Ulbin, et. al.Miran Ulbin ... Zoran Ren
20 Mar 2020
20 Mar 2020

Graphite
Mohammad Hasanzadeh Mofrad ... Yousuf Ahmad
Proceedings of the VLDB Endowment | VOL. 13
Mohammad Hasanzadeh Mofrad, et. al.Mohammad Hasanzadeh Mofrad ... Yousuf Ahmad
01 Feb 2020
Proceedings of the VLDB Endowment | VOL. 13

Compiler Optimization for Irregular Memory Access Patterns in PGAS Programs
Thomas B Rolinger ... Christopher D Krieger
-
Thomas B Rolinger, et. al.Thomas B Rolinger ... Christopher D Krieger
01 Jan 2023
01 Jan 2023

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Application Performance on the Tri-Lab Linux Capacity Cluster - TLCC

Abstract

Talk to us

Similar Papers

More From: International Journal of Distributed Systems and Technologies