Computation and Data Partitioning on Scalable Shared Memory Multiprocessors.
Sudarsan Tandri, Tarek S. Abdelrahman
Abstract
Sudarsan Tandri, Tarek S. Abdelrahman
Abstract
In this paper we identify the factors that affect the derivation of computation and data partitions on scalable shared memory multiprocessors (SSMMs). We show that these factors necessitate an SSMM-conscious approach. In addition to remote memory access, which is the sole factor on distributed memory multiprocessors, cache affinity, memory contention and false sharing are important factors that must be considered. Experimental evidence is presented to demonstrate the impact of these factors on performance using three applications on the KSR1 and the Hector multiprocessors. 1 Introduction Scalable shared memory multiprocessors (SSMMs) are becoming increasingly popular and a viable alternative to distributed memory multiprocessors (DMMs). The Stanford DASH [20], FLASH [14], the KSR1 [24], Toronto's Hector [26], NUMAchine [1], and the Cray T3D [23] are some SSMMs currently in use or under development. Processors in a SSMM share a single coherent address space. However, shared memory is p...
OpenAlex reports 7 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
In this paper we identify the factors that affect the derivation of computation and data partitions on scalable shared memory multiprocessors (SSMMs). We show that these factors necessitate an SSMM-conscious approach. In addition to remote memory access, which is the sole factor on distributed memory multiprocessors, cache affinity, memory contention and false sharing are important factors that must be considered. Experimental evidence is presented to demonstrate the impact of these factors on performance using three applications on the KSR1 and the Hector multiprocessors. 1 Introduction Scalable shared memory multiprocessors (SSMMs) are becoming increasingly popular and a viable alternative to distributed memory multiprocessors (DMMs). The Stanford DASH [20], FLASH [14], the KSR1 [24], Toronto's Hector [26], NUMAchine [1], and the Cray T3D [23] are some SSMMs currently in use or under development. Processors in a SSMM share a single coherent address space. However, shared memory is p...
Key concepts: Computer science, Parallel computing, Scalability, Shared memory, Distributed shared memory, Distributed memory, Computation, Cache-only memory architecture