Evaluation of various node configurations for fine-grain multithreading on stock processors
Jin‐Soo Kim, Soonhoi Ha, Chu Shik Jhon
Abstract
Jin‐Soo Kim, Soonhoi Ha, Chu Shik Jhon
Abstract
It becomes more and more interesting to construct multi-threaded parallel machines using stock processors due to their high performance/price ratio. However no quantitative analysis has been reported on the effectiveness of various node configurations and its impact on the overall performance. In this paper we explore three different node configurations in detail and compare their dynamic characteristics through instruction-level simulation with six benchmark programs. Our experiments show that employing a dedicated processor for communication and synchronization is a reasonable approach because it can almost double the performance. Several factors that limit the overall speedup are also presented.
OpenAlex reports 1 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
It becomes more and more interesting to construct multi-threaded parallel machines using stock processors due to their high performance/price ratio. However no quantitative analysis has been reported on the effectiveness of various node configurations and its impact on the overall performance. In this paper we explore three different node configurations in detail and compare their dynamic characteristics through instruction-level simulation with six benchmark programs. Our experiments show that employing a dedicated processor for communication and synchronization is a reasonable approach because it can almost double the performance. Several factors that limit the overall speedup are also presented.
Key concepts: Speedup, Computer science, Multithreading, Simultaneous multithreading, Parallel computing, Synchronization (alternating current), Benchmark (surveying), Yarn