An assessment of COMA multiprocessors
Gyungho Lee
Abstract
Gyungho Lee
Abstract
In Cache Only Memory Architecture (COMA) for distributed shared memory multiprocessors, the physical location of a datum is completely decoupled from its address by organizing the memory local to each node as a cache for shared address space. As in traditional cache-coherent multiprocessors, the overhead involved in coherence enforcement strongly affects the performance. The overhead seems more pronounced in a COMA machine than the traditional multiprocessors because of its relatively huge size of the local memory acting as cache and the level of memory hierarchy at which coherence needs to be enforced. Using trace driven simulations of the Perfect Club Benchmark Suite, this paper studies the change in the miss ratio and the network traffic with the two coherence policies, update and invalidate. Our study shows that the two policies provide distinct characteristics which reveal different opportunity of improving COMA multiprocessors for different choice of coherence policy.>
OpenAlex reports 6 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
In Cache Only Memory Architecture (COMA) for distributed shared memory multiprocessors, the physical location of a datum is completely decoupled from its address by organizing the memory local to each node as a cache for shared address space. As in traditional cache-coherent multiprocessors, the overhead involved in coherence enforcement strongly affects the performance. The overhead seems more pronounced in a COMA machine than the traditional multiprocessors because of its relatively huge size of the local memory acting as cache and the level of memory hierarchy at which coherence needs to be enforced. Using trace driven simulations of the Perfect Club Benchmark Suite, this paper studies the change in the miss ratio and the network traffic with the two coherence policies, update and invalidate. Our study shows that the two policies provide distinct characteristics which reveal different opportunity of improving COMA multiprocessors for different choice of coherence policy.>
Key concepts: Computer science, MESIF protocol, Cache-only memory architecture, Cache coherence, MESI protocol, Parallel computing, Bus sniffing, Cache