Exploiting instruction- and data-level parallelism
Roger Espasa, Mateo Valero
Abstract
Roger Espasa, Mateo Valero
Abstract
Simultaneous multithreaded vector architectures combine the best of data-level and instruction-level parallelism and perform better than either approach could separately. Our design achieves performance equivalent to executing 15 to 26 scalar instructions/cycle for numerical applications.
OpenAlex reports 41 citations for this work. Citation counts describe recorded attention and do not establish research quality.
A contribution statement is not available in the OpenAlex record.
Method details are not available in the OpenAlex metadata.
Findings are not separately available in the OpenAlex metadata.
Limitations are not available in the OpenAlex metadata.
Application details are not available in the OpenAlex metadata.
Simultaneous multithreaded vector architectures combine the best of data-level and instruction-level parallelism and perform better than either approach could separately. Our design achieves performance equivalent to executing 15 to 26 scalar instructions/cycle for numerical applications.
Key concepts: Computer science, Instruction-level parallelism, Parallel computing, Data parallelism, Parallelism (grammar), Task parallelism, Instruction set, Computer architecture