2016Unpublished venueRequires access

Compiler optimization for superscalar and pipelined processors

Vishnu P. Bharadwaj, Mahesh K. Rao

Open publisher page 3 citations

Abstract

The exploitation of parallelism at both the multiprocessor or multicore level and at the instruction level is the means to achieve high-performance. The compiler for VLIW and superscalar processors must expose sufficient parallelism to effectively utilize the parallel hardware. The amount of instruction level parallelism available to VLIW processors or superscalar processors can be limited. This will limit the performance of these processors to a certain extent. However, with compiler optimization techniques, its performance can be increased to greater extent. This evaluation shows that utilizing the existing resources of the processor with certain programmer constraints and an efficient scheduling of independent and dependent blocks of instructions, we can increase the performance of the processors. As compiler optimization interact with the micro-architecture in complex ways, certain programmer constraints can be added to reduce the complexity and help the compiler to structure the Assembly code in a manner which can be used for out-of-order execution of the code. This paper provides new methods and improvements for the structure of the Assembly code for execution on superscalar processors.

About this research paper

What this paper is about

The exploitation of parallelism at both the multiprocessor or multicore level and at the instruction level is the means to achieve high-performance. The compiler for VLIW and superscalar processors must expose sufficient parallelism to effectively utilize the parallel hardware. The amount of instruction level parallelism available to VLIW processors or superscalar processors can be limited. This will limit the performance of these processors to a certain extent. However, with compiler optimization techniques, its performance can be increased to greater extent. This evaluation shows that utilizing the existing resources of the processor with certain programmer constraints and an efficient scheduling of independent and dependent blocks of instructions, we can increase the performance of the processors. As compiler optimization interact with the micro-architecture in complex ways, certain programmer constraints can be added to reduce the complexity and help the compiler to structure the Assembly code in a manner which can be used for out-of-order execution of the code. This paper provides new methods and improvements for the structure of the Assembly code for execution on superscalar processors.

Why it matters

OpenAlex reports 3 citations for this work. Citation counts describe recorded attention and do not establish research quality.

Key contribution

A contribution statement is not available in the OpenAlex record.

Method / approach

Method details are not available in the OpenAlex metadata.

Main findings

Findings are not separately available in the OpenAlex metadata.

Limitations

Limitations are not available in the OpenAlex metadata.

Applications

Application details are not available in the OpenAlex metadata.

Available abstract

The exploitation of parallelism at both the multiprocessor or multicore level and at the instruction level is the means to achieve high-performance. The compiler for VLIW and superscalar processors must expose sufficient parallelism to effectively utilize the parallel hardware. The amount of instruction level parallelism available to VLIW processors or superscalar processors can be limited. This will limit the performance of these processors to a certain extent. However, with compiler optimization techniques, its performance can be increased to greater extent. This evaluation shows that utilizing the existing resources of the processor with certain programmer constraints and an efficient scheduling of independent and dependent blocks of instructions, we can increase the performance of the processors. As compiler optimization interact with the micro-architecture in complex ways, certain programmer constraints can be added to reduce the complexity and help the compiler to structure the Assembly code in a manner which can be used for out-of-order execution of the code. This paper provides new methods and improvements for the structure of the Assembly code for execution on superscalar processors.

Key concepts: Computer science, Very long instruction word, Parallel computing, Compiler, Instruction-level parallelism, Programmer, Superscalar, Instruction scheduling

Related papers

Back to paper searchBrowse research topicsOriginal source
Compiler optimization for superscalar and pipelined processors — Research Paper | ScholarLens