Compiler code transformations for superscalar-based high-performance systems

Tokuzo Kiyohara, William Y. Chen, John C. Gyllenhaal, Pohua P. Chang, Scott Mahlke, Wen-mei W. Hwu

doi:10.5555/147877.148138

A set of compiler transformations designed to increase instruction-level parallelism is described. The effectiveness of these transformations is evaluated using 40 loop nests extracted from a range of supercomputer applications. This evaluation shows that increasing execution resources in superscalar/VLIW node processors yields little performance improvement unless loop unrolling and register renaming are applied. It also reveals that these two transformations are sufficient for DOALL loops. However, more advanced transformations are required in order for serial and DOACROSS loops to fully benefit from the increased execution resources. The results show that the six additional transformations studied satisfy most of this need. >

Compiler code transformations for superscalar-based high-performance systems

説明

収録刊行物

詳細情報詳細情報について

書き出し

問題の指摘

Compiler code transformations for superscalar-based high-performance systems

説明

収録刊行物

詳細情報 詳細情報について

書き出し

問題の指摘

詳細情報詳細情報について