Optimal loop parallelization for maximizing iteration-level parallelism
Optimal loop parallelization for maximizing iteration-level parallelism
复制标题
用于最大化迭代级并行性的最佳循环并行化
DOI:
--
复制
发表时间:
2009
期刊:
影响因子:
--
通讯作者:
Jingling Xue
中科院分区:
文献类型:
--
作者:
Duo Liu;Z. Shao;M. Wang;M. Guo;Jingling Xue
This paper solves the open problem of extracting the maximal number of iterations from a loop that can be executed in parallel on chip multiprocessors. Our algorithm solves it optimally by migrating the weights of parallelism-inhibiting dependences on dependence cycles in two phases. First, we model dependence migration with retiming and formulate this classic loop parallelization into a graph optimization problem, i.e., one of finding retiming values for its nodes so that the minimum non-zero edge weight in the graph is maximized. We present our algorithm in three stages with each being built incrementally on the preceding one. Second, the optimal code for a loop is generated from the retimed graph of the loop found in the first phase. We demonstrate the effectiveness of our optimal algorithm by comparing with a number of representative non-optimal algorithms using a set of benchmarks frequently used in prior work.