A GPU Implementation of Dynamic Programming for the Optimal Polygon Triangulation

Yasuaki ITO; Koji NAKANO

doi:10.1587/transinf.E96.D.2596

IEICE TRANSACTIONS on Information

A GPU Implementation of Dynamic Programming for the Optimal Polygon Triangulation

Yasuaki ITO, Koji NAKANO

Full Text Views

0

Cite this

Summary :

This paper presents a GPU (Graphics Processing Units) implementation of dynamic programming for the optimal polygon triangulation. Recently, GPUs can be used for general purpose parallel computation. Users can develop parallel programs running on GPUs using programming architecture called CUDA (Compute Unified Device Architecture) provided by NVIDIA. The optimal polygon triangulation problem for a convex polygon is an optimization problem to find a triangulation with minimum total weight. It is known that this problem for a convex n-gon can be solved using the dynamic programming technique in O(n³) time using a work space of size O(n²). In this paper, we propose an efficient parallel implementation of this O(n³)-time algorithm on the GPU. In our implementation, we have used two new ideas to accelerate the dynamic programming. The first idea (adaptive granularity) is to partition the dynamic programming algorithm into many sequential kernel calls of CUDA, and to select the best parameters for the size and the number of blocks for each kernel call. The second idea (sliding and mirroring arrangements) is to arrange the working data for coalesced access of the global memory in the GPU to minimize the memory access overhead. Our implementation using these two ideas solves the optimal polygon triangulation problem for a convex 8192-gon in 5.57 seconds on the NVIDIA GeForce GTX 680, while a conventional CPU implementation runs in 1939.02 seconds. Thus, our GPU implementation attains a speedup factor of 348.02.

Publication: IEICE TRANSACTIONS on Information Vol.E96-D No.12 pp.2596-2603

Publication Date: 2013/12/01

Publicized

Online ISSN: 1745-1361

DOI: 10.1587/transinf.E96.D.2596

Type of Manuscript: Special Section PAPER (Special Section on Parallel and Distributed Computing and Networking)

Category

Authors

Yasuaki ITO
Hiroshima University
Koji NAKANO
Hiroshima University

Keyword

dynamic programming, parallel algorithms, coalesced memory access, GPGPU, CUDA

Cite this

Copy

Yasuaki ITO, Koji NAKANO, "A GPU Implementation of Dynamic Programming for the Optimal Polygon Triangulation" in IEICE TRANSACTIONS on Information, vol. E96-D, no. 12, pp. 2596-2603, December 2013, doi: 10.1587/transinf.E96.D.2596.
Abstract: This paper presents a GPU (Graphics Processing Units) implementation of dynamic programming for the optimal polygon triangulation. Recently, GPUs can be used for general purpose parallel computation. Users can develop parallel programs running on GPUs using programming architecture called CUDA (Compute Unified Device Architecture) provided by NVIDIA. The optimal polygon triangulation problem for a convex polygon is an optimization problem to find a triangulation with minimum total weight. It is known that this problem for a convex n-gon can be solved using the dynamic programming technique in O(n³) time using a work space of size O(n²). In this paper, we propose an efficient parallel implementation of this O(n³)-time algorithm on the GPU. In our implementation, we have used two new ideas to accelerate the dynamic programming. The first idea (adaptive granularity) is to partition the dynamic programming algorithm into many sequential kernel calls of CUDA, and to select the best parameters for the size and the number of blocks for each kernel call. The second idea (sliding and mirroring arrangements) is to arrange the working data for coalesced access of the global memory in the GPU to minimize the memory access overhead. Our implementation using these two ideas solves the optimal polygon triangulation problem for a convex 8192-gon in 5.57 seconds on the NVIDIA GeForce GTX 680, while a conventional CPU implementation runs in 1939.02 seconds. Thus, our GPU implementation attains a speedup factor of 348.02.
URL: https://global.ieice.org/en_transactions/information/10.1587/transinf.E96.D.2596/_p

Copy

@ARTICLE{e96-d_12_2596,
author={Yasuaki ITO, Koji NAKANO, },
journal={IEICE TRANSACTIONS on Information},
title={A GPU Implementation of Dynamic Programming for the Optimal Polygon Triangulation},
year={2013},
volume={E96-D},
number={12},
pages={2596-2603},
abstract={This paper presents a GPU (Graphics Processing Units) implementation of dynamic programming for the optimal polygon triangulation. Recently, GPUs can be used for general purpose parallel computation. Users can develop parallel programs running on GPUs using programming architecture called CUDA (Compute Unified Device Architecture) provided by NVIDIA. The optimal polygon triangulation problem for a convex polygon is an optimization problem to find a triangulation with minimum total weight. It is known that this problem for a convex n-gon can be solved using the dynamic programming technique in O(n³) time using a work space of size O(n²). In this paper, we propose an efficient parallel implementation of this O(n³)-time algorithm on the GPU. In our implementation, we have used two new ideas to accelerate the dynamic programming. The first idea (adaptive granularity) is to partition the dynamic programming algorithm into many sequential kernel calls of CUDA, and to select the best parameters for the size and the number of blocks for each kernel call. The second idea (sliding and mirroring arrangements) is to arrange the working data for coalesced access of the global memory in the GPU to minimize the memory access overhead. Our implementation using these two ideas solves the optimal polygon triangulation problem for a convex 8192-gon in 5.57 seconds on the NVIDIA GeForce GTX 680, while a conventional CPU implementation runs in 1939.02 seconds. Thus, our GPU implementation attains a speedup factor of 348.02.},
keywords={},
doi={10.1587/transinf.E96.D.2596},
ISSN={1745-1361},
month={December},}

Copy

TY - JOUR
TI - A GPU Implementation of Dynamic Programming for the Optimal Polygon Triangulation
T2 - IEICE TRANSACTIONS on Information
SP - 2596
EP - 2603
AU - Yasuaki ITO
AU - Koji NAKANO
PY - 2013
DO - 10.1587/transinf.E96.D.2596
JO - IEICE TRANSACTIONS on Information
SN - 1745-1361
VL - E96-D
IS - 12
JA - IEICE TRANSACTIONS on Information
Y1 - December 2013
AB - This paper presents a GPU (Graphics Processing Units) implementation of dynamic programming for the optimal polygon triangulation. Recently, GPUs can be used for general purpose parallel computation. Users can develop parallel programs running on GPUs using programming architecture called CUDA (Compute Unified Device Architecture) provided by NVIDIA. The optimal polygon triangulation problem for a convex polygon is an optimization problem to find a triangulation with minimum total weight. It is known that this problem for a convex n-gon can be solved using the dynamic programming technique in O(n³) time using a work space of size O(n²). In this paper, we propose an efficient parallel implementation of this O(n³)-time algorithm on the GPU. In our implementation, we have used two new ideas to accelerate the dynamic programming. The first idea (adaptive granularity) is to partition the dynamic programming algorithm into many sequential kernel calls of CUDA, and to select the best parameters for the size and the number of blocks for each kernel call. The second idea (sliding and mirroring arrangements) is to arrange the working data for coalesced access of the global memory in the GPU to minimize the memory access overhead. Our implementation using these two ideas solves the optimal polygon triangulation problem for a convex 8192-gon in 5.57 seconds on the NVIDIA GeForce GTX 680, while a conventional CPU implementation runs in 1939.02 seconds. Thus, our GPU implementation attains a speedup factor of 348.02.
ER -

IEICE TRANSACTIONS on Information