Performance Comparison of Asynchronous Transfer Configurations for UHD Game Image Compression with GPGPU

Youngsik Kim

Performance Comparison of Asynchronous Transfer Configurations for UHD Game Image Compression with GPGPU

원문정보

Youngsik Kim

보안공학연구지원센터(IJMUE) International Journal of Multimedia and Ubiquitous Engineering Vol.11 No.12 2016.12 pp.161-171 SCOPUS

피인용수 : 0건 (자료제공 : 네이버학술정보)

초록

영어

Ultra high definition (UHD) game scenes have caused the memory bandwidth problem. The lossless DPCM-GR based compression algorithm [12] using NVIDIA CUDA(Compute Unified Device Architecture) like general purpose GPU (GPGPU) computing relieves the bandwidth problem without sacrificing image quality, which supports bit parallel pipelining. This paper increases the memory bandwidth efficiency using the shared memory of CUDA based on the compression algorithm [12]. Also, various asynchronous transfer configurations which can overlap the kernel execution and data transfer between the Host and the CUDA device are implemented with the page-locked host memory. Experimental results show that GPGPU CUDA computing obtains the maximum 87.5 and 30.6 times speedups for GTX650Ti and GT330, respectively, comparing to Host CPU. Also, the maximum reductions of the compression time for GTX650Ti and GT330 are 54.1% and 30.3%, respectively, among various concurrency transfer configurations.

키워드

저자정보

Youngsik Kim Department of Game and Multimedia Engineering, Korea Polytechnic University

참고문헌

자료제공 : 네이버학술정보

함께 이용한 논문

※ 원문제공기관과의 협약기간이 종료되어 열람이 제한될 수 있습니다.

0개의 논문이 장바구니에 담겼습니다.

earticle