high performance computing on graphics processing units: hgpu.org

hgpu.org » Applications » Computer science » Parallel Programming using OpenCL on Modern Architectures

Parallel Programming using OpenCL on Modern Architectures

Allan Svejstrup Nielsen, Allan Peter Engsig-Karup, Bernd Dammann

Technical University of Denmark

Technical University of Denmark, IMM Technical Report 2012-05, 2012

@techreport{nielsen2012parallel,

title={Parallel Programming using OpenCL on Modern Architectures},

author={Nielsen, A.S. and Engsig-Karup, A.P. and Dammann, B.},

year={2012},

institution={Technical University of Denmark}

}

Download (PDF)

View

Source

2021

views

This report is intended as a quick introduction to the OpenCL framework and the aim is to facilitate a smooth transfer into the use OpenCL C for developers with previous GPGPU experience. The purpose of OpenCL is to allow for developers to use all compute resources available on a heterogeneous hardware platform. As well as being an introduction to OpenCL, the report also presents an overview of AMD GPU hardware, covering both the VLIW5/4 architectures and the upcoming Graphics-Core-Next architecture which is to form the basis of AMDs future generation GPUs that are to be as capable at compute as they are at graphics. To conclude the presentation of OpenCL as a language for compute, a matrix-matrix multiplication example is devised and optimized for the VLIW4, Tesla and Fermi architectures. The performance is measured as a function of both matrix and work-group size and results are discussed. Where applicable, the equivalent CUDA implementation is tested for comparison.

Tags: ATI, ATI Radeon HD 6990, Computer science, Heterogeneous systems, Matrix multiplication, nVidia, nVidia GeForce GTX 280, nVidia GeForce GTX 590, OpenCL, Overview, Tutorial

October 15, 2012 by hgpu

Rating: 2.0/5. From 1 vote.

Please wait...

gpu_tracker: Context manager and CLI that tracks the computational-resource-usage of a code block or shell command, particularly the GPU usage

gpu_tracker: Python package for tracking and profiling GPU utilization in both desktop and high-performance computing environments

* * *

high performance computing on graphics processing units: hgpu.org

Parallel Programming using OpenCL on Modern Architectures

Recent source codes

QArray

Celerity: High-level C++ for Accelerator Clusters

CIFAR-10 Airbench: 94% on CIFAR-10 in 3.29 second

gpu_tracker: Context manager and CLI that tracks the computational-resource-usage of a code block or shell command, particularly the GPU usage

LOOPer: a polyhedral compiler for expressing fast and portable data parallel algorithms

OpenMC Monte Carlo Code

Polygeist: C/C++ frontend for MLIR

Parallel Gaussian process with kernel approximation in CUDA

Optical flow algorithms for SYCL

OpenMP5-Offload-OpenMC-Intel-PVC

Most viewed papers (last 30 days)

Parallel Programming using OpenCL on Modern Architectures

Share this:

Recent source codes

Most viewed papers (last 30 days)