May, 16

Efficient Resource Scheduling for Big Data Processing on Accelerator-based Heterogeneous Systems

The involvement of accelerators is becoming widespread in the field of heterogeneous processing, performing computation tasks through a wide range of applications. In this paper, we examine the heterogeneity in modern computing systems, particularly, how to achieve a good level of resource utilization and fairness, when multiple tasks with different load and computation ratios are […]
May, 16

Using Butterfly-Patterned Partial Sums to Optimize GPU Memory Accesses for Drawing from Discrete Distributions

We describe a technique for drawing values from discrete distributions, such as sampling from the random variables of a mixture model, that avoids computing a complete table of partial sums of the relative probabilities. A table of alternate ("butterfly-patterned") form is faster to compute, making better use of coalesced memory accesses. From this table, complete […]
May, 16

Performance Analysis and Efficient Execution on Systems with multi-core CPUs, GPUs and MICs

We carry out a comparative performance study of multi-core CPUs, GPUs and Intel Xeon Phi (Many Integrated Core – MIC) with a microscopy image analysis application. We experimentally evaluate the performance of computing devices on core operations of the application. We correlate the observed performance with the characteristics of computing devices and data access patterns, […]
May, 15

Speeding up Automatic Hyperparameter Optimization of Deep Neural Networks by Extrapolation of Learning Curves

Deep neural networks (DNNs) show very strong performance on many machine learning problems, but they are very sensitive to the setting of their hyperparameters. Automated hyperparameter optimization methods have recently been shown to yield settings competitive with those found by human experts, but their widespread adoption is hampered by the fact that they require more […]
May, 15

MRCUDA: MapReduce Acceleration Framework Based on GPU

GPU programming model for general purpose computing is complex and difficult to be maintained. A MapReduce acceleration framework named MRCUDA is designed and implemented in this paper. There are four loosely coupled stages in MRCUDA, including Pre-Processing, Map, Group and Reduce, which can support flexible configurations for different applications. In order to take full advantage […]
May, 15

The 3D Flow Field Around an Embedded Planet

Understanding the 3D flow topology around a planet embedded in its natal disk is crucial to the study of planet formation. 3D modifications to the well-studied 2D flow topology have the potential to resolve longstanding problems in both planet migration and accretion. We present a detailed analysis of the 3D isothermal flow field around a […]
May, 15

Adaptive discrete cosine transform-based image compression method on a heterogeneous system platform using Open Computing Language

Discrete cosine transform (DCT) is one of the major operations in image compression standards and it requires intensive and complex computations. Recent computer systems and handheld devices are equipped with high computing capability devices such as a general-purpose graphics processing unit (GPGPU) in addition to the traditional multicores CPU. We develop an optimized parallel implementation […]
May, 15

Power, Energy and Speed of Embedded and Server Multi-Cores applied to Distributed Simulation of Spiking Neural Networks: ARM in NVIDIA Tegra vs Intel Xeon quad-cores

This short note regards a comparison of instantaneous power, total energy consumption, execution time and energetic cost per synaptic event of a spiking neural network simulator (DPSNN-STDP) distributed on MPI processes when executed either on an embedded platform (based on a dual socket quad-core ARM platform) or a server platform (INTEL-based quad-core dual socket platform). […]
May, 15

5th International Conference on Computer and Communication Devices (ICCCD), 2015

Publication: All accepted papers will be published in one of the indexed Journals after being selected. * International Journal of Future Computer and Communication (IJFCC, ISSN: 2010-3751) Abstracting/ Indexing: Google Scholar, Engineering & Technology Digital Library, and Crossref, DOAJ, Electronic Journals Library, EI (INSPEC, IET). * International Journal of Computer and Communication Engineering (IJCCE, ISSN: […]
May, 15

5th International Conference on Robotics and Automation Sciences (ICRAS, former ICSIA), 2015

Publication: All the papers of ICRAS 2015 will be indexed by Ei Compendex and ISI.   Topics: AREA 1: Intelligent Control Systems and Optimization • Genetic Algorithms • Fuzzy Control • Decision Support Systems • Machine Learning in Control Applications • Knowledge-based Systems Applications • Hybrid Learning Systems • Distributed Control Systems • Evolutionary Computation […]
May, 15

6th International Conference on Networking and Information Technology (ICNIT), 2015

Topics: Antennas & Propagation Bioinformatics and Scientific Computing Broadband & Intelligent networks Business Information Systems Communication Systems and Networks Complex Systems: Modeling and Simulation Computational Intelligence Applications Computer Vision & Pattern Recognition Data Base Management Data Mining and Data Fusion Data Warehousing, Ontologies and Databases Distributed Sensor Networks E-Commerce & E-government E-Health & Biomedical Applications […]
May, 14

OpenMPCon 2015 – Developer Conference

OpenMPCon is the annual, face-to-face developer gathering organized by the OpenMP community, for the community. Enjoy keynotes, inspirational talks, and a friendly atmosphere that helps attendees meet interesting people, learn more about OpenMP from each other, and have a stimulating experience. Multiple diverse technical tracks are being formulated that will appeal to anyone: from the […]
Page 3 of 80412345...102030...Last »

* * *

* * *

Like us on Facebook

HGPU group

244 people like HGPU on Facebook

Follow us on Twitter

HGPU group

1473 peoples are following HGPU @twitter

* * *

Free GPU computing nodes at hgpu.org

Registered users can now run their OpenCL application at hgpu.org. We provide 1 minute of computer time per each run on two nodes with two AMD and one nVidia graphics processing units, correspondingly. There are no restrictions on the number of starts.

The platforms are

Node 1
  • GPU device 0: nVidia GeForce GTX 560 Ti 2GB, 822MHz
  • GPU device 1: AMD/ATI Radeon HD 6970 2GB, 880MHz
  • CPU: AMD Phenom II X6 @ 2.8GHz 1055T
  • RAM: 12GB
  • OS: OpenSUSE 13.1
  • SDK: nVidia CUDA Toolkit 6.5.14, AMD APP SDK 3.0
Node 2
  • GPU device 0: AMD/ATI Radeon HD 7970 3GB, 1000MHz
  • GPU device 1: AMD/ATI Radeon HD 5870 2GB, 850MHz
  • CPU: Intel Core i7-2600 @ 3.4GHz
  • RAM: 16GB
  • OS: OpenSUSE 12.3
  • SDK: AMD APP SDK 3.0

Completed OpenCL project should be uploaded via User dashboard (see instructions and example there), compilation and execution terminal output logs will be provided to the user.

The information send to hgpu.org will be treated according to our Privacy Policy

HGPU group © 2010-2015 hgpu.org

All rights belong to the respective authors

Contact us: