high performance computing on graphics processing units: hgpu.org

hgpu.org » Programming » Algorithms » linalg: Matrix Computations in Apache Spark

linalg: Matrix Computations in Apache Spark

Reza Bosagh Zadeh, Xiangrui Meng, Burak Yavuz, Aaron Staple, Li Pu, Shivaram Venkataraman, Evan Sparks, Alexander Ulanov, Matei Zaharia

Stanford and Databricks, 475 Via Ortega, Stanford, CA 94305

arXiv:1509.02256 [cs.DC], (8 Sep 2015)

@article{zadeh2015linalg,

title={linalg: Matrix Computations in Apache Spark},

author={Zadeh, Reza Bosagh and Meng, Xiangrui and Yavuz, Burak and Staple, Aaron and Pu, Li and Venkataraman, Shivaram and Sparks, Evan and Ulanov, Alexander and Zaharia, Matei},

year={2015},

month={sep},

archivePrefix={"arXiv"},

primaryClass={cs.DC}

}

Download (PDF)

View

Source

Source codes

Package:

scala-blas: Benchmarks of BLAS libraries with Scala interface

2209

views

We describe matrix computations available in the cluster programming framework, Apache Spark. Out of the box, Spark comes with the mllib.linalg library, which provides abstractions and implementations for distributed matrices. Using these abstractions, we highlight the computations that were more challenging to distribute. When translating single-node algorithms to run on a distributed cluster, we observe that often a simple idea is enough: separating matrix operations from vector operations and shipping the matrix operations to be ran on the cluster, while keeping vector operations local to the driver. In the case of the Singular Value Decomposition, by taking this idea to an extreme, we are able to exploit the computational power of a cluster, while running code written decades ago for a single core. We conclude with a comprehensive set of benchmarks for hardware accelerated matrix computations from the JVM, which is interesting in its own right, as many cluster programming frameworks use the JVM.

Tags: Algorithms, Benchmarking, Computer science, CUBLAS, CUDA, Linear Algebra, Machine learning, Matrix multiplication, nVidia, Package, Scala, Tesla M2050

September 15, 2015 by hgpu

Rating: 1.5/5. From 2 votes.

Please wait...

gpu_tracker: Context manager and CLI that tracks the computational-resource-usage of a code block or shell command, particularly the GPU usage

gpu_tracker: Python package for tracking and profiling GPU utilization in both desktop and high-performance computing environments

* * *

high performance computing on graphics processing units: hgpu.org

linalg: Matrix Computations in Apache Spark

Package:

Recent source codes

QArray

Celerity: High-level C++ for Accelerator Clusters

CIFAR-10 Airbench: 94% on CIFAR-10 in 3.29 second

gpu_tracker: Context manager and CLI that tracks the computational-resource-usage of a code block or shell command, particularly the GPU usage

LOOPer: a polyhedral compiler for expressing fast and portable data parallel algorithms

OpenMC Monte Carlo Code

Polygeist: C/C++ frontend for MLIR

Parallel Gaussian process with kernel approximation in CUDA

Optical flow algorithms for SYCL

OpenMP5-Offload-OpenMC-Intel-PVC

Most viewed papers (last 30 days)

linalg: Matrix Computations in Apache Spark

Package:

Share this:

Recent source codes

Most viewed papers (last 30 days)